Pith. sign in

Paper Citation Record · LEDGER

SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 59 inbound Pith citation observations for arXiv:1904.08779.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1904.08779 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 59 of 59 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:55.485946Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T01:44:25.918448Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a04ce41b-27bd-47cb-a813-ef28d8c5a7e3 · inbound

Sound source detection, localization and classification using consecutive ensemble of CRNN models cites this paper.

Sound source detection, localization and classification using consecutive ensemble of CRNN models SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T15:36:32.154928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:36:32.154928Z digest=sha256:770be977f1d508b129493350ecf56557e92fd8de56a9271bd9b9bdf582e27091

Observation c703f176-8d9f-438c-b1fd-f5906d419a7a · inbound

Deepfake audio as a data augmentation technique for training automatic speech to text transcription models cites this paper.

Deepfake audio as a data augmentation technique for training automatic speech to text transcription models SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:04:03.132657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T07:01:04.631071Z digest=sha256:38d552687963a2b5afe65cd1ef5146934f2d16d99bb2e18f93bb108da4d64d08

Observation 7d529bb3-52c8-4509-8fde-faf419367190 · inbound

DGSNA: Dynamic Generative Scene-based Noise Addition method cites this paper.

DGSNA: Dynamic Generative Scene-based Noise Addition method SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T17:45:46.293633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T17:43:47.086524Z digest=sha256:8054501c0d17841ca29312aa9ede9a1337f9f4c668d8a7667b086efd4c9b9596

Observation 9185ab94-0d7e-4bd8-869e-ba3796cee21b · inbound

The SVASR System for Text-dependent Speaker Verification (TdSV) AAIC Challenge 2024 cites this paper.

The SVASR System for Text-dependent Speaker Verification (TdSV) AAIC Challenge 2024 SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:27:06.836908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:27:06.836908Z digest=sha256:fdf7320e01afa93ca5cd5cc79aa53e265eb3acb964e43d3e9e600bb0096c4bb7

Observation f47c79d8-4c03-43eb-ba5d-94721ba4c816 · inbound

Complexity boosted adaptive training for better low resource ASR performance cites this paper.

Complexity boosted adaptive training for better low resource ASR performance SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:59:13.220242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:59:13.220242Z digest=sha256:d86b38d2beee86ebf6cdcb438618296ff4d1d6804d6cc2d992d339d911f98a70

Observation bf63702e-89b8-4d4b-a882-2d9c1d41c6f3 · inbound

Efficient Adaptation of Multilingual Models for Japanese ASR cites this paper.

Efficient Adaptation of Multilingual Models for Japanese ASR SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:44:46.873668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:44:46.873668Z digest=sha256:0aed6b3ac982609c1d044e722cabdd915a7e6494f5eb342e42c84a4cf62a7be2

Observation bae8b542-5df0-438b-b3fa-eeaa837fabf5 · inbound

Efficient Speech Command Recognition Leveraging Spiking Neural Network and Curriculum Learning-based Knowledge Distillation cites this paper.

Efficient Speech Command Recognition Leveraging Spiking Neural Network and Curriculum Learning-based Knowledge Distillation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T13:44:44.089147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:44:44.089147Z digest=sha256:176bc67b2e96d82edd6f70a684b5b41257daa69a3844548c66543b7e21123a74

Observation 2413dcaa-af59-4b29-83df-05f6b82ac1e6 · inbound

JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts cites this paper.

JoVALE: Detecting Human Actions in Video Using Audiovisual and Language Contexts SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:55:40.573731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:55:40.573731Z digest=sha256:50b00744205738bae4c88906966c95d84d2f9044e20c13c6d35a273a02fb4b9a

Observation 7c8d5a98-8fde-479c-baac-b39705c6c96b · inbound

LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging cites this paper.

LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:55:44.790298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:55:44.790298Z digest=sha256:f0eddd4df8a949b8c0812aa3666dda156844046c8345ce5c98cacf5e6ccf4e42

Observation 4012f9d8-3fcf-4840-a6e4-8a5fdca92675 · inbound

Adaptive Data Augmentation with NaturalSpeech3 for Far-field Speaker Verification cites this paper.

Adaptive Data Augmentation with NaturalSpeech3 for Far-field Speaker Verification SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:25:17.209834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:25:17.209834Z digest=sha256:b17bef5cd9413a07882d3620ccb6d6d2cd44478d9210e04b8a0331ce2803a176

Observation 7701a5c4-ff0c-4c81-bba2-b716a10662c7 · inbound

A Non-autoregressive Model for Joint STT and TTS cites this paper.

A Non-autoregressive Model for Joint STT and TTS SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:28.496296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:28.496296Z digest=sha256:a8f93a851b994478554c433f50d1b54655dd278f88d4992a88a42995fa53faec

Observation fed1feb7-e486-4175-bffa-60c67e2761ce · inbound

Enhancing Neural Spoken Language Recognition: An Exploration with Multilingual Datasets cites this paper.

Enhancing Neural Spoken Language Recognition: An Exploration with Multilingual Datasets SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T18:43:47.065183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:43:47.065183Z digest=sha256:fc1a84e7747e493487101dc52987ac585b4069c116db4611ac233506d4a7c333

Observation 817baa08-039e-4e67-92b3-105838da1f9c · inbound

Generalizable Audio Deepfake Detection via Latent Space Refinement and Augmentation cites this paper.

Generalizable Audio Deepfake Detection via Latent Space Refinement and Augmentation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:19:16.860363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:19:16.860363Z digest=sha256:316485ba80aaa1be1e76483becb3a1ad477dde47e4b3932f3a18afc7a44002e0

Observation 568b8659-2a80-48ed-9869-2a8fa73ca546 · inbound

Unifying Prediction and Explanation in Time-Series Transformers via Shapley-based Pretraining cites this paper.

Unifying Prediction and Explanation in Time-Series Transformers via Shapley-based Pretraining SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:23.982192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:23.982192Z digest=sha256:d16c3ff4bc3f1ffad3df2351b065f31cf39785942c40134e7c9b2089b2295fb1

Observation 90ed51b8-2383-4dc1-a35a-3860a476e953 · inbound

Variational Bayesian Adaptive Learning of Deep Latent Variables for Acoustic Knowledge Transfer cites this paper.

Variational Bayesian Adaptive Learning of Deep Latent Variables for Acoustic Knowledge Transfer SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T14:18:44.567922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:18:44.567922Z digest=sha256:4b0d689b11c83e2c8b87fa3003912c687dac4824998f3c90b9f4c7584849563d

Observation 039c12cd-7bc5-4e1f-8277-9b37faf27336 · inbound

Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models cites this paper.

Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:49:20.114058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:49:20.114058Z digest=sha256:64cb56b5917a9fd8cbf068180146d46e90e3bc87fa0b4c814266d0861790b4eb

Observation 2111c041-ac5b-4b68-afe0-8a2460f2ce45 · inbound

Measuring Diversity in Synthetic Datasets cites this paper.

Measuring Diversity in Synthetic Datasets SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:50.830221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T04:54:50.830221Z digest=sha256:a605f947804fdc733baf1a0ea208f9b181809a22cb51777465b37810d1b67c82

Observation cba0f5e3-0cb4-491f-ba91-4405453af26e · inbound

Quantum Approaches for Dysphonia Assessment in Small Speech Datasets cites this paper.

Quantum Approaches for Dysphonia Assessment in Small Speech Datasets SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T23:06:36.031969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:06:36.031969Z digest=sha256:cb87e26f8860a50cbe1e82ad7959cdb44207aebba2d097857bc8a31f1f185488

Observation 03cb205e-8e0b-4d46-9b65-52875d06e5ff · inbound

Audio-Visual Class-Incremental Learning for Fish Feeding intensity Assessment in Aquaculture cites this paper.

Audio-Visual Class-Incremental Learning for Fish Feeding intensity Assessment in Aquaculture SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:55.485946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:55.485946Z digest=sha256:0d62fb1af4da09dfc88f3dee938ae266c16d1af6bde705d1a4a1e6578417a4d1

Observation 57e2f672-4b24-4820-88b7-fb975733de86 · inbound

Waveform-Logmel Audio Neural Networks for Respiratory Sound Classification cites this paper.

Waveform-Logmel Audio Neural Networks for Respiratory Sound Classification SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T10:52:01.408763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:52:01.408763Z digest=sha256:0e3c60c45e51976625d27bb5c9ae405f21d163998fa79f8e5d3e6ec8e98d8376

Observation 961d4178-ebeb-4354-a09b-4d127b1c0fc0 · inbound

Improving Out-of-Domain Robustness with Targeted Augmentation in Frequency and Pixel Spaces cites this paper.

Improving Out-of-Domain Robustness with Targeted Augmentation in Frequency and Pixel Spaces SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:31.754609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:31.754609Z digest=sha256:7178e907b14aa75cf38e3c447311812a3d9d19192e5b51e6773259d381d9b25d

Observation a16d5591-2f9b-4db3-95cc-40a8f436e322 · inbound

MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion cites this paper.

MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:23:09.286245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:23:09.286245Z digest=sha256:f88ced994476460657c68cdeb100d5dd868e6cf4c67988d697d69c5b415d8594

Observation 591a8169-8538-457d-ada0-0d00587ee056 · inbound

IIITH-BUT system for IWSLT 2025 low-resource Bhojpuri to Hindi speech translation cites this paper.

IIITH-BUT system for IWSLT 2025 low-resource Bhojpuri to Hindi speech translation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:04.625481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:39:04.625481Z digest=sha256:f67be510a1067b814d63c5de846ae0c46187e53a01be4eb8ea0e637f0df25317

Observation cff84f6b-ee52-4d80-9ec4-c86016e53867 · inbound

Technical Report: A Practical Guide to Kaldi ASR Optimization cites this paper.

Technical Report: A Practical Guide to Kaldi ASR Optimization SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:12.042353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:12.042353Z digest=sha256:b0df4f4cac35632c57421304d38c55a1cb5a97da46dcb2b457d004b320e4f48b

Observation 05030d40-2461-40e1-a152-33082675f082 · inbound

FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition cites this paper.

FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:25.896412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:24:25.896412Z digest=sha256:6f827a3eefef0c2c4ee483f842b45c5742d4f4a8fb0e76796fc76fd8fdb1a43a

Observation c02b99e4-d8be-44d5-873e-b6db9d2f2670 · inbound

From Sharpness to Better Generalization for Speech Deepfake Detection cites this paper.

From Sharpness to Better Generalization for Speech Deepfake Detection SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:06.052690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:08:06.052690Z digest=sha256:b5140fe0c8688889e0d5adb32ac136c9cd6e2d4860d21e12e6263b785981a6dd

Observation 85231792-3c9a-4dfc-b166-7954c14aa931 · inbound

SSLAM: Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes cites this paper.

SSLAM: Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T01:05:44.764826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:05:44.764826Z digest=sha256:eb9093b8f32ad8991f8d35934b7e26e54c3dee1b3b7adfd076226baf1ae8cc52

Observation 72e2873c-942a-42b9-bb99-33a848072dd0 · inbound

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition cites this paper.

State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T19:19:59.756520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:19:59.756520Z digest=sha256:abffe144128c2e3cdbdc25c14bc169a91c418a3d2dbc1e01282060dcde9f7054

Observation 9c330987-b0b0-432e-9436-8f0cfdca80e7 · inbound

Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization cites this paper.

Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:04.150837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:04.150837Z digest=sha256:2631db431cd2d38ab55afb8ff4cf60de9bf8f18e78ff24579f3ae798e5a726b1

Observation c12421a7-57f4-4b49-8354-68c297a431cf · inbound

A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods cites this paper.

A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 216

Resolution
unresolved
no resolver link, observed 2026-08-06T21:06:25.771242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:06:25.771242Z digest=sha256:871808e5bc209f0b99bb762b5932d630d89b90d5bdd35954f9bf670d77db6538

Observation 21ed4ddb-510b-4e3d-b7bf-60e72a320205 · inbound

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World cites this paper.

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:12.124426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:46:12.124426Z digest=sha256:51889d8f7e1da2b9506b293b5627e9e540995717cbd8859b79aef11524dc5de5

Observation d70f57b8-a892-4019-809c-e742ae54d57d · inbound

Adversarial Training Improves Generalization Under Distribution Shifts in Bioacoustics cites this paper.

Adversarial Training Improves Generalization Under Distribution Shifts in Bioacoustics SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T16:24:11.175788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:24:11.175788Z digest=sha256:e8328fe53582218aa2e68cf7e5f49c1233af7d5dbb622c94c318ab54168b968b

Observation b73be1b6-38ab-48fe-9807-0a0d3a789b23 · inbound

Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation cites this paper.

Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T14:42:46.340920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:42:46.340920Z digest=sha256:8537023719513aacea4404bd5754bc9907ed390cb98447640f7c8756983176e0

Observation 480da5c3-5868-46bf-b1d6-fb90ed5bab4e · inbound

Scaling and Distilling Transformer Models for sEMG cites this paper.

Scaling and Distilling Transformer Models for sEMG SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:28:55.409987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:28:55.409987Z digest=sha256:d2ca26b399bd9cb3ea5c1aa56a76d462973d58ef85e743fbad63aabaf2fd54b2

Observation 128a4bcf-b829-44a6-8a29-262746fb2229 · inbound

AD-AVSR: Asymmetric Dual-stream Enhancement for Robust Audio-Visual Speech Recognition cites this paper.

AD-AVSR: Asymmetric Dual-stream Enhancement for Robust Audio-Visual Speech Recognition SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T22:04:12.045872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:04:12.045872Z digest=sha256:d1c694096af69d7f22db205f260a93d9da3fc28d855a7926169f74cc48ed47db

Observation 7f388d3b-f2e9-4dfd-9ee3-e38c5d61dfcc · inbound

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model cites this paper.

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T17:56:54.263652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:56:54.263652Z digest=sha256:ca2cfd2fe5eb9d4102aca46f771fb3e18fbb97bddec83b6cbf9b0c1b90e5194f

Observation a445bdf4-d2b3-4a19-b1e8-3619529f9e55 · inbound

Full-Frequency Temporal Patching and Structured Masking for Enhanced Audio Classification cites this paper.

Full-Frequency Temporal Patching and Structured Masking for Enhanced Audio Classification SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T14:30:47.078357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:30:47.078357Z digest=sha256:1837aa2d51777d5e299ebe20a16eb7ae058105d00bd68fa0249e85d8ed884ed9

Observation d4da8133-e2b0-4b16-b874-062b585b7e8c · inbound

Text Reinforcement for Multimodal Time Series Forecasting cites this paper.

Text Reinforcement for Multimodal Time Series Forecasting SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:55.892351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:55.892351Z digest=sha256:29878286b13745d1b9c5bdf450451d08ee8bb93d6a788ee6c67f57309f9c27f3

Observation 7fc97b93-18f1-4d4e-b276-96f56f85380b · inbound

Beyond Words: Interjection Classification for Improved Human-Computer Interaction cites this paper.

Beyond Words: Interjection Classification for Improved Human-Computer Interaction SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:29.658446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:29.658446Z digest=sha256:4489f9419298757acaa9d1659c3591af3b9646923935ea9a86a74c1198c3fade

Observation 50e7e4b4-82cd-4515-bdde-58683159c110 · inbound

Effective Modeling of Critical Contextual Information for TDNN-based Speaker Verification cites this paper.

Effective Modeling of Critical Contextual Information for TDNN-based Speaker Verification SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T18:29:41.136871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:29:41.136871Z digest=sha256:6e7cb510efcb7351435641e3efa76798cc8586d55d8e8fe25eaec9ed1a70ccca

Observation da4f9378-fd45-4eb5-b02d-d77d40ffa2fc · inbound

DHAuDS: A Dynamic and Heterogeneous Audio Benchmark for Test-Time Adaptation cites this paper.

DHAuDS: A Dynamic and Heterogeneous Audio Benchmark for Test-Time Adaptation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T20:49:53.925309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:49:53.925309Z digest=sha256:e6b5be2c0c9627bfcf4cf13c4e5ec8468defaa9f06f07cb8f85230e907c073c0

Observation 1c1c9ce6-fdff-4b42-b3d6-21c2255a12c6 · inbound

EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting cites this paper.

EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T08:42:32.895398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:42:32.895398Z digest=sha256:31524a34597b226d1ad4cde20a5b86255dea08b9ec154884551c1657a58f1784

Observation 16425bd0-557e-44cb-9957-f8fa741e463f · inbound

SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise cites this paper.

SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:50:22.377406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T22:47:12.521942Z digest=sha256:77f968b6fdc453921211f7aaaee4e2d7490bc8b0e41cdcc8834c6762c156fd46

Observation 9f0e0270-8045-49ae-84e7-e6cd30596cf5 · inbound

SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise cites this paper.

SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T23:47:00.282041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:47:00.282041Z digest=sha256:061a17e6c75d566fe28375ce6c839c5f6eaea1d4c8ce16e95ed9f1232ed260d8

Observation 97421021-e006-4c0d-852c-688b90cedfb6 · inbound

IQRA 2026: Interspeech Challenge on Automatic Pronunciation Assessment for Modern Standard Arabic (MSA) cites this paper.

IQRA 2026: Interspeech Challenge on Automatic Pronunciation Assessment for Modern Standard Arabic (MSA) SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:18:30.012412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T00:15:10.797922Z digest=sha256:d8aa60483e5e707cdbdd7c0b66b34abe3005d1d4681aa8fbcdbdeb77692739e0

Observation 343b8353-94b7-4200-bdbf-e8d82f5fbf0d · inbound

"OK Aura, Be Fair With Me": Demographics-Agnostic Training for Bias Mitigation in Wake-up Word Detection cites this paper.

"OK Aura, Be Fair With Me": Demographics-Agnostic Training for Bias Mitigation in Wake-up Word Detection SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:40:53.659843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T18:55:56.713141Z digest=sha256:043d6185c8b8074293dfd1d3e0bbdfe34166b8e891eece8035fa64d9d8e1206b

Observation 8609d7bd-80f5-4678-b469-98c19eb17637 · inbound

A-SLIP: Acoustic Sensing for Continuous In-hand Slip Estimation cites this paper.

A-SLIP: Acoustic Sensing for Continuous In-hand Slip Estimation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:14.489466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T17:31:08.224669Z digest=sha256:8b0554febccb894864b533451c25af4bcc6bf889a36fbbbbfd7f8e77218ef150

Observation c4194cdd-0786-4f1d-879f-7207401a970e · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:21:05.990621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T03:48:14.211240Z digest=sha256:3e8c9835cbdb051442166984d95fcd0c403afafb2268dcbeb75390a60acf1de4

Observation 47b3835b-a2f2-481c-9359-8a2863bdaeba · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-05T13:21:06.249364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-05T13:15:50.794969Z digest=sha256:68d01c94abafd32819a4efb17debe0a0a7c83ddcb89fde27938e8751d513d448

Observation 9709bd04-8be2-4fd4-92ee-2ab92d186317 · inbound

Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR cites this paper.

Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:00:28.015524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T13:59:47.796221Z digest=sha256:c189a166839bc3755139c6e80c4b3dc8b03e80d4c4c572faf0360199ea1b9722

Observation e68f615b-5a73-4fe1-a3a9-ba536ac47fcb · inbound

MedASR: An Open-Source Model for High-Accuracy Medical Dictation cites this paper.

MedASR: An Open-Source Model for High-Accuracy Medical Dictation SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:22:48.105606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T21:18:56.084055Z digest=sha256:fa50667d6d5b43e595b95dbc893e0063bc096a6dad3a93a8c5d70e3bb26a7ba5

Observation 49eb96ab-2f0c-4c38-bf8a-8cf1af2ea33d · inbound

Executable Boundary Contracts for Sound Event Traces cites this paper.

Executable Boundary Contracts for Sound Event Traces SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T02:12:58.552426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T02:08:44.066554Z digest=sha256:bd458785b5c78d5310e033846b72fa7c47b10b81dd28e1b9167011af88440c3d

Observation 4b9b1d7a-eb45-4a7f-814a-af4aca1be086 · inbound

EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs cites this paper.

EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:45:45.531103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T22:52:43.396119Z digest=sha256:97e87ade1e29c22ded6e581f54e03aeec7dc8ac42629debe86e026ee572ae880

Observation 1f05bee2-f28a-4b83-bae7-1d178923fe57 · inbound

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions cites this paper.

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T15:03:32.128909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T05:32:33.771381Z digest=sha256:d795a424ca406df802c2e4a65857315b0253624657a8430dd7896c6f187c1b04

Observation 13aaa9cb-755a-4873-9c5a-509752bcbdf6 · inbound

C2GA: A Class-Controllable Generative Augmentation Framework for Respiratory Sound Classification cites this paper.

C2GA: A Class-Controllable Generative Augmentation Framework for Respiratory Sound Classification SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:06:24.451197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T12:39:47.668919Z digest=sha256:36ca5447bd551a2910851671049dedf83a53fe656ab52a925c483e76610e014c

Observation bf4924ae-db79-4ebc-95db-b63f8ad0222d · inbound

TRADE: Transducer-Augmented Decoder for Speech LLM cites this paper.

TRADE: Transducer-Augmented Decoder for Speech LLM SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:25.721259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T18:40:19.688550Z digest=sha256:7642c1885aae1135e25f19ec7588fd644a2c9eb1100759b91e93aaef2e347572

Observation a5d0a818-1dd1-4aa8-bb32-944d347593b7 · inbound

Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource Settings cites this paper.

Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource Settings SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:59:26.680654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T20:03:28.491546Z digest=sha256:9d836caa61a3ff51db488975de8fcd782dcc4e5c5991ff5334e2f4c0cf587821

Observation 7c5a3410-c7d9-4c56-ae47-80359d2b001c · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T05:56:39.865166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:73ff4a2e3b28a9f91f82567f849ee5d701741ac8937a0cfa69a04b5a9fa0bb26

Observation 257f0093-4bf6-4879-a083-3610fe1fc45f · inbound

Physiological Noise Augmentation Improves Non-Invasive Brain-to-Speech cites this paper.

Physiological Noise Augmentation Improves Non-Invasive Brain-to-Speech SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T01:44:25.919968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-08T01:38:05.160539Z digest=sha256:543b9095c16078a66d9ea260563f7fdb3d927f5486c7863d31c43c04f9939b22