Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 7
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 7 of 31: papers 601 to 700 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
5 Jun 2018 1 repository listed
-
20 May 2018 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
1 May 2018 1 repository listed
-
28 Apr 2018 1 repository listed
-
3 Apr 2018 1 repository listed
-
26 Mar 2018 1 repository listed
-
9 Feb 2018 1 repository listed
-
7 Feb 2018 1 repository listed
-
2 Jan 2018 1 repository listed
-
13 Sep 2017 1 repository listed
-
6 Sep 2017 1 repository listed
-
16 Aug 2017 1 repository listed
-
4 Aug 2017 1 repository listed
-
4 Jul 2017 1 repository listed
-
3 Jul 2017 1 repository listed
-
1 Jul 2017 1 repository listed
-
1 May 2017 1 repository listed
-
1 Dec 2016 1 repository listed
-
1 Nov 2016 1 repository listed
-
1 May 2016 1 repository listed
-
5 Apr 2016 1 repository listed
-
NonverbalTTS: A Public English Corpus of Text-Aligned Nonverbal Vocalizations with Emotion Annotations for Text-to-Speech17 Jul 2025 0 repositories listed
-
WhisperKit: On-device Real-time ASR with Billion-Scale Transformers14 Jul 2025 0 repositories listed
-
Lightweight Target-Speaker-Based Overlap Transcription for Practical Streaming ASR25 Jun 2025 0 repositories listed
-
AI-Generated Song Detection via Lyrics Transcripts23 Jun 2025 0 repositories listed
-
End-to-End Spoken Grammatical Error Correction23 Jun 2025 0 repositories listed
-
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages20 Jun 2025 0 repositories listed
-
LM-SPT: LM-Aligned Semantic Distillation for Speech Tokenization20 Jun 2025 0 repositories listed
-
Automatic Speech Recognition Biases in Newcastle English: an Error Analysis19 Jun 2025 0 repositories listed
-
Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios17 Jun 2025 0 repositories listed
-
Bi-directional Context-Enhanced Speech Large Language Models for Multilingual Conversational ASR16 Jun 2025 0 repositories listed
-
BUT System for the MLC-SLM Challenge16 Jun 2025 0 repositories listed
-
Seewo's Submission to MLC-SLM: Lessons learned from Speech Reasoning Language Models16 Jun 2025 0 repositories listed
-
Enabling automatic transcription of child-centered audio recordings from real-world environments13 Jun 2025 0 repositories listed
-
Lightweight and Robust Multi-Channel End-to-End Speech Recognition with Spherical Harmonic Transform13 Jun 2025 0 repositories listed
-
(SimPhon Speech Test): A Data-Driven Method for In Silico Design and Validation of a Phonetically Balanced Speech Test13 Jun 2025 0 repositories listed
-
Improving Named Entity Transcription with Contextual LLM-based Revision12 Jun 2025 0 repositories listed
-
Regularizing Learnable Feature Extraction for Automatic Speech Recognition11 Jun 2025 0 repositories listed
-
Benchmarking Foundation Speech and Language Models for Alzheimer's Disease and Related Dementia Detection from Spontaneous Speech9 Jun 2025 0 repositories listed
-
Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation9 Jun 2025 0 repositories listed
-
Speech Recognition on TV Series with Video-guided Post-Correction8 Jun 2025 0 repositories listed
-
Automatic Speech Recognition of African American English: Lexical and Contextual Effects7 Jun 2025 0 repositories listed
-
Bridging the Modality Gap: Softly Discretizing Audio Representation for LLM-based Automatic Speech Recognition6 Jun 2025 0 repositories listed
-
Lightweight Prompt Biasing for Contextualized End-to-End ASR Systems6 Jun 2025 0 repositories listed
-
Low-Resource Domain Adaptation for Speech LLMs via Text-Only Fine-Tuning6 Jun 2025 0 repositories listed
-
Better Pseudo-labeling with Multi-ASR Fusion and Error Correction by SpeechLLM5 Jun 2025 0 repositories listed
-
Customizing Speech Recognition Model with Large Language Model Feedback5 Jun 2025 0 repositories listed
-
LESS: Large Language Model Enhanced Semi-Supervised Learning for Speech Foundational Models5 Jun 2025 0 repositories listed
-
LLM-based phoneme-to-grapheme for phoneme-based speech recognition5 Jun 2025 0 repositories listed
-
Effects of Speaker Count, Duration, and Accent Diversity on Zero-Shot Accent Robustness in Low-Resource ASR4 Jun 2025 0 repositories listed
-
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation3 Jun 2025 0 repositories listed
-
Enhancing Lyrics Transcription on Music Mixtures with Consistency Loss3 Jun 2025 0 repositories listed
-
Overcoming Data Scarcity in Multi-Dialectal Arabic ASR via Whisper Fine-Tuning3 Jun 2025 0 repositories listed
-
DNCASR: End-to-End Training for Speaker-Attributed ASR2 Jun 2025 0 repositories listed
-
HENT-SRT: Hierarchical Efficient Neural Transducer with Self-Distillation for Joint Speech Recognition and Translation2 Jun 2025 0 repositories listed
-
Causal Structure Discovery for Error Diagnostics of Children's ASR31 May 2025 0 repositories listed
-
Dynamic Context-Aware Streaming Pretrained Language Model For Inverse Text Normalization30 May 2025 0 repositories listed
-
Fewer Hallucinations, More Verification: A Three-Stage LLM-Based Framework for ASR Error Correction30 May 2025 0 repositories listed
-
Improving Multilingual Speech Models on ML-SUPERB 2.0: Fine-tuning with Data Augmentation and LID-Aware CTC30 May 2025 0 repositories listed
-
MSDA: Combining Pseudo-labeling and Self-Supervision for Unsupervised Domain Adaptation in ASR30 May 2025 0 repositories listed
-
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation29 May 2025 0 repositories listed
-
Prompting Whisper for Improved Verbatim Transcription and End-to-end Miscue Detection29 May 2025 0 repositories listed
-
Advancing Hearing Assessment: An ASR-Based Frequency-Specific Speech Test for Diagnosing Presbycusis28 May 2025 0 repositories listed
-
NGPU-LM: GPU-Accelerated N-Gram Language Model for Context-Biasing in Greedy ASR Decoding28 May 2025 0 repositories listed
-
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use27 May 2025 0 repositories listed
-
PSRB: A Comprehensive Benchmark for Evaluating Persian ASR Systems27 May 2025 0 repositories listed
-
Towards Pretraining Robust ASR Foundation Model with Acoustic-Aware Data Augmentation27 May 2025 0 repositories listed
-
Beyond Manual Transcripts: The Potential of Automated Speech Recognition Errors in Improving Alzheimer's Disease Detection26 May 2025 0 repositories listed
-
Continuous Learning for Children's ASR: Overcoming Catastrophic Forgetting with Elastic Weight Consolidation and Synaptic Intelligence26 May 2025 0 repositories listed
-
In-context Language Learning for Endangered Languages in Speech Recognition26 May 2025 0 repositories listed
-
KIT's Low-resource Speech Translation Systems for IWSLT2025: System Enhancement with Synthetic Data and Model Regularization26 May 2025 0 repositories listed
-
Mixture of LoRA Experts for Low-Resourced Multi-Accent Automatic Speech Recognition26 May 2025 0 repositories listed
-
Robust fine-tuning of speech recognition models via model merging: application to disordered speech26 May 2025 0 repositories listed
-
VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining23 May 2025 0 repositories listed
-
An Effective Training Framework for Light-Weight Automatic Speech Recognition Models22 May 2025 0 repositories listed
-
Large Language Models based ASR Error Correction for Child Conversations22 May 2025 0 repositories listed
-
SoccerChat: Integrating Multimodal Data for Enhanced Soccer Game Understanding22 May 2025 0 repositories listed
-
From Weak Labels to Strong Results: Utilizing 5,000 Hours of Noisy Classroom Transcripts with Minimal Accurate Data20 May 2025 0 repositories listed
-
In-Context Learning Boosts Speech Recognition via Human-like Adaptation to Speakers and Language Varieties20 May 2025 0 repositories listed
-
Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio16 May 2025 0 repositories listed
-
LegoSLM: Connecting LLM with Speech Encoder using CTC Posteriors16 May 2025 0 repositories listed
-
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems16 May 2025 0 repositories listed
-
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions16 May 2025 0 repositories listed
-
LipDiffuser: Lip-to-Speech Generation with Conditional Diffusion Models16 May 2025 0 repositories listed
-
Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients9 May 2025 0 repositories listed
-
Teochew-Wild: The First In-the-wild Teochew Dataset with Orthographic Annotations8 May 2025 0 repositories listed
-
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech6 May 2025 0 repositories listed
-
SepALM: Audio Language Models Are Error Correctors for Robust Speech Separation6 May 2025 0 repositories listed
-
Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments2 May 2025 0 repositories listed
-
Retrieval-Enhanced Few-Shot Prompting for Speech Event Extraction30 Apr 2025 0 repositories listed
-
Chinese-LiPS: A Chinese audio-visual speech recognition dataset with Lip-reading and Presentation Slides21 Apr 2025 0 repositories listed
-
StableQuant: Layer Adaptive Post-Training Quantization for Speech Foundation Models21 Apr 2025 0 repositories listed
-
Acoustic to Articulatory Inversion of Speech; Data Driven Approaches, Challenges, Applications, and Future Scope17 Apr 2025 0 repositories listed
-
Advancing Arabic Speech Recognition Through Large-Scale Weakly Supervised Learning16 Apr 2025 0 repositories listed
-
Spatial Audio Processing with Large Language Model on Wearable Devices11 Apr 2025 0 repositories listed
-
Visual-Aware Speech Recognition for Noisy Scenarios9 Apr 2025 0 repositories listed
-
LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect3 Apr 2025 0 repositories listed
-
Chain of Correction for Full-text Speech Recognition with Large Language Models2 Apr 2025 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.