Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 16
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 16 of 31: papers 1,501 to 1,600 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Investigating the Impact of Cross-lingual Acoustic-Phonetic Similarities on Multilingual Speech Recognition7 Jul 2022 0 repositories listed
-
Improving Streaming End-to-End ASR on Transformer-based Causal Models with Encoder States Revision Strategies6 Jul 2022 0 repositories listed
-
Compute Cost Amortized Transformer for Streaming ASR5 Jul 2022 0 repositories listed
-
Vietnamese Capitalization and Punctuation Recovery Models4 Jul 2022 0 repositories listed
-
Leveraging Acoustic Contextual Representation by Audio-textual Cross-modal Learning for Conversational ASR3 Jul 2022 0 repositories listed
-
Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism2 Jul 2022 0 repositories listed
-
Tree-constrained Pointer Generator with Graph Neural Network Encodings for Contextual Speech Recognition2 Jul 2022 0 repositories listed
-
Activity focused Speech Recognition of Preschool Children in Early Childhood Classrooms1 Jul 2022 0 repositories listed
-
Exploring the Effect of Dialect Mismatched Language Models in Telugu Automatic Speech Recognition1 Jul 2022 0 repositories listed
-
Improving Low-Resource Speech Recognition with Pretrained Speech Models: Continued Pretraining vs. Semi-Supervised Training1 Jul 2022 0 repositories listed
-
Non-Autoregressive Chinese ASR Error Correction with Phonological Training1 Jul 2022 0 repositories listed
-
Updating Only Encoders Prevents Catastrophic Forgetting of End-to-End ASR Models1 Jul 2022 0 repositories listed
-
FeaRLESS: Feature Refinement Loss for Ensembling Self-Supervised Learning Features in Robust End-to-end Speech Recognition30 Jun 2022 0 repositories listed
-
Space-Efficient Representation of Entity-centric Query Language Models29 Jun 2022 0 repositories listed
-
The THUEE System Description for the IARPA OpenASR21 Challenge29 Jun 2022 0 repositories listed
-
Bengali Common Voice Speech Dataset for Automatic Speech Recognition28 Jun 2022 0 repositories listed
-
Challenges and Opportunities in Multi-device Speech Processing27 Jun 2022 0 repositories listed
-
TALCS: An Open-Source Mandarin-English Code-Switching Corpus and a Speech Recognition Baseline27 Jun 2022 0 repositories listed
-
Annotated Speech Corpus for Low Resource Indian Languages: Awadhi, Bhojpuri, Braj and Magahi26 Jun 2022 0 repositories listed
-
Improving the Training Recipe for a Robust Conformer-based Hybrid Model26 Jun 2022 0 repositories listed
-
Meta Auxiliary Learning for Low-resource Spoken Language Understanding26 Jun 2022 0 repositories listed
-
On Comparison of Encoders for Attention based End to End Speech Recognition in Standalone and Rescoring Mode26 Jun 2022 0 repositories listed
-
Confidence Score Based Conformer Speaker Adaptation for Speech Recognition24 Jun 2022 0 repositories listed
-
Two-pass Decoding and Cross-adaptation Based System Combination of End-to-end Conformer and Hybrid TDNN ASR Systems23 Jun 2022 0 repositories listed
-
A Simple Baseline for Domain Adaptation in End to End ASR Systems Using Synthetic Data22 Jun 2022 0 repositories listed
-
Supervision-Guided Codebooks for Masked Prediction in Speech Pre-training21 Jun 2022 0 repositories listed
-
The Makerere Radio Speech Corpus: A Luganda Radio Corpus for Automatic Speech Recognition20 Jun 2022 0 repositories listed
-
Transfer Learning for Robust Low-Resource Children's Speech ASR with Transformers and Source-Filter Warping19 Jun 2022 0 repositories listed
-
Decoupled Federated Learning for ASR with Non-IID Data18 Jun 2022 0 repositories listed
-
A CTC Triggered Siamese Network with Spatial-Temporal Dropout for Speech Recognition16 Jun 2022 0 repositories listed
-
DRAFT: A Novel Framework to Reduce Domain Shifting in Self-supervised Learning and Its Application to Children's ASR16 Jun 2022 0 repositories listed
-
Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recognition15 Jun 2022 0 repositories listed
-
Exploring Capabilities of Monolingual Audio Transformers using Large Datasets in Automatic Speech Recognition of Czech15 Jun 2022 0 repositories listed
-
Residual Language Model for End-to-end Speech Recognition15 Jun 2022 0 repositories listed
-
The ZevoMOS entry to VoiceMOS Challenge 202215 Jun 2022 0 repositories listed
-
Transformer-based Automatic Speech Recognition of Formal and Colloquial Czech in MALACH Project15 Jun 2022 0 repositories listed
-
Toward Zero Oracle Word Error Rate on the Switchboard Benchmark13 Jun 2022 0 repositories listed
-
Investigation of Ensemble features of Self-Supervised Pretrained Models for Automatic Speech Recognition11 Jun 2022 0 repositories listed
-
Context-based out-of-vocabulary word recovery for ASR systems in Indian languages9 Jun 2022 0 repositories listed
-
Face-Dubbing++: Lip-Synchronous, Voice Preserving Translation of Videos9 Jun 2022 0 repositories listed
-
Joint Encoder-Decoder Self-Supervised Pre-training for ASR9 Jun 2022 0 repositories listed
-
LegoNN: Building Modular Encoder-Decoder Models7 Jun 2022 0 repositories listed
-
FedNST: Federated Noisy Student Training for Automatic Speech Recognition6 Jun 2022 0 repositories listed
-
Pronunciation Dictionary-Free Multilingual Speech Synthesis by Combining Unsupervised and Supervised Phonetic Representations2 Jun 2022 0 repositories listed
-
A Semi-Automated Live Interlingual Communication Workflow Featuring Intralingual Respeaking: Evaluation and Benchmarking1 Jun 2022 0 repositories listed
-
Automatic Speech Recognition for Irish: the ABAIR-ÉIST System1 Jun 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Jun 2022 0 repositories listed
-
Building Open-source Speech Technology for Low-resource Minority Languages with SáMi as an Example – Tools, Methods and Experiments1 Jun 2022 0 repositories listed
-
Conversational Speech Recognition Needs Data? Experiments with Austrian German1 Jun 2022 0 repositories listed
-
Developing Automatic Speech Recognition for Scottish Gaelic1 Jun 2022 0 repositories listed
-
Development of Automatic Speech Recognition for the Documentation of Cook Islands Māori1 Jun 2022 0 repositories listed
-
Evaluation of Off-the-shelf Speech Recognizers on Different Accents in a Dialogue Domain1 Jun 2022 0 repositories listed
-
Generating Synthetic Clinical Speech Data through Simulated ASR Deletion Error1 Jun 2022 0 repositories listed
-
Huqariq: A Multilingual Speech Corpus of Native Languages of Peru forSpeech Recognition1 Jun 2022 0 repositories listed
-
LiSTra Automatic Speech Translation: English to Lingala Case Study1 Jun 2022 0 repositories listed
-
Mesures linguistiques automatiques pour l’évaluation des systèmes de Reconnaissance Automatique de la Parole (Automated linguistic measures for automatic speech recognition systems’ evaluation)1 Jun 2022 0 repositories listed
-
Multilingual Transfer Learning for Children Automatic Speech Recognition1 Jun 2022 0 repositories listed
-
ParlamentParla: A Speech Corpus of Catalan Parliamentary Sessions1 Jun 2022 0 repositories listed
-
ParlaSpeech-HR - a Freely Available ASR Dataset for Croatian Bootstrapped from the ParlaMint Corpus1 Jun 2022 0 repositories listed
-
Post-Stroke Speech Transcription Challenge (Task B): Correctness Detection in Anomia Diagnosis with Imperfect Transcripts1 Jun 2022 0 repositories listed
-
Progress in Multilingual Speech Recognition for Low Resource Languages Kurmanji Kurdish, Cree and Inuktut1 Jun 2022 0 repositories listed
-
Samrómur Children: An Icelandic Speech Corpus1 Jun 2022 0 repositories listed
-
Snow Mountain: Dataset of Audio Recordings of The Bible in Low Resource Languages1 Jun 2022 0 repositories listed
-
Towards a Unified ASR System for the Armenian Standards1 Jun 2022 0 repositories listed
-
Towards an Open-Source Dutch Speech Recognition System for the Healthcare Domain1 Jun 2022 0 repositories listed
-
Adversarial synthesis based data-augmentation for code-switched spoken language identification30 May 2022 0 repositories listed
-
Adaptive Activation Network For Low Resource Multilingual Speech Recognition28 May 2022 0 repositories listed
-
Acoustic-to-articulatory Speech Inversion with Multi-task Learning27 May 2022 0 repositories listed
-
Punctuation Restoration in Spanish Customer Support Transcripts using Transfer Learning27 May 2022 0 repositories listed
-
Clinical Dialogue Transcription Error Correction using Seq2Seq Models26 May 2022 0 repositories listed
-
Contextual Adapters for Personalized Speech Recognition in Neural Transducers26 May 2022 0 repositories listed
-
Joint Training of Speech Enhancement and Self-supervised Model for Noise-robust ASR26 May 2022 0 repositories listed
-
An Investigation on Applying Acoustic Feature Conversion to ASR of Adult and Child Speech25 May 2022 0 repositories listed
-
Heterogeneous Reservoir Computing Models for Persian Speech Recognition25 May 2022 0 repositories listed
-
Improving CTC-based ASR Models with Gated Interlayer Collaboration25 May 2022 0 repositories listed
-
Investigating Lexical Replacements for Arabic-English Code-Switched Data Augmentation25 May 2022 0 repositories listed
-
On Building Spoken Language Understanding Systems for Low Resourced Languages25 May 2022 0 repositories listed
-
Multi-Level Modeling Units for End-to-End Mandarin Speech Recognition24 May 2022 0 repositories listed
-
Calibrate and Refine! A Novel and Agile Framework for ASR-error Robust Intent Detection23 May 2022 0 repositories listed
-
Self-Supervised Speech Representation Learning: A Review21 May 2022 0 repositories listed
-
Automatic Spoken Language Identification using a Time-Delay Neural Network19 May 2022 0 repositories listed
-
Insights on Neural Representations for End-to-End Speech Recognition19 May 2022 0 repositories listed
-
Deploying self-supervised learning in the wild for hybrid automatic speech recognition17 May 2022 0 repositories listed
-
Streaming Noise Context Aware Enhancement For Automatic Speech Recognition in Multi-Talker Environments17 May 2022 0 repositories listed
-
Improved Consistency Training for Semi-Supervised Sequence-to-Sequence ASR via Speech Chain Reconstruction and Self-Transcribing14 May 2022 0 repositories listed
-
Pretraining Approaches for Spoken Language Recognition: TalTech Submission to the OLR 2021 Challenge14 May 2022 0 repositories listed
-
Personalized Adversarial Data Augmentation for Dysarthric and Elderly Speech Recognition13 May 2022 0 repositories listed
-
Unified Modeling of Multi-Domain Multi-Device ASR Systems13 May 2022 0 repositories listed
-
A Closer Look at Audio-Visual Multi-Person Speech Recognition and Active Speaker Selection11 May 2022 0 repositories listed
-
End-to-End Multi-Person Audio/Visual Automatic Speech Recognition11 May 2022 0 repositories listed
-
Best of Both Worlds: Multi-task Audio-Visual Automatic Speech Recognition and Active Speaker Detection10 May 2022 0 repositories listed
-
Speaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition9 May 2022 0 repositories listed
-
A Conformer-based Waveform-domain Neural Acoustic Echo Canceller Optimized for ASR Accuracy6 May 2022 0 repositories listed
-
ON-TRAC Consortium Systems for the IWSLT 2022 Dialect and Low-resource Speech Translation Tasks4 May 2022 0 repositories listed
-
A Meeting Transcription System for an Ad-Hoc Acoustic Sensor Network2 May 2022 0 repositories listed
-
Bilingual End-to-End ASR with Byte-Level Subwords1 May 2022 0 repositories listed
-
Discourse on ASR Measurement: Introducing the ARPOCA Assessment Tool1 May 2022 0 repositories listed
-
Enhancing Documentation of Hupa with Automatic Speech Recognition1 May 2022 0 repositories listed
-
Findings of the Shared Task on Speech Recognition for Vulnerable Individuals in Tamil1 May 2022 0 repositories listed
-
Fine-tuning pre-trained models for Automatic Speech Recognition, experiments on a fieldwork corpus of Japhug (Trans-Himalayan family)1 May 2022 0 repositories listed