Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 21
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 21 of 31: papers 2,001 to 2,100 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System9 May 2021 0 repositories listed
-
Latency-Controlled Neural Architecture Search for Streaming Speech Recognition8 May 2021 0 repositories listed
-
Robustness of end-to-end Automatic Speech Recognition Models -- A Case Study using Mozilla DeepSpeech8 May 2021 0 repositories listed
-
Accent Recognition with Hybrid Phonetic Features5 May 2021 0 repositories listed
-
Spectral modification for recognition of children’s speech undermismatched conditions1 May 2021 0 repositories listed
-
Personalized Keyphrase Detection using Speaker and Environment Information28 Apr 2021 0 repositories listed
-
Head-synchronous Decoding for Transformer-based Streaming ASR26 Apr 2021 0 repositories listed
-
Multi-Task Learning for End-to-End ASR Word and Utterance Confidence with Deletion Prediction26 Apr 2021 0 repositories listed
-
Semantic Data Augmentation for End-to-End Mandarin Speech Recognition26 Apr 2021 0 repositories listed
-
Bridging the gap between streaming and non-streaming ASR systems bydistilling ensembles of CTC and RNN-T models25 Apr 2021 0 repositories listed
-
Quantization of Deep Neural Networks for Accurate Edge Computing25 Apr 2021 0 repositories listed
-
Accented Speech Recognition: A Survey21 Apr 2021 0 repositories listed
-
Discriminative Self-training for Punctuation Prediction21 Apr 2021 0 repositories listed
-
Disfluency Detection with Unlabeled Data and Small BERT Models21 Apr 2021 0 repositories listed
-
Label-Synchronous Speech-to-Text Alignment for ASR Using Forward and Backward Transformers21 Apr 2021 0 repositories listed
-
On Sampling-Based Training Criteria for Neural Language Modeling21 Apr 2021 0 repositories listed
-
Pre-training for Spoken Language Understanding with Joint Textual and Phonetic Representation Learning21 Apr 2021 0 repositories listed
-
Scene-aware Far-field Automatic Speech Recognition21 Apr 2021 0 repositories listed
-
On the Impact of Word Error Rate on Acoustic-Linguistic Speech Emotion Recognition: An Update for the Deep Learning Era20 Apr 2021 0 repositories listed
-
Acoustic Data-Driven Subword Modeling for End-to-End Speech Recognition19 Apr 2021 0 repositories listed
-
Advanced Long-context End-to-end Speech Recognition Using Context-expanded Transformers19 Apr 2021 0 repositories listed
-
MIMO Self-attentive RNN Beamformer for Multi-speaker Speech Separation17 Apr 2021 0 repositories listed
-
Bridging the Gap Between Clean Data Training and Real-World Inference for Spoken Language Understanding13 Apr 2021 0 repositories listed
-
Equivalence of Segmental and Neural Transducer Modeling: A Proof of Concept13 Apr 2021 0 repositories listed
-
Source and Target Bidirectional Knowledge Distillation for End-to-end Speech Translation13 Apr 2021 0 repositories listed
-
Comparing the Benefit of Synthetic Training Data for Various Automatic Speech Recognition Architectures12 Apr 2021 0 repositories listed
-
Improved Conformer-based End-to-End Speech Recognition Using Neural Architecture Search12 Apr 2021 0 repositories listed
-
Innovative Bert-based Reranking Language Models for Speech Recognition11 Apr 2021 0 repositories listed
-
Non-autoregressive Transformer-based End-to-end ASR using BERT10 Apr 2021 0 repositories listed
-
Accented Speech Recognition Inspired by Human Perception9 Apr 2021 0 repositories listed
-
On Architectures and Training for Raw Waveform Feature Extraction in ASR9 Apr 2021 0 repositories listed
-
8 Apr 2021 0 repositories listed
-
Contextual Semi-Supervised Learning: An Approach To Leverage Air-Surveillance and Untranscribed ATC Data in ASR Systems8 Apr 2021 0 repositories listed
-
Exploring Machine Speech Chain for Domain Adaptation and Few-Shot Speaker Adaptation8 Apr 2021 0 repositories listed
-
WNARS: WFST based Non-autoregressive Streaming End-to-End Speech Recognition8 Apr 2021 0 repositories listed
-
Capturing Multi-Resolution Context by Dilated Self-Attention7 Apr 2021 0 repositories listed
-
Pushing the Limits of Non-Autoregressive Speech Recognition7 Apr 2021 0 repositories listed
-
Comparing CTC and LFMMI for out-of-domain adaptation of wav2vec 2.0 acoustic model6 Apr 2021 0 repositories listed
-
Dissecting User-Perceived Latency of On-Device E2E Speech Recognition6 Apr 2021 0 repositories listed
-
Exploring Targeted Universal Adversarial Perturbations to End-to-end ASR Models6 Apr 2021 0 repositories listed
-
Flexi-Transducer: Optimizing Latency, Accuracy and Compute forMulti-Domain On-Device Scenarios6 Apr 2021 0 repositories listed
-
Relaxing the Conditional Independence Assumption of CTC-based ASR by Conditioning on Intermediate Predictions6 Apr 2021 0 repositories listed
-
Citrinet: Closing the Gap between Non-Autoregressive and Autoregressive End-to-End Models for Automatic Speech Recognition5 Apr 2021 0 repositories listed
-
End-to-End Speaker-Attributed ASR with Transformer5 Apr 2021 0 repositories listed
-
Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding5 Apr 2021 0 repositories listed
-
Speaker conditioned acoustic modeling for multi-speaker conversational ASR5 Apr 2021 0 repositories listed
-
Talk, Don't Write: A Study of Direct Speech-Based Image Retrieval5 Apr 2021 0 repositories listed
-
Towards Lifelong Learning of End-to-end ASR4 Apr 2021 0 repositories listed
-
Adversarial Joint Training with Self-Attention Mechanism for Robust End-to-End Speech Recognition3 Apr 2021 0 repositories listed
-
Configurable Privacy-Preserving Automatic Speech Recognition1 Apr 2021 0 repositories listed
-
Context-sensitive evaluation of automatic speech recognition: considering user experience & language variation1 Apr 2021 0 repositories listed
-
Dialect Identification through Adversarial Learning and Knowledge Distillation on Romanian BERT1 Apr 2021 0 repositories listed
-
Leveraging End-to-End ASR for Endangered Language Documentation: An Empirical Study on Yolóxochitl Mixtec1 Apr 2021 0 repositories listed
-
Tutorial Proposal: End-to-End Speech Translation1 Apr 2021 0 repositories listed
-
Adversarial Attacks and Defenses for Speech Recognition Systems31 Mar 2021 0 repositories listed
-
Large-Scale Pre-Training of End-to-End Multi-Talker ASR for Meeting Transcription with Single Distant Microphone31 Mar 2021 0 repositories listed
-
Multi-Encoder Learning and Stream Fusion for Transformer-Based End-to-End Automatic Speech Recognition31 Mar 2021 0 repositories listed
-
Multiple-hypothesis CTC-based semi-supervised adaptation of end-to-end speech recognition29 Mar 2021 0 repositories listed
-
BART based semantic correction for Mandarin automatic speech recognition system26 Mar 2021 0 repositories listed
-
26 Mar 2021 0 repositories listed
-
Mutually-Constrained Monotonic Multihead Attention for Online ASR26 Mar 2021 0 repositories listed
-
An Approach to Improve Robustness of NLP Systems against ASR Errors25 Mar 2021 0 repositories listed
-
Residual Energy-Based Models for End-to-End Speech Recognition25 Mar 2021 0 repositories listed
-
Voice Privacy with Smart Digital Assistants in Educational Settings24 Mar 2021 0 repositories listed
-
Hallucination of speech recognition errors with sequence to sequence learning23 Mar 2021 0 repositories listed
-
Contextual Biasing of Language Models for Speech Recognition in Goal-Oriented Conversational Agents18 Mar 2021 0 repositories listed
-
17 Mar 2021 0 repositories listed
-
14 Mar 2021 0 repositories listed
-
A Distributed Optimisation Framework Combining Natural Gradient with Hessian-Free for Discriminative Sequence Training12 Mar 2021 0 repositories listed
-
Dynamic Acoustic Unit Augmentation With BPE-Dropout for Low-Resource End-to-End Speech Recognition12 Mar 2021 0 repositories listed
-
Learning Word-Level Confidence For Subword End-to-End ASR11 Mar 2021 0 repositories listed
-
Best of Both Worlds: Robust Accented Speech Recognition with Adversarial Transfer Learning10 Mar 2021 0 repositories listed
-
Contrastive Semi-supervised Learning for ASR9 Mar 2021 0 repositories listed
-
An Ultra-low Power RNN Classifier for Always-On Voice Wake-Up Detection Robust to Real-World Scenarios8 Mar 2021 0 repositories listed
-
Neural model robustness for skill routing in large-scale conversational AI systems: A design choice exploration4 Mar 2021 0 repositories listed
-
Incorporating VAD into ASR System by Multi-task Learning2 Mar 2021 0 repositories listed
-
Alignment Knowledge Distillation for Online Streaming Attention-based Speech Recognition28 Feb 2021 0 repositories listed
-
Brain Signals to Rescue Aphasia, Apraxia and Dysarthria Speech Recognition28 Feb 2021 0 repositories listed
-
Meta-Learning for improving rare word recognition in end-to-end ASR25 Feb 2021 0 repositories listed
-
MixSpeech: Data Augmentation for Low-resource Automatic Speech Recognition25 Feb 2021 0 repositories listed
-
Speech Enhancement Using Multi-Stage Self-Attentive Temporal Convolutional Networks24 Feb 2021 0 repositories listed
-
Thoughts on the potential to compensate a hearing loss in noise24 Feb 2021 0 repositories listed
-
Evolutionary optimization of contexts for phonetic correction in speech recognition systems23 Feb 2021 0 repositories listed
-
Generating Human Readable Transcript for Automatic Speech Recognition with Pre-trained Language Model22 Feb 2021 0 repositories listed
-
Echo State Speech Recognition18 Feb 2021 0 repositories listed
-
Fundamental Frequency Feature Normalization and Data Augmentation for Child Speech Recognition18 Feb 2021 0 repositories listed
-
Gaussian Kernelized Self-Attention for Long Sequence Data and Its Application to CTC-based Speech Recognition18 Feb 2021 0 repositories listed
-
ATCSpeechNet: A multilingual end-to-end speech recognition framework for air traffic control systems17 Feb 2021 0 repositories listed
-
Deep Learning based Multi-Source Localization with Source Splitting and its Effectiveness in Multi-Talker Speech Recognition16 Feb 2021 0 repositories listed
-
End-to-End Automatic Speech Recognition with Deep Mutual Learning16 Feb 2021 0 repositories listed
-
Hierarchical Transformer-based Large-Context End-to-end ASR with Large-Context Knowledge Distillation16 Feb 2021 0 repositories listed
-
Improving speech recognition models with small samples for air traffic control systems16 Feb 2021 0 repositories listed
-
Thank you for Attention: A survey on Attention-based Artificial Neural Networks for Automatic Speech Recognition14 Feb 2021 0 repositories listed
-
Bi-APC: Bidirectional Autoregressive Predictive Coding for Unsupervised Pre-training and Its Application to Children's ASR12 Feb 2021 0 repositories listed
-
Content-Aware Speaker Embeddings for Speaker Diarisation12 Feb 2021 0 repositories listed
-
Do as I mean, not as I say: Sequence Loss Training for Spoken Language Understanding12 Feb 2021 0 repositories listed
-
Multimodal Punctuation Prediction with Contextual Dropout12 Feb 2021 0 repositories listed
-
NUVA: A Naming Utterance Verifier for Aphasia Treatment10 Feb 2021 0 repositories listed
-
Sparsification via Compressed Sensing for Automatic Speech Recognition9 Feb 2021 0 repositories listed
-
Train your classifier first: Cascade Neural Networks Training from upper layers to lower layers9 Feb 2021 0 repositories listed