Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 24
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 24 of 32: papers 2,301 to 2,400 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Dissecting User-Perceived Latency of On-Device E2E Speech Recognition6 Apr 2021 0 repositories listed
-
Exploring Targeted Universal Adversarial Perturbations to End-to-end ASR Models6 Apr 2021 0 repositories listed
-
Flexi-Transducer: Optimizing Latency, Accuracy and Compute forMulti-Domain On-Device Scenarios6 Apr 2021 0 repositories listed
-
Relaxing the Conditional Independence Assumption of CTC-based ASR by Conditioning on Intermediate Predictions6 Apr 2021 0 repositories listed
-
Citrinet: Closing the Gap between Non-Autoregressive and Autoregressive End-to-End Models for Automatic Speech Recognition5 Apr 2021 0 repositories listed
-
End-to-End Speaker-Attributed ASR with Transformer5 Apr 2021 0 repositories listed
-
Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding5 Apr 2021 0 repositories listed
-
Speaker conditioned acoustic modeling for multi-speaker conversational ASR5 Apr 2021 0 repositories listed
-
Talk, Don't Write: A Study of Direct Speech-Based Image Retrieval5 Apr 2021 0 repositories listed
-
Towards Lifelong Learning of End-to-end ASR4 Apr 2021 0 repositories listed
-
Adversarial Joint Training with Self-Attention Mechanism for Robust End-to-End Speech Recognition3 Apr 2021 0 repositories listed
-
Configurable Privacy-Preserving Automatic Speech Recognition1 Apr 2021 0 repositories listed
-
Context-sensitive evaluation of automatic speech recognition: considering user experience & language variation1 Apr 2021 0 repositories listed
-
Dialect Identification through Adversarial Learning and Knowledge Distillation on Romanian BERT1 Apr 2021 0 repositories listed
-
Leveraging End-to-End ASR for Endangered Language Documentation: An Empirical Study on Yolóxochitl Mixtec1 Apr 2021 0 repositories listed
-
Tutorial Proposal: End-to-End Speech Translation1 Apr 2021 0 repositories listed
-
Adversarial Attacks and Defenses for Speech Recognition Systems31 Mar 2021 0 repositories listed
-
Large-Scale Pre-Training of End-to-End Multi-Talker ASR for Meeting Transcription with Single Distant Microphone31 Mar 2021 0 repositories listed
-
Multi-Encoder Learning and Stream Fusion for Transformer-Based End-to-End Automatic Speech Recognition31 Mar 2021 0 repositories listed
-
Multiple-hypothesis CTC-based semi-supervised adaptation of end-to-end speech recognition29 Mar 2021 0 repositories listed
-
BART based semantic correction for Mandarin automatic speech recognition system26 Mar 2021 0 repositories listed
-
26 Mar 2021 0 repositories listed
-
Mutually-Constrained Monotonic Multihead Attention for Online ASR26 Mar 2021 0 repositories listed
-
An Approach to Improve Robustness of NLP Systems against ASR Errors25 Mar 2021 0 repositories listed
-
Residual Energy-Based Models for End-to-End Speech Recognition25 Mar 2021 0 repositories listed
-
Voice Privacy with Smart Digital Assistants in Educational Settings24 Mar 2021 0 repositories listed
-
Hallucination of speech recognition errors with sequence to sequence learning23 Mar 2021 0 repositories listed
-
Contextual Biasing of Language Models for Speech Recognition in Goal-Oriented Conversational Agents18 Mar 2021 0 repositories listed
-
17 Mar 2021 0 repositories listed
-
14 Mar 2021 0 repositories listed
-
A Distributed Optimisation Framework Combining Natural Gradient with Hessian-Free for Discriminative Sequence Training12 Mar 2021 0 repositories listed
-
Dynamic Acoustic Unit Augmentation With BPE-Dropout for Low-Resource End-to-End Speech Recognition12 Mar 2021 0 repositories listed
-
Learning Word-Level Confidence For Subword End-to-End ASR11 Mar 2021 0 repositories listed
-
Best of Both Worlds: Robust Accented Speech Recognition with Adversarial Transfer Learning10 Mar 2021 0 repositories listed
-
Contrastive Semi-supervised Learning for ASR9 Mar 2021 0 repositories listed
-
An Ultra-low Power RNN Classifier for Always-On Voice Wake-Up Detection Robust to Real-World Scenarios8 Mar 2021 0 repositories listed
-
Neural model robustness for skill routing in large-scale conversational AI systems: A design choice exploration4 Mar 2021 0 repositories listed
-
Incorporating VAD into ASR System by Multi-task Learning2 Mar 2021 0 repositories listed
-
Alignment Knowledge Distillation for Online Streaming Attention-based Speech Recognition28 Feb 2021 0 repositories listed
-
Brain Signals to Rescue Aphasia, Apraxia and Dysarthria Speech Recognition28 Feb 2021 0 repositories listed
-
Meta-Learning for improving rare word recognition in end-to-end ASR25 Feb 2021 0 repositories listed
-
MixSpeech: Data Augmentation for Low-resource Automatic Speech Recognition25 Feb 2021 0 repositories listed
-
Speech Enhancement Using Multi-Stage Self-Attentive Temporal Convolutional Networks24 Feb 2021 0 repositories listed
-
Thoughts on the potential to compensate a hearing loss in noise24 Feb 2021 0 repositories listed
-
Evolutionary optimization of contexts for phonetic correction in speech recognition systems23 Feb 2021 0 repositories listed
-
Generating Human Readable Transcript for Automatic Speech Recognition with Pre-trained Language Model22 Feb 2021 0 repositories listed
-
Echo State Speech Recognition18 Feb 2021 0 repositories listed
-
Fundamental Frequency Feature Normalization and Data Augmentation for Child Speech Recognition18 Feb 2021 0 repositories listed
-
Gaussian Kernelized Self-Attention for Long Sequence Data and Its Application to CTC-based Speech Recognition18 Feb 2021 0 repositories listed
-
ATCSpeechNet: A multilingual end-to-end speech recognition framework for air traffic control systems17 Feb 2021 0 repositories listed
-
Deep Learning based Multi-Source Localization with Source Splitting and its Effectiveness in Multi-Talker Speech Recognition16 Feb 2021 0 repositories listed
-
End-to-End Automatic Speech Recognition with Deep Mutual Learning16 Feb 2021 0 repositories listed
-
Hierarchical Transformer-based Large-Context End-to-end ASR with Large-Context Knowledge Distillation16 Feb 2021 0 repositories listed
-
Improving speech recognition models with small samples for air traffic control systems16 Feb 2021 0 repositories listed
-
Thank you for Attention: A survey on Attention-based Artificial Neural Networks for Automatic Speech Recognition14 Feb 2021 0 repositories listed
-
Bi-APC: Bidirectional Autoregressive Predictive Coding for Unsupervised Pre-training and Its Application to Children's ASR12 Feb 2021 0 repositories listed
-
Content-Aware Speaker Embeddings for Speaker Diarisation12 Feb 2021 0 repositories listed
-
Do as I mean, not as I say: Sequence Loss Training for Spoken Language Understanding12 Feb 2021 0 repositories listed
-
Multimodal Punctuation Prediction with Contextual Dropout12 Feb 2021 0 repositories listed
-
NUVA: A Naming Utterance Verifier for Aphasia Treatment10 Feb 2021 0 repositories listed
-
Sparsification via Compressed Sensing for Automatic Speech Recognition9 Feb 2021 0 repositories listed
-
Train your classifier first: Cascade Neural Networks Training from upper layers to lower layers9 Feb 2021 0 repositories listed
-
A bandit approach to curriculum generation for automatic speech recognition6 Feb 2021 0 repositories listed
-
Intermediate Loss Regularization for CTC-based Speech Recognition5 Feb 2021 0 repositories listed
-
Multi-Task Self-Supervised Pre-Training for Music Classification5 Feb 2021 0 repositories listed
-
Two-Stage Augmentation and Adaptive CTC Fusion for Improved Robustness of Multi-Stream End-to-End ASR5 Feb 2021 0 repositories listed
-
Effects of Number of Filters of Convolutional Layers on Speech Recognition Model Accuracy3 Feb 2021 0 repositories listed
-
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition2 Feb 2021 0 repositories listed
-
Speech Recognition by Simply Fine-tuning BERT30 Jan 2021 0 repositories listed
-
BCN2BRNO: ASR System Fusion for Albayzin 2020 Speech to Text Challenge29 Jan 2021 0 repositories listed
-
Leveraging End-to-End ASR for Endangered Language Documentation: An Empirical Study on Yoloxóchitl Mixtec26 Jan 2021 0 repositories listed
-
Exploiting Beam Search Confidence for Energy-Efficient Speech Recognition22 Jan 2021 0 repositories listed
-
Streaming Models for Joint Speech Recognition and Translation22 Jan 2021 0 repositories listed
-
Efficiently Fusing Pretrained Acoustic and Linguistic Encoders for Low-resource Speech Recognition17 Jan 2021 0 repositories listed
-
An evaluation of word-level confidence estimation for end-to-end automatic speech recognition14 Jan 2021 0 repositories listed
-
Fast offline Transformer-based end-to-end automatic speech recognition for real-world applications14 Jan 2021 0 repositories listed
-
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm14 Jan 2021 0 repositories listed
-
Hypothesis Stitcher for End-to-End Speaker-attributed ASR on Long-form Multi-talker Recordings6 Jan 2021 0 repositories listed
-
Learning without Forgetting: Task Aware Multitask Learning for Multi-Modality Tasks1 Jan 2021 0 repositories listed
-
NAS-Bench-ASR: Reproducible Neural Architecture Search for Speech Recognition1 Jan 2021 0 repositories listed
-
Why Does Decentralized Training Outperform Synchronous Training In The Large Batch Setting?1 Jan 2021 0 repositories listed
-
Multi-channel Multi-frame ADL-MVDR for Target Speech Separation24 Dec 2020 0 repositories listed
-
A Hierarchical Reasoning Graph Neural Network for The Automatic Scoring of Answer Transcriptions in Video Job Interviews22 Dec 2020 0 repositories listed
-
Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition22 Dec 2020 0 repositories listed
-
Adjust-free adversarial example generation in speech recognition using evolutionary multi-objective optimization under black-box condition21 Dec 2020 0 repositories listed
-
Toward Streaming ASR with Non-Autoregressive Insertion-based Model18 Dec 2020 0 repositories listed
-
15 Dec 2020 0 repositories listed
-
User-friendly automatic transcription of low-resource languages: Plugging ESPnet into Elpis15 Dec 2020 0 repositories listed
-
A review of on-device fully neural end-to-end automatic speech recognition algorithms14 Dec 2020 0 repositories listed
-
Less Is More: Improved RNN-T Decoding Using Limited Label Context and Path Merging12 Dec 2020 0 repositories listed
-
Improved Robustness to Disfluencies in RNN-Transducer Based Speech Recognition11 Dec 2020 0 repositories listed
-
Bayesian Learning of LF-MMI Trained Time Delay Neural Networks for Speech Recognition8 Dec 2020 0 repositories listed
-
Using multiple ASR hypotheses to boost i18n NLU performance7 Dec 2020 0 repositories listed
-
1 Dec 2020 0 repositories listed
-
ASR for Non-standardised Languages with Dialectal Variation: the case of Swiss German1 Dec 2020 0 repositories listed
-
German-Arabic Speech-to-Speech Translation for Psychiatric Diagnosis1 Dec 2020 0 repositories listed
-
Multi-task Learning of Spoken Language Understanding by Integrating N-Best Hypotheses with Hierarchical Attention1 Dec 2020 0 repositories listed
-
On-Device detection of sentence completion for voice assistants with low-memory footprint1 Dec 2020 0 repositories listed
-
Sparse Transcription1 Dec 2020 0 repositories listed
-
The Indigenous Languages Technology project at NRC Canada: An empowerment-oriented approach to developing language software1 Dec 2020 0 repositories listed