Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 22
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 22 of 31: papers 2,101 to 2,200 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A bandit approach to curriculum generation for automatic speech recognition6 Feb 2021 0 repositories listed
-
Intermediate Loss Regularization for CTC-based Speech Recognition5 Feb 2021 0 repositories listed
-
Multi-Task Self-Supervised Pre-Training for Music Classification5 Feb 2021 0 repositories listed
-
Two-Stage Augmentation and Adaptive CTC Fusion for Improved Robustness of Multi-Stream End-to-End ASR5 Feb 2021 0 repositories listed
-
Effects of Number of Filters of Convolutional Layers on Speech Recognition Model Accuracy3 Feb 2021 0 repositories listed
-
Internal Language Model Training for Domain-Adaptive End-to-End Speech Recognition2 Feb 2021 0 repositories listed
-
Speech Recognition by Simply Fine-tuning BERT30 Jan 2021 0 repositories listed
-
BCN2BRNO: ASR System Fusion for Albayzin 2020 Speech to Text Challenge29 Jan 2021 0 repositories listed
-
Leveraging End-to-End ASR for Endangered Language Documentation: An Empirical Study on Yoloxóchitl Mixtec26 Jan 2021 0 repositories listed
-
Exploiting Beam Search Confidence for Energy-Efficient Speech Recognition22 Jan 2021 0 repositories listed
-
Streaming Models for Joint Speech Recognition and Translation22 Jan 2021 0 repositories listed
-
Efficiently Fusing Pretrained Acoustic and Linguistic Encoders for Low-resource Speech Recognition17 Jan 2021 0 repositories listed
-
An evaluation of word-level confidence estimation for end-to-end automatic speech recognition14 Jan 2021 0 repositories listed
-
Fast offline Transformer-based end-to-end automatic speech recognition for real-world applications14 Jan 2021 0 repositories listed
-
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm14 Jan 2021 0 repositories listed
-
Hypothesis Stitcher for End-to-End Speaker-attributed ASR on Long-form Multi-talker Recordings6 Jan 2021 0 repositories listed
-
Learning without Forgetting: Task Aware Multitask Learning for Multi-Modality Tasks1 Jan 2021 0 repositories listed
-
NAS-Bench-ASR: Reproducible Neural Architecture Search for Speech Recognition1 Jan 2021 0 repositories listed
-
Why Does Decentralized Training Outperform Synchronous Training In The Large Batch Setting?1 Jan 2021 0 repositories listed
-
Multi-channel Multi-frame ADL-MVDR for Target Speech Separation24 Dec 2020 0 repositories listed
-
A Hierarchical Reasoning Graph Neural Network for The Automatic Scoring of Answer Transcriptions in Video Job Interviews22 Dec 2020 0 repositories listed
-
Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition22 Dec 2020 0 repositories listed
-
Adjust-free adversarial example generation in speech recognition using evolutionary multi-objective optimization under black-box condition21 Dec 2020 0 repositories listed
-
Toward Streaming ASR with Non-Autoregressive Insertion-based Model18 Dec 2020 0 repositories listed
-
15 Dec 2020 0 repositories listed
-
User-friendly automatic transcription of low-resource languages: Plugging ESPnet into Elpis15 Dec 2020 0 repositories listed
-
A review of on-device fully neural end-to-end automatic speech recognition algorithms14 Dec 2020 0 repositories listed
-
Less Is More: Improved RNN-T Decoding Using Limited Label Context and Path Merging12 Dec 2020 0 repositories listed
-
Improved Robustness to Disfluencies in RNN-Transducer Based Speech Recognition11 Dec 2020 0 repositories listed
-
Bayesian Learning of LF-MMI Trained Time Delay Neural Networks for Speech Recognition8 Dec 2020 0 repositories listed
-
Using multiple ASR hypotheses to boost i18n NLU performance7 Dec 2020 0 repositories listed
-
1 Dec 2020 0 repositories listed
-
ASR for Non-standardised Languages with Dialectal Variation: the case of Swiss German1 Dec 2020 0 repositories listed
-
German-Arabic Speech-to-Speech Translation for Psychiatric Diagnosis1 Dec 2020 0 repositories listed
-
Multi-task Learning of Spoken Language Understanding by Integrating N-Best Hypotheses with Hierarchical Attention1 Dec 2020 0 repositories listed
-
On-Device detection of sentence completion for voice assistants with low-memory footprint1 Dec 2020 0 repositories listed
-
Sparse Transcription1 Dec 2020 0 repositories listed
-
The Indigenous Languages Technology project at NRC Canada: An empowerment-oriented approach to developing language software1 Dec 2020 0 repositories listed
-
Improving accuracy of rare words for RNN-Transducer through unigram shallow fusion30 Nov 2020 0 repositories listed
-
Transformer-Transducers for Code-Switched Speech Recognition30 Nov 2020 0 repositories listed
-
Unsupervised Domain Adaptation for Speech Recognition via Uncertainty Driven Self-Training26 Nov 2020 0 repositories listed
-
Bootstrap an end-to-end ASR system by multilingual training, transfer learning, text-to-text mapping and synthetic audio25 Nov 2020 0 repositories listed
-
Adam^+: A Stochastic Method with Adaptive Variance Reduction24 Nov 2020 0 repositories listed
-
Multi-task Language Modeling for Improving Speech Recognition of Rare Words23 Nov 2020 0 repositories listed
-
Using Synthetic Audio to Improve The Recognition of Out-Of-Vocabulary Words in End-To-End ASR Systems23 Nov 2020 0 repositories listed
-
Improving RNN-T ASR Accuracy Using Context Audio20 Nov 2020 0 repositories listed
-
Cascade RNN-Transducer: Syllable Based Streaming On-device Mandarin Speech Recognition with a Syllable-to-Character Converter17 Nov 2020 0 repositories listed
-
Refining Automatic Speech Recognition System for older adults17 Nov 2020 0 repositories listed
-
Audio-visual Multi-channel Integration and Recognition of Overlapped Speech16 Nov 2020 0 repositories listed
-
Deep Shallow Fusion for RNN-T Personalization16 Nov 2020 0 repositories listed
-
Improving Speech Enhancement Performance by Leveraging Contextual Broad Phonetic Class Information15 Nov 2020 0 repositories listed
-
Self-supervised reinforcement learning for speaker localisation with the iCub humanoid robot12 Nov 2020 0 repositories listed
-
Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS10 Nov 2020 0 repositories listed
-
Benchmarking LF-MMI, CTC and RNN-T Criteria for Streaming ASR9 Nov 2020 0 repositories listed
-
Gated Recurrent Fusion with Joint Training Framework for Robust End-to-End Speech Recognition9 Nov 2020 0 repositories listed
-
Personalized Query Rewriting in Conversational AI Agents9 Nov 2020 0 repositories listed
-
On the Usefulness of Self-Attention for Automatic Speech Recognition with Transformers8 Nov 2020 0 repositories listed
-
Acoustics Based Intent Recognition Using Discovered Phonetic Units for Low Resource Languages7 Nov 2020 0 repositories listed
-
ESPnet-se: end-to-end speech enhancement and separation toolkit designed for asr integration7 Nov 2020 0 repositories listed
-
Alignment Restricted Streaming Recurrent Neural Network Transducer5 Nov 2020 0 repositories listed
-
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework4 Nov 2020 0 repositories listed
-
Data Augmentation for End-to-end Code-switching Speech Recognition4 Nov 2020 0 repositories listed
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
Sequence-to-Sequence Learning via Attention Transfer for Incremental Speech Recognition4 Nov 2020 0 repositories listed
-
Improving RNN transducer with normalized jointer network3 Nov 2020 0 repositories listed
-
Integration of speech separation, diarization, and recognition for multi-speaker meetings: System description, comparison, and analysis3 Nov 2020 0 repositories listed
-
Internal Language Model Estimation for Domain-Adaptive End-to-End Speech Recognition3 Nov 2020 0 repositories listed
-
Streaming Attention-Based Models with Augmented Memory for End-to-End Speech Recognition3 Nov 2020 0 repositories listed
-
Unsupervised Pattern Discovery from Thematic Speech Archives Based on Multilingual Bottleneck Features3 Nov 2020 0 repositories listed
-
Warped Language Models for Noise Robust Language Understanding3 Nov 2020 0 repositories listed
-
DNN-Based Semantic Model for Rescoring N-best Speech Recognition List2 Nov 2020 0 repositories listed
-
SapAugment: Learning A Sample Adaptive Policy for Data Augmentation2 Nov 2020 0 repositories listed
-
Effectively pretraining a speech translation decoder with Machine Translation data1 Nov 2020 0 repositories listed
-
ELITR: European Live Translator1 Nov 2020 0 repositories listed
-
Impact of ASR on Alzheimer’s Disease Detection: All Errors are Equal, but Deletions are More Equal than Others1 Nov 2020 0 repositories listed
-
Improving End-to-End Bangla Speech Recognition with Semi-supervised Training1 Nov 2020 0 repositories listed
-
Directional ASR: A New Paradigm for E2E Multi-Speaker Speech Recognition with Source Localization30 Oct 2020 0 repositories listed
-
Streaming Simultaneous Speech Translation with Augmented Memory Transformer30 Oct 2020 0 repositories listed
-
Semi-Supervised Speech Recognition via Graph-based Temporal Classification29 Oct 2020 0 repositories listed
-
Decoupling Pronunciation and Language for End-to-end Code-switching Automatic Speech Recognition28 Oct 2020 0 repositories listed
-
Fusion Models for Improved Visual Captioning28 Oct 2020 0 repositories listed
-
INT8 Winograd Acceleration for Conv1D Equipped ASR Models Deployed on Mobile Devices28 Oct 2020 0 repositories listed
-
Non-Autoregressive Transformer ASR with CTC-Enhanced Decoder Input28 Oct 2020 0 repositories listed
-
Cascaded encoders for unifying streaming and non-streaming ASR27 Oct 2020 0 repositories listed
-
Effective Decoder Masking for Transformer Based End-to-End Speech Recognition27 Oct 2020 0 repositories listed
-
Emotion recognition by fusing time synchronous and time asynchronous representations27 Oct 2020 0 repositories listed
-
Improved Mask-CTC for Non-Autoregressive End-to-End ASR26 Oct 2020 0 repositories listed
-
Improving Noise Robustness of an End-to-End Neural Model for Automatic Speech Recognition23 Oct 2020 0 repositories listed
-
Improving Streaming Automatic Speech Recognition With Non-Streaming Model Distillation On Unsupervised Data22 Oct 2020 0 repositories listed
-
MAM: Masked Acoustic Modeling for End-to-End Speech-to-Text Translation22 Oct 2020 0 repositories listed
-
SlimIPL: Language-Model-Free Iterative Pseudo-Labeling22 Oct 2020 0 repositories listed
-
A General Multi-Task Learning Framework to Leverage Text Data for Speech to Text Tasks21 Oct 2020 0 repositories listed
-
Cascaded Models With Cyclic Feedback For Direct Speech Translation21 Oct 2020 0 repositories listed
-
Knowledge Distillation for Improved Accuracy in Spoken Question Answering21 Oct 2020 0 repositories listed
-
Sentence Boundary Augmentation For Neural Machine Translation Robustness21 Oct 2020 0 repositories listed
-
Knowledge Transfer for Efficient On-device False Trigger Mitigation20 Oct 2020 0 repositories listed
-
Replacing Human Audio with Synthetic Audio for On-device Unspoken Punctuation Prediction20 Oct 2020 0 repositories listed
-
Ensemble Chinese End-to-End Spoken Language Understanding for Abnormal Event Detection from audio stream19 Oct 2020 0 repositories listed
-
Towards Data Distillation for End-to-end Spoken Conversational Question Answering18 Oct 2020 0 repositories listed
-
Studying the Similarity of COVID-19 Sounds based on Correlation Analysis of MFCC17 Oct 2020 0 repositories listed