Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 22
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 22 of 32: papers 2,101 to 2,200 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Spell my name: keyword boosted speech recognition6 Oct 2021 0 repositories listed
-
ASR Rescoring and Confidence Estimation with ELECTRA5 Oct 2021 0 repositories listed
-
Fast Contextual Adaptation with Neural Associative Memory for On-Device Personalized Speech Recognition5 Oct 2021 0 repositories listed
-
5 Oct 2021 0 repositories listed
-
Building a Noisy Audio Dataset to Evaluate Machine Learning Approaches for Automatic Speech Recognition Systems4 Oct 2021 0 repositories listed
-
Exploiting Pre-Trained ASR Models for Alzheimer's Disease Recognition Through Spontaneous Speech4 Oct 2021 0 repositories listed
-
Towards efficient end-to-end speech recognition with biologically-inspired neural networks4 Oct 2021 0 repositories listed
-
Chinese Medical Speech Recognition with Punctuated Hypothesis1 Oct 2021 0 repositories listed
-
Employing low-pass filtered temporal speech features for the training of ideal ratio mask in speech enhancement1 Oct 2021 0 repositories listed
-
Exploring the Integration of E2E ASR and Pronunciation Modeling for English Mispronunciation Detection1 Oct 2021 0 repositories listed
-
Improving Punctuation Restoration for Speech Transcripts via External Data1 Oct 2021 0 repositories listed
-
Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English with Transfer Learning1 Oct 2021 0 repositories listed
-
SpliceOut: A Simple and Efficient Audio Augmentation Method30 Sep 2021 0 repositories listed
-
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch29 Sep 2021 0 repositories listed
-
Conditioning Sequence-to-sequence Networks with Learned Activations29 Sep 2021 0 repositories listed
-
Demystifying Limited Adversarial Transferability in Automatic Speech Recognition Systems29 Sep 2021 0 repositories listed
-
MLP-based architecture with variable length input for automatic speech recognition29 Sep 2021 0 repositories listed
-
PhaseFool: Phase-oriented Audio Adversarial Examples via Energy Dissipation29 Sep 2021 0 repositories listed
-
Synthesising Audio Adversarial Examples for Automatic Speech Recognition29 Sep 2021 0 repositories listed
-
Understanding the Role of Self Attention for Efficient Speech Recognition29 Sep 2021 0 repositories listed
-
W-CTC: a Connectionist Temporal Classification Loss with Wild Cards29 Sep 2021 0 repositories listed
-
Private Language Model Adaptation for Speech Recognition28 Sep 2021 0 repositories listed
-
Word-level confidence estimation for RNN transducers28 Sep 2021 0 repositories listed
-
27 Sep 2021 0 repositories listed
-
Challenges and Opportunities of Speech Recognition for Bengali Language27 Sep 2021 0 repositories listed
-
Topic Model Robustness to Automatic Speech Recognition Errors in Podcast Transcripts25 Sep 2021 0 repositories listed
-
Learning Domain Specific Language Models for Automatic Speech Recognition through Machine Translation21 Sep 2021 0 repositories listed
-
Audio-Visual Speech Recognition is Worth 32×32×8 Voxels20 Sep 2021 0 repositories listed
-
iRNN: Integer-only Recurrent Neural Network20 Sep 2021 0 repositories listed
-
MeetDot: Videoconferencing with Live Translation Captions20 Sep 2021 0 repositories listed
-
Model-Based Approach for Measuring the Fairness in ASR19 Sep 2021 0 repositories listed
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding17 Sep 2021 0 repositories listed
-
PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription16 Sep 2021 0 repositories listed
-
Utterance-level neural confidence measure for end-to-end children speech recognition16 Sep 2021 0 repositories listed
-
Improving Accent Identification and Accented Speech Recognition Under a Framework of Self-supervised Learning15 Sep 2021 0 repositories listed
-
Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning15 Sep 2021 0 repositories listed
-
Non-autoregressive Transformer with Unified Bidirectional Decoder for Automatic Speech Recognition14 Sep 2021 0 repositories listed
-
Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech14 Sep 2021 0 repositories listed
-
Unsupervised Domain Adaptation Schemes for Building ASR in Low-resource Languages12 Sep 2021 0 repositories listed
-
Remember the context! ASR slot error correction through memorization10 Sep 2021 0 repositories listed
-
Self-Attention Channel Combinator Frontend for End-to-End Multichannel Far-field Speech Recognition10 Sep 2021 0 repositories listed
-
Coarse-To-Fine And Cross-Lingual ASR Transfer2 Sep 2021 0 repositories listed
-
Robustness of end-to-end Automatic Speech Recognition Models – A Case Study using Mozilla DeepSpeech1 Sep 2021 0 repositories listed
-
Tree-constrained Pointer Generator for End-to-end Contextual Speech Recognition1 Sep 2021 0 repositories listed
-
30 Aug 2021 0 repositories listed
-
Investigations on Speech Recognition Systems for Low-Resource Dialectal Arabic-English Code-Switching Speech29 Aug 2021 0 repositories listed
-
4-bit Quantization of LSTM-based Speech Recognition Models27 Aug 2021 0 repositories listed
-
Grammar Based Speaker Role Identification for Air Traffic Control Speech Recognition27 Aug 2021 0 repositories listed
-
Improving callsign recognition with air-surveillance data in air-traffic communication27 Aug 2021 0 repositories listed
-
Task-aware Warping Factors in Mask-based Speech Enhancement27 Aug 2021 0 repositories listed
-
Cross-domain Single-channel Speech Enhancement Model with Bi-projection Fusion Module for Noise-robust ASR26 Aug 2021 0 repositories listed
-
Reducing Exposure Bias in Training Recurrent Neural Network Transducers24 Aug 2021 0 repositories listed
-
A Unified Transformer-based Framework for Duplex Text Normalization23 Aug 2021 0 repositories listed
-
Automatic Speech Recognition And Limited Vocabulary: A Survey23 Aug 2021 0 repositories listed
-
Hierarchical Summarization for Longform Spoken Dialog21 Aug 2021 0 repositories listed
-
A Multi-level Acoustic Feature Extraction Framework for Transformer Based End-to-End Speech Recognition18 Aug 2021 0 repositories listed
-
A Light-weight contextual spelling correction model for customizing transducer-based speech recognition systems17 Aug 2021 0 repositories listed
-
StarGAN-VC+ASR: StarGAN-based Non-Parallel Voice Conversion Regularized by Automatic Speech Recognition10 Aug 2021 0 repositories listed
-
The HW-TSC's Offline Speech Translation Systems for IWSLT 2021 Evaluation9 Aug 2021 0 repositories listed
-
Blind and neural network-guided convolutional beamformer for joint denoising, dereverberation, and source separation4 Aug 2021 0 repositories listed
-
Dyn-ASR: Compact, Multilingual Speech Recognition via Spoken Language and Accent Identification4 Aug 2021 0 repositories listed
-
Improving Distinction between ASR Errors and Speech Disfluencies with Feature Space Interpolation4 Aug 2021 0 repositories listed
-
Unsupervised Domain Adaptation in Speech Recognition using Phonetic Features4 Aug 2021 0 repositories listed
-
3 Aug 2021 0 repositories listed
-
Learning a Neural Diff for Speech Models3 Aug 2021 0 repositories listed
-
Automatic recognition of suprasegmentals in speech2 Aug 2021 0 repositories listed
-
Decoupling recognition and transcription in Mandarin ASR2 Aug 2021 0 repositories listed
-
BTS: Back TranScription for Speech-to-Text Post-Processor using Text-to-Speech-to-Text1 Aug 2021 0 repositories listed
-
How Might We Create Better Benchmarks for Speech Recognition?1 Aug 2021 0 repositories listed
-
IMS’ Systems for the IWSLT 2021 Low-Resource Speech Translation Task1 Aug 2021 0 repositories listed
-
Interactive Reinforcement Learning for Table Balancing Robot1 Aug 2021 0 repositories listed
-
On Knowledge Distillation for Translating Erroneous Speech Transcriptions1 Aug 2021 0 repositories listed
-
ON-TRAC’ systems for the IWSLT 2021 low-resource speech translation and multilingual speech translation shared tasks1 Aug 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus1 Aug 2021 0 repositories listed
-
Technology-Augmented Multilingual Communication Models: New Interaction Paradigms, Shifts in the Language Services Industry, and Implications for Training Programs1 Aug 2021 0 repositories listed
-
Without Further Ado: Direct and Simultaneous Speech Translation by AppTek in 20211 Aug 2021 0 repositories listed
-
ZJU’s IWSLT 2021 Speech Translation System1 Aug 2021 0 repositories listed
-
Can You Hear It? Backdoor Attacks via Ultrasonic Triggers30 Jul 2021 0 repositories listed
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition29 Jul 2021 0 repositories listed
-
An Adapter Based Pre-Training for Efficient and Scalable Self-Supervised Speech Representation Learning26 Jul 2021 0 repositories listed
-
Facetron: A Multi-speaker Face-to-Speech Model based on Cross-modal Latent Representations26 Jul 2021 0 repositories listed
-
23 Jul 2021 0 repositories listed
-
CarneliNet: Neural Mixture Model for Automatic Speech Recognition22 Jul 2021 0 repositories listed
-
Multitask-Based Joint Learning Approach To Robust ASR For Radio Communication Speech22 Jul 2021 0 repositories listed
-
On Prosody Modeling for ASR+TTS based Voice Conversion20 Jul 2021 0 repositories listed
-
A baseline model for computationally inexpensive speech recognition for Kazakh using the Coqui STT framework19 Jul 2021 0 repositories listed
-
Multi-task Learning with Cross Attention for Keyword Spotting15 Jul 2021 0 repositories listed
-
VAD-free Streaming Hybrid CTC/Attention ASR for Unsegmented Recording15 Jul 2021 0 repositories listed
-
A Configurable Multilingual Model is All You Need to Recognize All Languages13 Jul 2021 0 repositories listed
-
The IWSLT 2021 BUT Speech Translation Systems13 Jul 2021 0 repositories listed
-
Zero-shot Speech Translation13 Jul 2021 0 repositories listed
-
Perceptual-based deep-learning denoiser as a defense against adversarial attacks on ASR systems12 Jul 2021 0 repositories listed
-
Loss Prediction: End-to-End Active Learning Approach For Speech Recognition9 Jul 2021 0 repositories listed
-
Noisy Training Improves E2E ASR for the Edge9 Jul 2021 0 repositories listed
-
On lattice-free boosted MMI training of HMM and CTC-based full-context ASR models9 Jul 2021 0 repositories listed
-
Improved Language Identification Through Cross-Lingual Self-Supervised Learning8 Jul 2021 0 repositories listed
-
End-to-End Rich Transcription-Style Automatic Speech Recognition with Semi-Supervised Learning7 Jul 2021 0 repositories listed
-
A Comparative Study of Modular and Joint Approaches for Speaker-Attributed ASR on Monaural Long-Form Audio6 Jul 2021 0 repositories listed
-
Investigation of Practical Aspects of Single Channel Speech Separation for ASR5 Jul 2021 0 repositories listed
-
Arabic Code-Switching Speech Recognition using Monolingual Data4 Jul 2021 0 repositories listed