Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 23
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 23 of 32: papers 2,201 to 2,300 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Cross-Modal Transformer-Based Neural Correction Models for Automatic Speech Recognition4 Jul 2021 0 repositories listed
-
Unified Autoregressive Modeling for Joint End-to-End Multi-Talker Overlapped Speech Recognition and Speaker Attribute Estimation4 Jul 2021 0 repositories listed
-
Dual Causal/Non-Causal Self-Attention for Streaming End-to-End Speech Recognition2 Jul 2021 0 repositories listed
-
Multi-user VoiceFilter-Lite via Attentive Speaker Embedding2 Jul 2021 0 repositories listed
-
Improving Named Entity Recognition in Spoken Dialog Systems by Context and Speech Pattern Modeling1 Jul 2021 0 repositories listed
-
SmarTerp: A CAI System to Support Simultaneous Interpreters in Real-Time1 Jul 2021 0 repositories listed
-
StableEmit: Selection Probability Discount for Reducing Emission Latency of Streaming Monotonic Attention ASR1 Jul 2021 0 repositories listed
-
Word-Free Spoken Language Understanding for Mandarin-Chinese1 Jul 2021 0 repositories listed
-
On joint training with interfaces for spoken language understanding30 Jun 2021 0 repositories listed
-
IMS' Systems for the IWSLT 2021 Low-Resource Speech Translation Task30 Jun 2021 0 repositories listed
-
Sequence-level Confidence Classifier for ASR Utterance Accuracy and Application to Acoustic Models30 Jun 2021 0 repositories listed
-
Rethinking End-to-End Evaluation of Decomposable Tasks: A Case Study on Spoken Language Understanding29 Jun 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus24 Jun 2021 0 repositories listed
-
Where are we in semantic concept extraction for Spoken Language Understanding?24 Jun 2021 0 repositories listed
-
Mixtures of Deep Neural Experts for Automated Speech Scoring23 Jun 2021 0 repositories listed
-
Zero-Shot Joint Modeling of Multiple Spoken-Text-Style Conversion Tasks using Switching Tokens23 Jun 2021 0 repositories listed
-
A Discriminative Entity-Aware Language Model for Virtual Assistants21 Jun 2021 0 repositories listed
-
An Improved Single Step Non-autoregressive Transformer for Automatic Speech Recognition18 Jun 2021 0 repositories listed
-
Low Resource German ASR with Untranscribed Data Spoken by Non-native Children -- INTERSPEECH 2021 Shared Task SPAPL System18 Jun 2021 0 repositories listed
-
On-Device Personalization of Automatic Speech Recognition Models for Disordered Speech18 Jun 2021 0 repositories listed
-
Layer Pruning on Demand with Intermediate CTC17 Jun 2021 0 repositories listed
-
Multi-mode Transformer Transducer with Stochastic Future Context17 Jun 2021 0 repositories listed
-
Topic Classification on Spoken Documents Using Deep Acoustic and Linguistic Features16 Jun 2021 0 repositories listed
-
A Study into Pre-training Strategies for Spoken Language Understanding on Dysarthric Speech15 Jun 2021 0 repositories listed
-
ASR Adaptation for E-commerce Chatbots using Cross-Utterance Context and Multi-Task Language Modeling15 Jun 2021 0 repositories listed
-
Dialectal Speech Recognition and Translation of Swiss German Speech to Standard German Text: Microsoft's Submission to SwissText 202115 Jun 2021 0 repositories listed
-
Multi-channel Opus compression for far-field automatic speech recognition with a fixed bitrate budget15 Jun 2021 0 repositories listed
-
Overcoming Domain Mismatch in Low Resource Sequence-to-Sequence ASR Models using Hybrid Generated Pseudotranscripts14 Jun 2021 0 repositories listed
-
SynthASR: Unlocking Synthetic Data for Speech Recognition14 Jun 2021 0 repositories listed
-
Using heterogeneity in semi-supervised transcription hypotheses to improve code-switched speech recognition14 Jun 2021 0 repositories listed
-
Cross-utterance Reranking Models with BERT and Graph Convolutional Networks for Conversational Speech Recognition13 Jun 2021 0 repositories listed
-
Improving RNN-T ASR Performance with Date-Time and Location Awareness11 Jun 2021 0 repositories listed
-
Leveraging Pre-trained Language Model for Speech Sentiment Analysis11 Jun 2021 0 repositories listed
-
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition10 Jun 2021 0 repositories listed
-
10 Jun 2021 0 repositories listed
-
A Comparative Study on Neural Architectures and Training Methods for Japanese Speech Recognition9 Jun 2021 0 repositories listed
-
Unsupervised Automatic Speech Recognition: A Review9 Jun 2021 0 repositories listed
-
8 Jun 2021 0 repositories listed
-
Data Augmentation Methods for End-to-end Speech Recognition on Distant-Talk Scenarios7 Jun 2021 0 repositories listed
-
Human Listening and Live Captioning: Multi-Task Training for Speech Enhancement5 Jun 2021 0 repositories listed
-
Do You Listen with One or Two Microphones? A Unified ASR Model for Single and Multi-Channel Audio4 Jun 2021 0 repositories listed
-
Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability3 Jun 2021 0 repositories listed
-
Dual Script E2E framework for Multilingual and Code-Switching ASR2 Jun 2021 0 repositories listed
-
Improving low-resource ASR performance with untranscribed out-of-domain data2 Jun 2021 0 repositories listed
-
Should We Always Separate?: Switching Between Enhanced and Observed Signals for Overlapping Speech Recognition2 Jun 2021 0 repositories listed
-
A Neural Acoustic Echo Canceller Optimized Using An Automatic Speech Recognizer And Large Scale Synthetic Data1 Jun 2021 0 repositories listed
-
Developing ASR for Indonesian-English Bilingual Language Teaching1 Jun 2021 0 repositories listed
-
End-to-end ASR to jointly predict transcriptions and linguistic annotations1 Jun 2021 0 repositories listed
-
End-to-End Automatic Speech Recognition: Its Impact on the Workflowin Documenting Yoloxóchitl Mixtec1 Jun 2021 0 repositories listed
-
Evaluating Automatic Speech Recognition Quality and Its Impact on Counselor Utterance Coding1 Jun 2021 0 repositories listed
-
Highland Puebla Nahuatl Speech Translation Corpus for Endangered Language Documentation1 Jun 2021 0 repositories listed
-
Towards One Model to Rule All: Multilingual Strategy for Dialectal Code-Switching Arabic ASR31 May 2021 0 repositories listed
-
Training Speech Enhancement Systems with Noisy Speech Datasets26 May 2021 0 repositories listed
-
Mondegreen: A Post-Processing Solution to Speech Recognition Error Correction for Voice Search Queries20 May 2021 0 repositories listed
-
LiSTra, Automatic Speech Translation: English to Lingala case study16 May 2021 0 repositories listed
-
Listen with Intent: Improving Speech Recognition with Audio-to-Intent Front-End14 May 2021 0 repositories listed
-
Streaming Transformer for Hardware Efficient Voice Trigger Detection and False Trigger Mitigation14 May 2021 0 repositories listed
-
Exploring CTC Based End-to-End Techniques for Myanmar Speech Recognition13 May 2021 0 repositories listed
-
Stacked Acoustic-and-Textual Encoding: Integrating the Pre-trained Models into Speech Translation Encoders12 May 2021 0 repositories listed
-
StutterNet: Stuttering Detection Using Time Delay Neural Network12 May 2021 0 repositories listed
-
10 May 2021 0 repositories listed
-
English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System9 May 2021 0 repositories listed
-
Latency-Controlled Neural Architecture Search for Streaming Speech Recognition8 May 2021 0 repositories listed
-
Robustness of end-to-end Automatic Speech Recognition Models -- A Case Study using Mozilla DeepSpeech8 May 2021 0 repositories listed
-
Accent Recognition with Hybrid Phonetic Features5 May 2021 0 repositories listed
-
Spectral modification for recognition of children’s speech undermismatched conditions1 May 2021 0 repositories listed
-
Personalized Keyphrase Detection using Speaker and Environment Information28 Apr 2021 0 repositories listed
-
Head-synchronous Decoding for Transformer-based Streaming ASR26 Apr 2021 0 repositories listed
-
Multi-Task Learning for End-to-End ASR Word and Utterance Confidence with Deletion Prediction26 Apr 2021 0 repositories listed
-
Semantic Data Augmentation for End-to-End Mandarin Speech Recognition26 Apr 2021 0 repositories listed
-
Bridging the gap between streaming and non-streaming ASR systems bydistilling ensembles of CTC and RNN-T models25 Apr 2021 0 repositories listed
-
Quantization of Deep Neural Networks for Accurate Edge Computing25 Apr 2021 0 repositories listed
-
Accented Speech Recognition: A Survey21 Apr 2021 0 repositories listed
-
Discriminative Self-training for Punctuation Prediction21 Apr 2021 0 repositories listed
-
Disfluency Detection with Unlabeled Data and Small BERT Models21 Apr 2021 0 repositories listed
-
Label-Synchronous Speech-to-Text Alignment for ASR Using Forward and Backward Transformers21 Apr 2021 0 repositories listed
-
On Sampling-Based Training Criteria for Neural Language Modeling21 Apr 2021 0 repositories listed
-
Pre-training for Spoken Language Understanding with Joint Textual and Phonetic Representation Learning21 Apr 2021 0 repositories listed
-
Scene-aware Far-field Automatic Speech Recognition21 Apr 2021 0 repositories listed
-
On the Impact of Word Error Rate on Acoustic-Linguistic Speech Emotion Recognition: An Update for the Deep Learning Era20 Apr 2021 0 repositories listed
-
Acoustic Data-Driven Subword Modeling for End-to-End Speech Recognition19 Apr 2021 0 repositories listed
-
Advanced Long-context End-to-end Speech Recognition Using Context-expanded Transformers19 Apr 2021 0 repositories listed
-
Best Practices for Noise-Based Augmentation to Improve the Performance of Deployable Speech-Based Emotion Recognition Systems18 Apr 2021 0 repositories listed
-
MIMO Self-attentive RNN Beamformer for Multi-speaker Speech Separation17 Apr 2021 0 repositories listed
-
Bridging the Gap Between Clean Data Training and Real-World Inference for Spoken Language Understanding13 Apr 2021 0 repositories listed
-
Equivalence of Segmental and Neural Transducer Modeling: A Proof of Concept13 Apr 2021 0 repositories listed
-
Source and Target Bidirectional Knowledge Distillation for End-to-end Speech Translation13 Apr 2021 0 repositories listed
-
Comparing the Benefit of Synthetic Training Data for Various Automatic Speech Recognition Architectures12 Apr 2021 0 repositories listed
-
Improved Conformer-based End-to-End Speech Recognition Using Neural Architecture Search12 Apr 2021 0 repositories listed
-
Innovative Bert-based Reranking Language Models for Speech Recognition11 Apr 2021 0 repositories listed
-
Non-autoregressive Transformer-based End-to-end ASR using BERT10 Apr 2021 0 repositories listed
-
Accented Speech Recognition Inspired by Human Perception9 Apr 2021 0 repositories listed
-
On Architectures and Training for Raw Waveform Feature Extraction in ASR9 Apr 2021 0 repositories listed
-
8 Apr 2021 0 repositories listed
-
Contextual Semi-Supervised Learning: An Approach To Leverage Air-Surveillance and Untranscribed ATC Data in ASR Systems8 Apr 2021 0 repositories listed
-
Exploring Machine Speech Chain for Domain Adaptation and Few-Shot Speaker Adaptation8 Apr 2021 0 repositories listed
-
WNARS: WFST based Non-autoregressive Streaming End-to-End Speech Recognition8 Apr 2021 0 repositories listed
-
Capturing Multi-Resolution Context by Dilated Self-Attention7 Apr 2021 0 repositories listed
-
Pushing the Limits of Non-Autoregressive Speech Recognition7 Apr 2021 0 repositories listed
-
Comparing CTC and LFMMI for out-of-domain adaptation of wav2vec 2.0 acoustic model6 Apr 2021 0 repositories listed