Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 23
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 23 of 31: papers 2,201 to 2,300 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Multimodal Speech Recognition with Unstructured Audio Masking16 Oct 2020 0 repositories listed
-
Non-intrusive speech intelligibility prediction using automatic speech recognition derived measures16 Oct 2020 0 repositories listed
-
Lightweight End-to-End Speech Recognition from Raw Audio Data Using Sinc-Convolutions15 Oct 2020 0 repositories listed
-
Exploiting Spectral Augmentation for Code-Switched Spoken Language Identification14 Oct 2020 0 repositories listed
-
Improving Low Resource Code-switched ASR using Augmented Code-switched TTS12 Oct 2020 0 repositories listed
-
Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling12 Oct 2020 0 repositories listed
-
WER we are and WER we think we are7 Oct 2020 0 repositories listed
-
A Study on Lip Localization Techniques used for Lip reading from a Video28 Sep 2020 0 repositories listed
-
FluentNet: End-to-End Detection of Speech Disfluency with Deep Learning23 Sep 2020 0 repositories listed
-
EasyASR: A Distributed Machine Learning Platform for End-to-end Automatic Speech Recognition14 Sep 2020 0 repositories listed
-
Multi-modal embeddings using multi-task learning for emotion recognition10 Sep 2020 0 repositories listed
-
Unmanned Aerial Vehicle Control Through Domain-based Automatic Speech Recognition9 Sep 2020 0 repositories listed
-
Robust Spoken Language Understanding with RL-based Value Error Recovery7 Sep 2020 0 repositories listed
-
Silent Speech Interfaces for Speech Restoration: A Review4 Sep 2020 0 repositories listed
-
Voice Conversion by Cascading Automatic Speech Recognition and Text-to-Speech Synthesis with Prosody Transfer3 Sep 2020 0 repositories listed
-
Convolutional Speech Recognition with Pitch and Voice Quality Features2 Sep 2020 0 repositories listed
-
Multi-view Attention-based Speech Enhancement Model for Noise-robust Automatic Speech Recognition1 Sep 2020 0 repositories listed
-
Aphasic Speech Recognition using a Mixture of Speech Intelligibility Experts25 Aug 2020 0 repositories listed
-
Learned Transferable Architectures Can Surpass Hand-Designed Architectures for Large Scale Speech Recognition25 Aug 2020 0 repositories listed
-
Improving Tail Performance of a Deliberation E2E ASR Model Using a Large Text Corpus24 Aug 2020 0 repositories listed
-
Cross-Utterance Language Models with Acoustic Error Sampling19 Aug 2020 0 repositories listed
-
Speech To Semantics: Improve ASR and NLU Jointly via All-Neural Interfaces14 Aug 2020 0 repositories listed
-
Conv-Transformer Transducer: Low Latency, Low Frame Rate, Streamable End-to-End Speech Recognition13 Aug 2020 0 repositories listed
-
Large-scale Transfer Learning for Low-resource Spoken Language Understanding13 Aug 2020 0 repositories listed
-
13 Aug 2020 0 repositories listed
-
Online Automatic Speech Recognition with Listen, Attend and Spell Model12 Aug 2020 0 repositories listed
-
Transfer Learning Approaches for Streaming End-to-End Speech Recognition System12 Aug 2020 0 repositories listed
-
Transformer with Bidirectional Decoder for Speech Recognition11 Aug 2020 0 repositories listed
-
Subword Regularization: An Analysis of Scalability and Generalization for End-to-End Automatic Speech Recognition10 Aug 2020 0 repositories listed
-
LRSpeech: Extremely Low-Resource Speech Synthesis and Recognition9 Aug 2020 0 repositories listed
-
Deep Learning Based Dereverberation of Temporal Envelopesfor Robust Speech Recognition7 Aug 2020 0 repositories listed
-
Investigation of Speaker-adaptation methods in Transformer based ASR7 Aug 2020 0 repositories listed
-
A Transfer Learning Method for Speech Emotion Recognition from Automatic Speech Recognition6 Aug 2020 0 repositories listed
-
Iterative Compression of End-to-End ASR Model using AutoML6 Aug 2020 0 repositories listed
-
Shouted Speech Compensation for Speaker Verification Robust to Vocal Effort Conditions6 Aug 2020 0 repositories listed
-
Unsupervised Cross-Domain Singing Voice Conversion6 Aug 2020 0 repositories listed
-
"This is Houston. Say again, please". The Behavox system for the Apollo-11 Fearless Steps Challenge (phase II)4 Aug 2020 0 repositories listed
-
Weakly Supervised Construction of ASR Systems with Massive Video Data4 Aug 2020 0 repositories listed
-
Modular End-to-end Automatic Speech Recognition Framework for Acoustic-to-word Model31 Jul 2020 0 repositories listed
-
Utterance-Wise Meeting Transcription System Using Asynchronous Distributed Microphones31 Jul 2020 0 repositories listed
-
Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability30 Jul 2020 0 repositories listed
-
Exploiting Cross-Lingual Knowledge in Unsupervised Acoustic Modeling for Low-Resource Languages29 Jul 2020 0 repositories listed
-
Neural Kalman Filtering for Speech Enhancement28 Jul 2020 0 repositories listed
-
Effects of Language Relatedness for Cross-lingual Transfer Learning in Character-Based Language Models22 Jul 2020 0 repositories listed
-
Audio Adversarial Examples for Robust Hybrid CTC/Attention Speech Recognition21 Jul 2020 0 repositories listed
-
Neural Architecture Search For LF-MMI Trained Time Delay Neural Networks17 Jul 2020 0 repositories listed
-
Towards an Automated SOAP Note: Classifying Utterances from Medical Conversations17 Jul 2020 0 repositories listed
-
SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems13 Jul 2020 0 repositories listed
-
Massively Multilingual ASR: 50 Languages, 1 Model, 1 Billion Parameters6 Jul 2020 0 repositories listed
-
Deep Graph Random Process for Relational-Thinking-Based Speech Recognition4 Jul 2020 0 repositories listed
-
Robust Prediction of Punctuation and Truecasing for Medical ASR4 Jul 2020 0 repositories listed
-
CUNI Neural ASR with Phoneme-Level Intermediate Step for\textasciitildeNon-Native\textasciitildeSLT at IWSLT 20201 Jul 2020 0 repositories listed
-
1 Jul 2020 0 repositories listed
-
How Accents Confound: Probing for Accent Information in End-to-End Speech Recognition Systems1 Jul 2020 0 repositories listed
-
Investigating the effect of auxiliary objectives for the automated grading of learner English speech transcriptions1 Jul 2020 0 repositories listed
-
Large Vocabulary Read Speech Corpora for Four Ethiopian Languages: Amharic, Tigrigna, Oromo, and Wolaytta1 Jul 2020 0 repositories listed
-
Multimodal and Multiresolution Speech Recognition with Transformers1 Jul 2020 0 repositories listed
-
Robust Neural Machine Translation with ASR Errors1 Jul 2020 0 repositories listed
-
SimulSpeech: End-to-End Simultaneous Speech to Text Translation1 Jul 2020 0 repositories listed
-
Start-Before-End and End-to-End: Neural Speech Translation by AppTek and RWTH Aachen University1 Jul 2020 0 repositories listed
-
The AFRL IWSLT 2020 Systems: Work-From-Home Edition1 Jul 2020 0 repositories listed
-
Tigrinya Automatic Speech recognition with Morpheme based recognition units1 Jul 2020 0 repositories listed
-
Towards end-2-end learning for predicting behavior codes from spoken utterances in psychotherapy conversations1 Jul 2020 0 repositories listed
-
Towards Stream Translation: Adaptive Computation Time for Simultaneous Machine Translation1 Jul 2020 0 repositories listed
-
Towards Understanding ASR Error Correction for Medical Conversations1 Jul 2020 0 repositories listed
-
Multi-view Frequency LSTM: An Efficient Frontend for Automatic Speech Recognition30 Jun 2020 0 repositories listed
-
Neural Machine Translation for Multilingual Grapheme-to-Phoneme Conversion25 Jun 2020 0 repositories listed
-
Streaming Transformer ASR with Blockwise Synchronous Inference25 Jun 2020 0 repositories listed
-
Boosting Active Learning for Speech Recognition with Noisy Pseudo-labeled Samples19 Jun 2020 0 repositories listed
-
Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers19 Jun 2020 0 repositories listed
-
End-to-End Code Switching Language Models for Automatic Speech Recognition16 Jun 2020 0 repositories listed
-
Quantization of Acoustic Model Parameters in Automatic Speech Recognition Framework16 Jun 2020 0 repositories listed
-
Towards Automated Assessment of Stuttering and Stuttering Therapy16 Jun 2020 0 repositories listed
-
Exploration of End-to-End ASR for OpenSTT -- Russian Open Speech-to-Text Dataset15 Jun 2020 0 repositories listed
-
The JHU Multi-Microphone Multi-Speaker ASR System for the CHiME-6 Challenge14 Jun 2020 0 repositories listed
-
Data Augmentation for Training Dialog Models Robust to Speech Recognition Errors10 Jun 2020 0 repositories listed
-
Improving Cross-Lingual Transfer Learning for End-to-End Speech Recognition with Speech Translation9 Jun 2020 0 repositories listed
-
Learning not to Discriminate: Task Agnostic Learning for Improving Monolingual and Code-switched Speech Recognition9 Jun 2020 0 repositories listed
-
On the Effectiveness of Neural Text Generation based Data Augmentation for Recognition of Morphologically Rich Speech9 Jun 2020 0 repositories listed
-
Contextual RNN-T For Open Domain ASR4 Jun 2020 0 repositories listed
-
Multi-talker ASR for an unknown number of sources: Joint training of source counting, separation and ASR4 Jun 2020 0 repositories listed
-
Transfer Learning for British Sign Language Modelling3 Jun 2020 0 repositories listed
-
Analyzing the Quality and Stability of a Streaming End-to-End On-Device Speech Recognizer2 Jun 2020 0 repositories listed
-
Detecting Audio Attacks on ASR Systems with Dropout Uncertainty2 Jun 2020 0 repositories listed
-
Dilated U-net based approach for multichannel speech enhancement from First-Order Ambisonics recordings2 Jun 2020 0 repositories listed
-
An Effective Contextual Language Modeling Framework for Speech Summarization with Augmented Features1 Jun 2020 0 repositories listed
-
Analyse de l'effet de la réverbération sur la reconnaissance automatique de la parole (Analyzing how reverberation affects Automatic Speech Recognition)1 Jun 2020 0 repositories listed
-
Constrained Variational Autoencoder for improving EEG based Speech Recognition Systems1 Jun 2020 0 repositories listed
-
Introduction d'informations sémantiques dans un système de reconnaissance de la parole (Despite spectacular advances in recent years, the Automatic Speech Recognition (ASR) systems still make mistakes, especially in noisy environments)1 Jun 2020 0 repositories listed
-
Learning to Recognize Code-switched Speech Without Forgetting Monolingual Speech Recognition1 Jun 2020 0 repositories listed
-
Reconnaissance automatique de la parole : génération des prononciations non natives pour l'enrichissement du lexique (In this study we propose a method for lexicon adaptation in order to improve the automatic speech recognition (ASR) of non-native speakers)1 Jun 2020 0 repositories listed
-
Sur l'utilisation de la reconnaissance automatique de la parole pour l'aide au diagnostic différentiel entre la maladie de Parkinson et l'AMS (On using automatic speech recognition for the differential diagnosis of Parkinson's Disease and MSA This article presents a study regarding the contribution of automatic speech processing in the differential diagnosis between Parkinson's disease and MSA (Multi-System Atrophies))1 Jun 2020 0 repositories listed
-
Dynamic Masking for Improved Stability in Spoken Language Translation30 May 2020 0 repositories listed
-
An Audio-enriched BERT-based Framework for Spoken Multiple-choice Question Answering25 May 2020 0 repositories listed
-
An End-to-End Mispronunciation Detection System for L2 English Speech Leveraging Novel Anti-Phone Modeling25 May 2020 0 repositories listed
-
25 May 2020 0 repositories listed
-
Large scale evaluation of importance maps in automatic speech recognition21 May 2020 0 repositories listed
-
A Comparison of Label-Synchronous and Frame-Synchronous End-to-End Models for Speech Recognition20 May 2020 0 repositories listed
-
Early Stage LM Integration Using Local and Global Log-Linear Combination20 May 2020 0 repositories listed
-
Investigation of Large-Margin Softmax in Neural Language Modeling20 May 2020 0 repositories listed