Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 19
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 19 of 32: papers 1,801 to 1,900 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Context-based out-of-vocabulary word recovery for ASR systems in Indian languages9 Jun 2022 0 repositories listed
-
Face-Dubbing++: Lip-Synchronous, Voice Preserving Translation of Videos9 Jun 2022 0 repositories listed
-
Joint Encoder-Decoder Self-Supervised Pre-training for ASR9 Jun 2022 0 repositories listed
-
LegoNN: Building Modular Encoder-Decoder Models7 Jun 2022 0 repositories listed
-
FedNST: Federated Noisy Student Training for Automatic Speech Recognition6 Jun 2022 0 repositories listed
-
Pronunciation Dictionary-Free Multilingual Speech Synthesis by Combining Unsupervised and Supervised Phonetic Representations2 Jun 2022 0 repositories listed
-
A Semi-Automated Live Interlingual Communication Workflow Featuring Intralingual Respeaking: Evaluation and Benchmarking1 Jun 2022 0 repositories listed
-
Automatic Speech Recognition for Irish: the ABAIR-ÉIST System1 Jun 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Jun 2022 0 repositories listed
-
Building Open-source Speech Technology for Low-resource Minority Languages with SáMi as an Example – Tools, Methods and Experiments1 Jun 2022 0 repositories listed
-
Conversational Speech Recognition Needs Data? Experiments with Austrian German1 Jun 2022 0 repositories listed
-
Developing Automatic Speech Recognition for Scottish Gaelic1 Jun 2022 0 repositories listed
-
Development of Automatic Speech Recognition for the Documentation of Cook Islands Māori1 Jun 2022 0 repositories listed
-
Evaluation of Off-the-shelf Speech Recognizers on Different Accents in a Dialogue Domain1 Jun 2022 0 repositories listed
-
Generating Synthetic Clinical Speech Data through Simulated ASR Deletion Error1 Jun 2022 0 repositories listed
-
Huqariq: A Multilingual Speech Corpus of Native Languages of Peru forSpeech Recognition1 Jun 2022 0 repositories listed
-
LiSTra Automatic Speech Translation: English to Lingala Case Study1 Jun 2022 0 repositories listed
-
Mesures linguistiques automatiques pour l’évaluation des systèmes de Reconnaissance Automatique de la Parole (Automated linguistic measures for automatic speech recognition systems’ evaluation)1 Jun 2022 0 repositories listed
-
Multilingual Transfer Learning for Children Automatic Speech Recognition1 Jun 2022 0 repositories listed
-
ParlamentParla: A Speech Corpus of Catalan Parliamentary Sessions1 Jun 2022 0 repositories listed
-
ParlaSpeech-HR - a Freely Available ASR Dataset for Croatian Bootstrapped from the ParlaMint Corpus1 Jun 2022 0 repositories listed
-
Post-Stroke Speech Transcription Challenge (Task B): Correctness Detection in Anomia Diagnosis with Imperfect Transcripts1 Jun 2022 0 repositories listed
-
Progress in Multilingual Speech Recognition for Low Resource Languages Kurmanji Kurdish, Cree and Inuktut1 Jun 2022 0 repositories listed
-
Samrómur Children: An Icelandic Speech Corpus1 Jun 2022 0 repositories listed
-
Snow Mountain: Dataset of Audio Recordings of The Bible in Low Resource Languages1 Jun 2022 0 repositories listed
-
Towards a Unified ASR System for the Armenian Standards1 Jun 2022 0 repositories listed
-
Towards an Open-Source Dutch Speech Recognition System for the Healthcare Domain1 Jun 2022 0 repositories listed
-
Adversarial synthesis based data-augmentation for code-switched spoken language identification30 May 2022 0 repositories listed
-
Adaptive Activation Network For Low Resource Multilingual Speech Recognition28 May 2022 0 repositories listed
-
Acoustic-to-articulatory Speech Inversion with Multi-task Learning27 May 2022 0 repositories listed
-
Punctuation Restoration in Spanish Customer Support Transcripts using Transfer Learning27 May 2022 0 repositories listed
-
Clinical Dialogue Transcription Error Correction using Seq2Seq Models26 May 2022 0 repositories listed
-
Contextual Adapters for Personalized Speech Recognition in Neural Transducers26 May 2022 0 repositories listed
-
Joint Training of Speech Enhancement and Self-supervised Model for Noise-robust ASR26 May 2022 0 repositories listed
-
An Investigation on Applying Acoustic Feature Conversion to ASR of Adult and Child Speech25 May 2022 0 repositories listed
-
Heterogeneous Reservoir Computing Models for Persian Speech Recognition25 May 2022 0 repositories listed
-
Improving CTC-based ASR Models with Gated Interlayer Collaboration25 May 2022 0 repositories listed
-
Investigating Lexical Replacements for Arabic-English Code-Switched Data Augmentation25 May 2022 0 repositories listed
-
On Building Spoken Language Understanding Systems for Low Resourced Languages25 May 2022 0 repositories listed
-
Multi-Level Modeling Units for End-to-End Mandarin Speech Recognition24 May 2022 0 repositories listed
-
Calibrate and Refine! A Novel and Agile Framework for ASR-error Robust Intent Detection23 May 2022 0 repositories listed
-
Self-Supervised Speech Representation Learning: A Review21 May 2022 0 repositories listed
-
Automatic Spoken Language Identification using a Time-Delay Neural Network19 May 2022 0 repositories listed
-
Insights on Neural Representations for End-to-End Speech Recognition19 May 2022 0 repositories listed
-
Deploying self-supervised learning in the wild for hybrid automatic speech recognition17 May 2022 0 repositories listed
-
Streaming Noise Context Aware Enhancement For Automatic Speech Recognition in Multi-Talker Environments17 May 2022 0 repositories listed
-
Improved Consistency Training for Semi-Supervised Sequence-to-Sequence ASR via Speech Chain Reconstruction and Self-Transcribing14 May 2022 0 repositories listed
-
Pretraining Approaches for Spoken Language Recognition: TalTech Submission to the OLR 2021 Challenge14 May 2022 0 repositories listed
-
Personalized Adversarial Data Augmentation for Dysarthric and Elderly Speech Recognition13 May 2022 0 repositories listed
-
Unified Modeling of Multi-Domain Multi-Device ASR Systems13 May 2022 0 repositories listed
-
A Closer Look at Audio-Visual Multi-Person Speech Recognition and Active Speaker Selection11 May 2022 0 repositories listed
-
End-to-End Multi-Person Audio/Visual Automatic Speech Recognition11 May 2022 0 repositories listed
-
Best of Both Worlds: Multi-task Audio-Visual Automatic Speech Recognition and Active Speaker Detection10 May 2022 0 repositories listed
-
Speaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition9 May 2022 0 repositories listed
-
A Conformer-based Waveform-domain Neural Acoustic Echo Canceller Optimized for ASR Accuracy6 May 2022 0 repositories listed
-
ON-TRAC Consortium Systems for the IWSLT 2022 Dialect and Low-resource Speech Translation Tasks4 May 2022 0 repositories listed
-
A Meeting Transcription System for an Ad-Hoc Acoustic Sensor Network2 May 2022 0 repositories listed
-
Bilingual End-to-End ASR with Byte-Level Subwords1 May 2022 0 repositories listed
-
Discourse on ASR Measurement: Introducing the ARPOCA Assessment Tool1 May 2022 0 repositories listed
-
Enhancing Documentation of Hupa with Automatic Speech Recognition1 May 2022 0 repositories listed
-
Findings of the Shared Task on Speech Recognition for Vulnerable Individuals in Tamil1 May 2022 0 repositories listed
-
Fine-tuning pre-trained models for Automatic Speech Recognition, experiments on a fieldwork corpus of Japhug (Trans-Himalayan family)1 May 2022 0 repositories listed
-
JHU IWSLT 2022 Dialect Speech Translation System Description1 May 2022 0 repositories listed
-
MTL-SLT: Multi-Task Learning for Spoken Language Tasks1 May 2022 0 repositories listed
-
NVIDIA NeMo Offline Speech Translation Systems for IWSLT 20221 May 2022 0 repositories listed
-
Phoneme transcription of endangered languages: an evaluation of recent ASR architectures in the single speaker scenario1 May 2022 0 repositories listed
-
SSNCSE_NLP@LT-EDI-ACL2022: Speech Recognition for Vulnerable Individuals in Tamil using pre-trained XLSR models1 May 2022 0 repositories listed
-
SUH_ASR@LT-EDI-ACL2022: Transformer based Approach for Speech Recognition for Vulnerable Individuals in Tamil1 May 2022 0 repositories listed
-
The HW-TSC’s Offline Speech Translation System for IWSLT 2022 Evaluation1 May 2022 0 repositories listed
-
The Xiaomi Text-to-Text Simultaneous Speech Translation System for IWSLT 20221 May 2022 0 repositories listed
-
Mask scalar prediction for improving robust automatic speech recognition26 Apr 2022 0 repositories listed
-
Cleanformer: A multichannel array configuration-invariant neural enhancement frontend for ASR in smart speakers25 Apr 2022 0 repositories listed
-
Improved far-field speech recognition using Joint Variational Autoencoder24 Apr 2022 0 repositories listed
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment22 Apr 2022 0 repositories listed
-
An Investigation of Monotonic Transducers for Large-Scale Automatic Speech Recognition19 Apr 2022 0 repositories listed
-
Blockwise Streaming Transformer for Spoken Language Understanding and Simultaneous Speech Translation19 Apr 2022 0 repositories listed
-
Disappeared Command: Spoofing Attack On Automatic Speech Recognition Systems with Sound Masking19 Apr 2022 0 repositories listed
-
Automated speech tools for helping communities process restricted-access corpora for language revival efforts15 Apr 2022 0 repositories listed
-
Lombard Effect for Bilingual Speakers in Cantonese and English: importance of spectro-temporal features14 Apr 2022 0 repositories listed
-
A Unified Cascaded Encoder ASR Model for Dynamic Model Sizes13 Apr 2022 0 repositories listed
-
Self-critical Sequence Training for Automatic Speech Recognition13 Apr 2022 0 repositories listed
-
Study of Indian English Pronunciation Variabilities relative to Received Pronunciation13 Apr 2022 0 repositories listed
-
ASR in German: A Detailed Error Analysis12 Apr 2022 0 repositories listed
-
Building an ASR Error Robust Spoken Virtual Patient System in a Highly Class-Imbalanced Scenario Without Speech Data11 Apr 2022 0 repositories listed
-
Auditory-Based Data Augmentation for End-to-End Automatic Speech Recognition8 Apr 2022 0 repositories listed
-
Defense against Adversarial Attacks on Hybrid Speech Recognition using Joint Adversarial Fine-tuning with Denoiser8 Apr 2022 0 repositories listed
-
Three-Module Modeling For End-to-End Spoken Language Understanding Using Pre-trained DNN-HMM-Based Acoustic-Phonetic Model7 Apr 2022 0 repositories listed
-
A Wav2vec2-Based Experimental Study on Self-Supervised Learning Methods to Improve Child Speech Recognition6 Apr 2022 0 repositories listed
-
Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation6 Apr 2022 0 repositories listed
-
A Complementary Joint Training Approach Using Unpaired Speech and Text for Low-Resource Automatic Speech Recognition5 Apr 2022 0 repositories listed
-
Audio-visual multi-channel speech separation, dereverberation and recognition5 Apr 2022 0 repositories listed
-
Hear No Evil: Towards Adversarial Robustness of Automatic Speech Recognition via Multi-Task Learning5 Apr 2022 0 repositories listed
-
Unsupervised Data Selection via Discrete Speech Representation for ASR5 Apr 2022 0 repositories listed
-
A Study of Gender Impact in Self-supervised Models for Speech-to-Text Systems4 Apr 2022 0 repositories listed
-
An Analysis of Semantically-Aligned Speech-Text Embeddings4 Apr 2022 0 repositories listed
-
Cross-lingual Self-Supervised Speech Representations for Improved Dysarthric Speech Recognition4 Apr 2022 0 repositories listed
-
Deliberation Model for On-Device Spoken Language Understanding4 Apr 2022 0 repositories listed
-
End-to-end model for named entity recognition from speech without paired training data2 Apr 2022 0 repositories listed
-
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation2 Apr 2022 0 repositories listed
-
End-to-End Integration of Speech Recognition, Speech Enhancement, and Self-Supervised Learning Representation1 Apr 2022 0 repositories listed