Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 17
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 17 of 32: papers 1,601 to 1,700 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Improving Fast-slow Encoder based Transducer with Streaming Deliberation15 Dec 2022 0 repositories listed
-
Disentangling Prosody Representations with Unsupervised Speech Reconstruction14 Dec 2022 0 repositories listed
-
Speech and Natural Language Processing Technologies for Pseudo-Pilot Simulator14 Dec 2022 0 repositories listed
-
End-to-End Speech Translation of Arabic to English Broadcast News11 Dec 2022 0 repositories listed
-
Improved Self-Supervised Multilingual Speech Representation Learning Combined with Auxiliary Language Information7 Dec 2022 0 repositories listed
-
Improved Speech Pre-Training with Supervision-Enhanced Acoustic Unit7 Dec 2022 0 repositories listed
-
Lattice-Free Sequence Discriminative Training for Phoneme-Based Neural Transducers7 Dec 2022 0 repositories listed
-
Progressive Multi-Scale Self-Supervised Learning for Speech Recognition7 Dec 2022 0 repositories listed
-
Unsupervised Fine-Tuning Data Selection for ASR Using Self-Supervised Speech Models3 Dec 2022 0 repositories listed
-
Continual Learning for On-Device Speech Recognition using Disentangled Conformers2 Dec 2022 0 repositories listed
-
Preliminary Study on SSCF-derived Polar Coordinate for ASR30 Nov 2022 0 repositories listed
-
Better Transcription of UK Supreme Court Hearings29 Nov 2022 0 repositories listed
-
Evaluating and reducing the distance between synthetic and real speech distributions29 Nov 2022 0 repositories listed
-
Neural Transducer Training: Reduced Memory Consumption with Sample-wise Computation29 Nov 2022 0 repositories listed
-
Inter-KD: Intermediate Knowledge Distillation for CTC-Based Automatic Speech Recognition28 Nov 2022 0 repositories listed
-
Bidirectional Representations for Low Resource Spoken Language Understanding24 Nov 2022 0 repositories listed
-
Multitask Learning for Low Resource Spoken Language Understanding24 Nov 2022 0 repositories listed
-
Device Directedness with Contextual Cues for Spoken Dialog Systems23 Nov 2022 0 repositories listed
-
22 Nov 2022 0 repositories listed
-
Complex-Valued Time-Frequency Self-Attention for Speech Dereverberation22 Nov 2022 0 repositories listed
-
SSCFormer: Push the Limit of Chunk-wise Conformer for Streaming ASR Using Sequentially Sampled Chunks and Chunked Causal Convolution21 Nov 2022 0 repositories listed
-
SpeechNet: Weakly Supervised, End-to-End Speech Recognition at Industrial Scale21 Nov 2022 0 repositories listed
-
Hey ASR System! Why Aren't You More Inclusive? Automatic Speech Recognition Systems' Bias and Proposed Bias Mitigation Techniques. A Literature Review17 Nov 2022 0 repositories listed
-
LongFNT: Long-form Speech Recognition with Factorized Neural Transducer17 Nov 2022 0 repositories listed
-
Unsupervised Model-based speaker adaptation of end-to-end lattice-free MMI model for speech recognition17 Nov 2022 0 repositories listed
-
Improving Speech Emotion Recognition with Unsupervised Speaking Style Transfer16 Nov 2022 0 repositories listed
-
On using the UA-Speech and TORGO databases to validate automatic dysarthric speech classification approaches16 Nov 2022 0 repositories listed
-
Introducing Semantics into Speech Encoders15 Nov 2022 0 repositories listed
-
Align, Write, Re-order: Explainable End-to-End Speech Translation via Operation Sequence Generation11 Nov 2022 0 repositories listed
-
Handling Trade-Offs in Speech Separation with Sparsely-Gated Mixture of Experts11 Nov 2022 0 repositories listed
-
A Study on the Integration of Pre-trained SSL, ASR, LM and SLU Models for Spoken Language Understanding10 Nov 2022 0 repositories listed
-
Adaptive Multi-Corpora Language Model Training for Speech Recognition9 Nov 2022 0 repositories listed
-
Improving Noisy Student Training on Non-target Domain Data for Automatic Speech Recognition9 Nov 2022 0 repositories listed
-
End-to-End Evaluation of a Spoken Dialogue System for Learning Basic Mathematics7 Nov 2022 0 repositories listed
-
Streaming, fast and accurate on-device Inverse Text Normalization for Automatic Speech Recognition7 Nov 2022 0 repositories listed
-
Bridging Speech and Textual Pre-trained Models with Unsupervised ASR6 Nov 2022 0 repositories listed
-
Evaluation of Automated Speech Recognition Systems for Conversational Speech: A Linguistic Perspective5 Nov 2022 0 repositories listed
-
LAMASSU: Streaming Language-Agnostic Multilingual Speech Recognition and Translation Using Neural Transducers5 Nov 2022 0 repositories listed
-
Biased Self-supervised learning for ASR4 Nov 2022 0 repositories listed
-
Resource-Efficient Transfer Learning From Speech Foundation Model Using Hierarchical Feature Fusion4 Nov 2022 0 repositories listed
-
Stutter-TTS: Controlled Synthesis and Improved Recognition of Stuttered Speech4 Nov 2022 0 repositories listed
-
H_eval: A new hybrid evaluation metric for automatic speech recognition tasks3 Nov 2022 0 repositories listed
-
Leveraging Domain Features for Detecting Adversarial Attacks Against Deep Speech Recognition in Noise3 Nov 2022 0 repositories listed
-
Phonetic-assisted Multi-Target Units Modeling for Improving Conformer-Transducer ASR system3 Nov 2022 0 repositories listed
-
Probing Statistical Representations For End-To-End ASR3 Nov 2022 0 repositories listed
-
Streaming Audio-Visual Speech Recognition with Alignment Regularization3 Nov 2022 0 repositories listed
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder2 Nov 2022 0 repositories listed
-
Monolingual Recognizers Fusion for Code-switching Speech Recognition2 Nov 2022 0 repositories listed
-
More Speaking or More Speakers?2 Nov 2022 0 repositories listed
-
Towards Zero-Shot Code-Switched Speech Recognition2 Nov 2022 0 repositories listed
-
A Comparative Study on Multichannel Speaker-Attributed Automatic Speech Recognition in Multi-party Meetings1 Nov 2022 0 repositories listed
-
A Preliminary Study on Automated Speaking Assessment of English as a Second Language (ESL) Students1 Nov 2022 0 repositories listed
-
Adapting self-supervised models to multi-talker speech recognition using speaker embeddings1 Nov 2022 0 repositories listed
-
Mandarin-English Code-Switching Speech Recognition System for Specific Domain1 Nov 2022 0 repositories listed
-
Unified End-to-End Speech Recognition and Endpointing for Fast and Efficient Speech Systems1 Nov 2022 0 repositories listed
-
An analysis of degenerating speech due to progressive dysarthria on ASR performance31 Oct 2022 0 repositories listed
-
Audio-Visual Speech Enhancement and Separation by Utilizing Multi-Modal Self-Supervised Embeddings31 Oct 2022 0 repositories listed
-
DiaCorrect: End-to-end error correction for speaker diarization31 Oct 2022 0 repositories listed
-
FusionFormer: Fusing Operations in Transformer for Efficient Streaming Speech Recognition31 Oct 2022 0 repositories listed
-
Structured State Space Decoder for Speech Recognition and Synthesis31 Oct 2022 0 repositories listed
-
DuDe: Dual-Decoder Multilingual ASR for Indian Languages using Common Label Set30 Oct 2022 0 repositories listed
-
Phonemic Representation and Transcription for Speech to Text Applications for Under-resourced Indigenous African Languages: The Case of Kiswahili29 Oct 2022 0 repositories listed
-
Filter and evolve: progressive pseudo label refining for semi-supervised automatic speech recognition28 Oct 2022 0 repositories listed
-
Random Utterance Concatenation Based Data Augmentation for Improving Short-video Speech Recognition28 Oct 2022 0 repositories listed
-
Contextual-Utterance Training for Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Make More of Your Data: Minimal Effort Data Augmentation for Automatic Speech Recognition and Translation27 Oct 2022 0 repositories listed
-
SAN: a robust end-to-end ASR model architecture27 Oct 2022 0 repositories listed
-
Simulating realistic speech overlaps improves multi-talker ASR27 Oct 2022 0 repositories listed
-
Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance27 Oct 2022 0 repositories listed
-
TRScore: A Novel GPT-based Readability Scorer for ASR Segmentation and Punctuation model evaluation and selection27 Oct 2022 0 repositories listed
-
V-Cloak: Intelligibility-, Naturalness- & Timbre-Preserving Real-Time Voice Anonymization27 Oct 2022 0 repositories listed
-
Virtuoso: Massive Multilingual Speech-Text Joint Semi-Supervised Learning for Text-To-Speech27 Oct 2022 0 repositories listed
-
Weight Averaging: A Simple Yet Effective Method to Overcome Catastrophic Forgetting in Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Efficient Utilization of Large Pre-Trained Models for Low Resource ASR26 Oct 2022 0 repositories listed
-
End-to-End Speech to Intent Prediction to improve E-commerce Customer Support Voicebot in Hindi and English26 Oct 2022 0 repositories listed
-
Four-in-One: A Joint Approach to Inverse Text Normalization, Punctuation, Capitalization, and Disfluency for Automatic Speech Recognition26 Oct 2022 0 repositories listed
-
Smart Speech Segmentation using Acousto-Linguistic Features with look-ahead26 Oct 2022 0 repositories listed
-
UFO2: A unified pre-training framework for online and offline speech recognition26 Oct 2022 0 repositories listed
-
Linguistic-Enhanced Transformer with CTC Embedding for Speech Recognition25 Oct 2022 0 repositories listed
-
Investigating self-supervised, weakly supervised and fully supervised training approaches for multi-domain automatic speech recognition: a study on Bangladeshi Bangla24 Oct 2022 0 repositories listed
-
Time-Domain Speech Enhancement for Robust Automatic Speech Recognition24 Oct 2022 0 repositories listed
-
Guided contrastive self-supervised pre-training for automatic speech recognition22 Oct 2022 0 repositories listed
-
Can Visual Context Improve Automatic Speech Recognition for an Embodied Agent?21 Oct 2022 0 repositories listed
-
Optimizing Bilingual Neural Transducer with Synthetic Code-switching Text Generation21 Oct 2022 0 repositories listed
-
Improving Semi-supervised End-to-end Automatic Speech Recognition using CycleGAN and Inter-domain Losses20 Oct 2022 0 repositories listed
-
19 Oct 2022 0 repositories listed
-
Continuous Pseudo-Labeling from the Start17 Oct 2022 0 repositories listed
-
Language-agnostic Code-Switching in Sequence-To-Sequence Speech Recognition17 Oct 2022 0 repositories listed
-
Sub-8-bit quantization for on-device speech recognition: a regularization-free approach17 Oct 2022 0 repositories listed
-
Learning to Jointly Transcribe and Subtitle for End-to-End Spontaneous Speech Recognition14 Oct 2022 0 repositories listed
-
LeVoice ASR Systems for the ISCSLP 2022 Intelligent Cockpit Speech Recognition Challenge14 Oct 2022 0 repositories listed
-
Experiments on Turkish ASR with Self-Supervised Speech Representation Learning13 Oct 2022 0 repositories listed
-
Summary on the ISCSLP 2022 Chinese-English Code-Switching ASR Challenge12 Oct 2022 0 repositories listed
-
An Experimental Study on Private Aggregation of Teacher Ensemble Learning for End-to-End Speech Recognition11 Oct 2022 0 repositories listed
-
Automatic Speech Recognition of Low-Resource Languages Based on Chukchi11 Oct 2022 0 repositories listed
-
Comparison of Soft and Hard Target RNN-T Distillation for Large-scale ASR11 Oct 2022 0 repositories listed
-
CTC Alignments Improve Autoregressive Translation11 Oct 2022 0 repositories listed
-
Scaling Up Deliberation for Multilingual ASR11 Oct 2022 0 repositories listed
-
Streaming Punctuation for Long-form Dictation with Transformers11 Oct 2022 0 repositories listed