Browse State-of-the-Art › Speech Recognition › Papers, page 30
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 30 of 65: papers 2,901 to 3,000 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Contextual-Utterance Training for Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Explicit Intensity Control for Accented Text-to-speech27 Oct 2022 0 repositories listed
-
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Make More of Your Data: Minimal Effort Data Augmentation for Automatic Speech Recognition and Translation27 Oct 2022 0 repositories listed
-
SAN: a robust end-to-end ASR model architecture27 Oct 2022 0 repositories listed
-
Simulating realistic speech overlaps improves multi-talker ASR27 Oct 2022 0 repositories listed
-
Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance27 Oct 2022 0 repositories listed
-
Training Autoregressive Speech Recognition Models with Limited in-domain Supervision27 Oct 2022 0 repositories listed
-
TRScore: A Novel GPT-based Readability Scorer for ASR Segmentation and Punctuation model evaluation and selection27 Oct 2022 0 repositories listed
-
V-Cloak: Intelligibility-, Naturalness- & Timbre-Preserving Real-Time Voice Anonymization27 Oct 2022 0 repositories listed
-
Virtuoso: Massive Multilingual Speech-Text Joint Semi-Supervised Learning for Text-To-Speech27 Oct 2022 0 repositories listed
-
Weight Averaging: A Simple Yet Effective Method to Overcome Catastrophic Forgetting in Automatic Speech Recognition27 Oct 2022 0 repositories listed
-
Efficient Utilization of Large Pre-Trained Models for Low Resource ASR26 Oct 2022 0 repositories listed
-
End-to-End Speech to Intent Prediction to improve E-commerce Customer Support Voicebot in Hindi and English26 Oct 2022 0 repositories listed
-
Four-in-One: A Joint Approach to Inverse Text Normalization, Punctuation, Capitalization, and Disfluency for Automatic Speech Recognition26 Oct 2022 0 repositories listed
-
Improving Speech-to-Speech Translation Through Unlabeled Text26 Oct 2022 0 repositories listed
-
Pronunciation Generation for Foreign Language Words in Intra-Sentential Code-Switching Speech Recognition26 Oct 2022 0 repositories listed
-
Smart Speech Segmentation using Acousto-Linguistic Features with look-ahead26 Oct 2022 0 repositories listed
-
UFO2: A unified pre-training framework for online and offline speech recognition26 Oct 2022 0 repositories listed
-
Linguistic-Enhanced Transformer with CTC Embedding for Speech Recognition25 Oct 2022 0 repositories listed
-
Investigating self-supervised, weakly supervised and fully supervised training approaches for multi-domain automatic speech recognition: a study on Bangladeshi Bangla24 Oct 2022 0 repositories listed
-
Time-Domain Speech Enhancement for Robust Automatic Speech Recognition24 Oct 2022 0 repositories listed
-
Guided contrastive self-supervised pre-training for automatic speech recognition22 Oct 2022 0 repositories listed
-
Can Visual Context Improve Automatic Speech Recognition for an Embodied Agent?21 Oct 2022 0 repositories listed
-
Deep LSTM Spoken Term Detection using Wav2Vec 2.0 Recognizer21 Oct 2022 0 repositories listed
-
Optimizing Bilingual Neural Transducer with Synthetic Code-switching Text Generation21 Oct 2022 0 repositories listed
-
Anchored Speech Recognition with Neural Transducers20 Oct 2022 0 repositories listed
-
Improving Semi-supervised End-to-end Automatic Speech Recognition using CycleGAN and Inter-domain Losses20 Oct 2022 0 repositories listed
-
19 Oct 2022 0 repositories listed
-
Speaker- and Age-Invariant Training for Child Acoustic Modeling Using Adversarial Multi-Task Learning19 Oct 2022 0 repositories listed
-
Tourist Guidance Robot Based on HyperCLOVA19 Oct 2022 0 repositories listed
-
Layer-wise Relevance Propagation for Echo State Networks applied to Earth System Variability18 Oct 2022 0 repositories listed
-
Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR18 Oct 2022 0 repositories listed
-
Towards Personalization of CTC Speech Recognition Models with Contextual Adapters and Adaptive Boosting18 Oct 2022 0 repositories listed
-
Simple and Effective Unsupervised Speech Translation18 Oct 2022 0 repositories listed
-
A Treatise On FST Lattice Based MMI Training17 Oct 2022 0 repositories listed
-
Continuous Pseudo-Labeling from the Start17 Oct 2022 0 repositories listed
-
Language-agnostic Code-Switching in Sequence-To-Sequence Speech Recognition17 Oct 2022 0 repositories listed
-
Sub-8-bit quantization for on-device speech recognition: a regularization-free approach17 Oct 2022 0 repositories listed
-
Acoustic-aware Non-autoregressive Spell Correction with Mask Sample Decoding16 Oct 2022 0 repositories listed
-
Learning Invariant Representation and Risk Minimized for Unsupervised Accent Domain Adaptation15 Oct 2022 0 repositories listed
-
Learning to Jointly Transcribe and Subtitle for End-to-End Spontaneous Speech Recognition14 Oct 2022 0 repositories listed
-
LeVoice ASR Systems for the ISCSLP 2022 Intelligent Cockpit Speech Recognition Challenge14 Oct 2022 0 repositories listed
-
Experiments on Turkish ASR with Self-Supervised Speech Representation Learning13 Oct 2022 0 repositories listed
-
Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models13 Oct 2022 0 repositories listed
-
An Ensemble Teacher-Student Learning Approach with Poisson Sub-sampling to Differential Privacy Preserving Speech Recognition12 Oct 2022 0 repositories listed
-
Summary on the ISCSLP 2022 Chinese-English Code-Switching ASR Challenge12 Oct 2022 0 repositories listed
-
An Experimental Study on Private Aggregation of Teacher Ensemble Learning for End-to-End Speech Recognition11 Oct 2022 0 repositories listed
-
Automatic Speech Recognition of Low-Resource Languages Based on Chukchi11 Oct 2022 0 repositories listed
-
Comparison of Soft and Hard Target RNN-T Distillation for Large-scale ASR11 Oct 2022 0 repositories listed
-
CTC Alignments Improve Autoregressive Translation11 Oct 2022 0 repositories listed
-
Inner speech recognition through electroencephalographic signals11 Oct 2022 0 repositories listed
-
Scaling Up Deliberation for Multilingual ASR11 Oct 2022 0 repositories listed
-
Streaming Punctuation for Long-form Dictation with Transformers11 Oct 2022 0 repositories listed
-
Cloud-based Automatic Speech Recognition Systems for Southeast Asian Languages7 Oct 2022 0 repositories listed
-
Pronunciation Modeling of Foreign Words for Mandarin ASR by Considering the Effect of Language Transfer7 Oct 2022 0 repositories listed
-
Damage Control During Domain Adaptation for Transducer Based Automatic Speech Recognition6 Oct 2022 0 repositories listed
-
Synthetic Dataset Generation for Privacy-Preserving Machine Learning6 Oct 2022 0 repositories listed
-
Code-Switching without Switching: Language Agnostic End-to-End Speech Translation4 Oct 2022 0 repositories listed
-
Efficient acoustic feature transformation in mismatched environments using a Guided-GAN3 Oct 2022 0 repositories listed
-
Learning with Limited Samples -- Meta-Learning and Applications to Communication Systems3 Oct 2022 0 repositories listed
-
A Comparison of Transformer, Convolutional, and Recurrent Neural Networks on Phoneme Recognition1 Oct 2022 0 repositories listed
-
Can We Train a Language Model Inside an End-to-End ASR Model? - Investigating Effective Implicit Language Modeling1 Oct 2022 0 repositories listed
-
Fashioning Local Designs from Generic Speech Technologies in an Australian Aboriginal Community1 Oct 2022 0 repositories listed
-
Improving Code-switched ASR with Linguistic Information1 Oct 2022 0 repositories listed
-
Investigating the Impact of ASR Errors on Spoken Implicit Discourse Relation Recognition1 Oct 2022 0 repositories listed
-
Keyphrase Prediction from Video Transcripts: New Dataset and Directions1 Oct 2022 0 repositories listed
-
Language-specific Effects on Automatic Speech Recognition Errors for World Englishes1 Oct 2022 0 repositories listed
-
面向 Transformer 模型的蒙古语语音识别词特征编码方法(Researching of the Mongolian word encoding method based on Transformer Mongolian speech recognition)1 Oct 2022 0 repositories listed
-
Multi-stage Progressive Compression of Conformer Transducer for On-device Speech Recognition1 Oct 2022 0 repositories listed
-
融合外部语言知识的流式越南语语音识别(Streaming Vietnamese Speech Recognition Based on Fusing External Vietnamese Language Knowledge)1 Oct 2022 0 repositories listed
-
Zero-shot Disfluency Detection for Indian Languages1 Oct 2022 0 repositories listed
-
Adaptive Sparse and Monotonic Attention for Transformer-based Automatic Speech Recognition30 Sep 2022 0 repositories listed
-
Blind Signal Dereverberation for Machine Speech Recognition30 Sep 2022 0 repositories listed
-
ConvRNN-T: Convolutional Augmented Recurrent Neural Network Transducers for Streaming Speech Recognition29 Sep 2022 0 repositories listed
-
An Effective, Performant Named Entity Recognition System for Noisy Business Telephone Conversation Transcripts27 Sep 2022 0 repositories listed
-
Unsupervised domain adaptation for speech recognition with unsupervised error correction24 Sep 2022 0 repositories listed
-
Assessing ASR Model Quality on Disordered Speech using BERTScore21 Sep 2022 0 repositories listed
-
Parameter-Efficient Conformers via Sharing Sparsely-Gated Experts for End-to-End Speech Recognition17 Sep 2022 0 repositories listed
-
Distribution Aware Metrics for Conditional Natural Language Generation15 Sep 2022 0 repositories listed
-
Non-Parallel Voice Conversion for ASR Augmentation15 Sep 2022 0 repositories listed
-
A Universally-Deployable ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement, and Voice Separation14 Sep 2022 0 repositories listed
-
ESSumm: Extractive Speech Summarization from Untranscribed Meeting14 Sep 2022 0 repositories listed
-
Federated Pruning: Improving Neural Network Efficiency with Federated Learning14 Sep 2022 0 repositories listed
-
Analysis of Self-Attention Head Diversity for Conformer-based Automatic Speech Recognition13 Sep 2022 0 repositories listed
-
Bangla-Wave: Improving Bangla Automatic Speech Recognition Utilizing N-gram Language Models13 Sep 2022 0 repositories listed
-
Learning ASR pathways: A sparse multilingual ASR model13 Sep 2022 0 repositories listed
-
Streaming End-to-End Multilingual Speech Recognition with Joint Language Identification13 Sep 2022 0 repositories listed
-
VarArray Meets t-SOT: Advancing the State of the Art of Streaming Distant Conversational Speech Recognition12 Sep 2022 0 repositories listed
-
Applying wav2vec2 for Speech Recognition on Bengali Common Voices Dataset11 Sep 2022 0 repositories listed
-
Lexicon and Attention based Handwritten Text Recognition System11 Sep 2022 0 repositories listed
-
Conversion of Acoustic Signal (Speech) Into Text By Digital Filter using Natural Language Processing9 Sep 2022 0 repositories listed
-
Streaming Target-Speaker ASR with Neural Transducer9 Sep 2022 0 repositories listed
-
Multilingual Transformer Language Model for Speech Recognition in Low-resource Languages8 Sep 2022 0 repositories listed
-
Modeling Dependent Structure for Utterances in ASR Evaluation7 Sep 2022 0 repositories listed
-
Plant Species Classification Using Transfer Learning by Pretrained Classifier VGG-197 Sep 2022 0 repositories listed
-
Distilling the Knowledge of BERT for CTC-based ASR5 Sep 2022 0 repositories listed
-
Towards Deep Learning-aided Wireless Channel Estimation and Channel State Information Feedback for 6G5 Sep 2022 0 repositories listed
-
A Review of Sparse Expert Models in Deep Learning4 Sep 2022 0 repositories listed
-
Universal Fourier Attack for Time Series2 Sep 2022 0 repositories listed