Browse State-of-the-Art › speech-recognition › Papers, page 25
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 25 of 58: papers 2,401 to 2,500 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Federated Self-Learning with Weak Supervision for Speech Recognition21 Jun 2023 0 repositories listed
-
Learning When to Trust Which Teacher for Weakly Supervised ASR21 Jun 2023 0 repositories listed
-
Mixture Encoder for Joint Speech Separation and Recognition21 Jun 2023 0 repositories listed
-
Strategies in Transfer Learning for Low-Resource Speech Synthesis: Phone Mapping, Features Input, and Source Language Selection21 Jun 2023 0 repositories listed
-
Multi-pass Training and Cross-information Fusion for Low-resource End-to-end Accented Speech Recognition20 Jun 2023 0 repositories listed
-
Distillation Strategies for Discriminative Speech Recognition Rescoring15 Jun 2023 0 repositories listed
-
Lexical Speaker Error Correction: Leveraging Language Models for Speaker Diarization Error Correction15 Jun 2023 0 repositories listed
-
MobileASR: A resource-aware on-device learning framework for user voice personalization applications on mobile phones15 Jun 2023 0 repositories listed
-
Automated Speaker Independent Visual Speech Recognition: A Comprehensive Survey14 Jun 2023 0 repositories listed
-
EM-Network: Oracle Guided Self-distillation for Sequence Learning14 Jun 2023 0 repositories listed
-
Feature Normalization for Fine-tuning Self-Supervised Models in Speech Enhancement14 Jun 2023 0 repositories listed
-
Improving Code-Switching and Named Entity Recognition in ASR with Speech Editing based Data Augmentation14 Jun 2023 0 repositories listed
-
Learning Cross-lingual Mappings for Data Augmentation to Improve Low-Resource Speech Recognition14 Jun 2023 0 repositories listed
-
Research on an improved Conformer end-to-end Speech Recognition Model with R-Drop Structure14 Jun 2023 0 repositories listed
-
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR13 Jun 2023 0 repositories listed
-
Large-scale Language Model Rescoring on Long-form Data13 Jun 2023 0 repositories listed
-
Statistical Beamformer Exploiting Non-stationarity and Sparsity with Spatially Constrained ICA for Robust Speech Recognition13 Jun 2023 0 repositories listed
-
Multi-View Frequency-Attention Alternative to CNN Frontends for Automatic Speech Recognition12 Jun 2023 0 repositories listed
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding12 Jun 2023 0 repositories listed
-
On the N-gram Approximation of Pre-trained Language Models12 Jun 2023 0 repositories listed
-
Parameter-efficient Dysarthric Speech Recognition Using Adapter Fusion and Householder Transformation12 Jun 2023 0 repositories listed
-
Impact of Experiencing Misrecognition by Teachable Agents on Learning and Rapport11 Jun 2023 0 repositories listed
-
Modality Influence in Multimodal Machine Learning10 Jun 2023 0 repositories listed
-
What Can an Accent Identifier Learn? Probing Phonetic and Prosodic Information in a Wav2vec2-based Accent Identification Model10 Jun 2023 0 repositories listed
-
Improving Frame-level Classifier for Word Timings with Non-peaky CTC in End-to-End Automatic Speech Recognition9 Jun 2023 0 repositories listed
-
Record Deduplication for Entity Distribution Modeling in ASR Transcripts9 Jun 2023 0 repositories listed
-
Improving Language Model Integration for Neural Machine Translation8 Jun 2023 0 repositories listed
-
Latent Phrase Matching for Dysarthric Speech8 Jun 2023 0 repositories listed
-
A study on the impact of Self-Supervised Learning on automatic dysarthric speech assessment7 Jun 2023 0 repositories listed
-
An ASR-Based Tutor for Learning to Read: How to Optimize Feedback to First Graders7 Jun 2023 0 repositories listed
-
FOOCTTS: Generating Arabic Speech with Acoustic Environment for Football Commentator7 Jun 2023 0 repositories listed
-
Label Aware Speech Representation Learning For Language Identification7 Jun 2023 0 repositories listed
-
Lenient Evaluation of Japanese Speech Recognition: Modeling Naturally Occurring Spelling Inconsistency7 Jun 2023 0 repositories listed
-
Text-only Domain Adaptation using Unified Speech-Text Representation in Transducer7 Jun 2023 0 repositories listed
-
Transfer Learning from Pre-trained Language Models Improves End-to-End Speech Summarization7 Jun 2023 0 repositories listed
-
Transfer Learning of Transformer-based Speech Recognition Models from Czech to Slovak7 Jun 2023 0 repositories listed
-
Alzheimer Disease Classification through ASR-based Transcriptions: Exploring the Impact of Punctuation and Pauses6 Jun 2023 0 repositories listed
-
Automatic Assessment of Oral Reading Accuracy for Reading Diagnostics6 Jun 2023 0 repositories listed
-
Improving Fairness and Robustness in End-to-End Speech Recognition through unsupervised clustering6 Jun 2023 0 repositories listed
-
Machine Unlearning: A Survey6 Jun 2023 0 repositories listed
-
RescueSpeech: A German Corpus for Speech Recognition in Search and Rescue Domain6 Jun 2023 0 repositories listed
-
Incorporating L2 Phonemes Using Articulatory Features for Robust Speech Recognition5 Jun 2023 0 repositories listed
-
N-Shot Benchmarking of Whisper on Diverse Arabic Speech Recognition5 Jun 2023 0 repositories listed
-
OTF: Optimal Transport based Fusion of Supervised and Self-Supervised Learning Models for Automatic Speech Recognition5 Jun 2023 0 repositories listed
-
End-to-End Joint Target and Non-Target Speakers ASR4 Jun 2023 0 repositories listed
-
Audio-Visual Speech Enhancement with Score-Based Generative Models2 Jun 2023 0 repositories listed
-
Improved Training for End-to-End Streaming Automatic Speech Recognition Model with Punctuation2 Jun 2023 0 repositories listed
-
On Crowdsourcing-design with Comparison Category Rating for Evaluating Speech Enhancement Algorithms2 Jun 2023 0 repositories listed
-
Streaming Speech-to-Confusion Network Speech Recognition2 Jun 2023 0 repositories listed
-
Tensor decomposition for minimization of E2E SLU model toward on-device processing2 Jun 2023 0 repositories listed
-
Adaptation and Optimization of Automatic Speech Recognition (ASR) for the Maritime Domain in the Field of VHF Communication1 Jun 2023 0 repositories listed
-
Adapting an Unadaptable ASR System1 Jun 2023 0 repositories listed
-
Adaptive Contextual Biasing for Transducer Based Streaming Speech Recognition1 Jun 2023 0 repositories listed
-
AfriNames: Most ASR models "butcher" African Names1 Jun 2023 0 repositories listed
-
Automatic Data Augmentation for Domain Adapted Fine-Tuning of Self-Supervised Speech Representations1 Jun 2023 0 repositories listed
-
Bypass Temporal Classification: Weakly Supervised Automatic Speech Recognition with Imperfect Transcripts1 Jun 2023 0 repositories listed
-
Encoder-decoder multimodal speaker change detection1 Jun 2023 0 repositories listed
-
Enhancing the Unified Streaming and Non-streaming Model with Contrastive Learning1 Jun 2023 0 repositories listed
-
Inspecting Spoken Language Understanding from Kids for Basic Math Learning at Home1 Jun 2023 0 repositories listed
-
On the Robustness of Arabic Speech Dialect Identification1 Jun 2023 0 repositories listed
-
Some voices are too common: Building fair speech recognition systems using the Common Voice dataset1 Jun 2023 0 repositories listed
-
Speech inpainting: Context-based speech synthesis guided by video1 Jun 2023 0 repositories listed
-
Towards hate speech detection in low-resource languages: Comparing ASR to acoustic word embeddings on Wolof and Swahili1 Jun 2023 0 repositories listed
-
Accurate and Structured Pruning for Efficient Automatic Speech Recognition31 May 2023 0 repositories listed
-
Simple yet Effective Code-Switching Language Identification with Multitask Pre-Training and Transfer Learning31 May 2023 0 repositories listed
-
Strategies for improving low resource speech to text translation relying on pre-trained ASR models31 May 2023 0 repositories listed
-
The Tag-Team Approach: Leveraging CLS and Language Tagging for Enhancing Multilingual ASR31 May 2023 0 repositories listed
-
VILAS: Exploring the Effects of Vision and Language Context in Automatic Speech Recognition31 May 2023 0 repositories listed
-
Zero-Shot Automatic Pronunciation Assessment31 May 2023 0 repositories listed
-
Adapting Multi-Lingual ASR Models for Handling Multiple Talkers30 May 2023 0 repositories listed
-
STT4SG-350: A Speech Corpus for All Swiss German Dialect Regions30 May 2023 0 repositories listed
-
The News Delivery Channel Recommendation Based on Granular Neural Network30 May 2023 0 repositories listed
-
Towards Selection of Text-to-speech Data to Augment ASR Training30 May 2023 0 repositories listed
-
An Experimental Review of Speaker Diarization methods with application to Two-Speaker Conversational Telephone Speech recordings29 May 2023 0 repositories listed
-
Building Accurate Low Latency ASR for Streaming Voice Search29 May 2023 0 repositories listed
-
Can We Trust Explainable AI Methods on ASR? An Evaluation on Phoneme Recognition29 May 2023 0 repositories listed
-
Improving Textless Spoken Language Understanding with Discrete Units as Intermediate Target29 May 2023 0 repositories listed
-
Retraining-free Customized ASR for Enharmonic Words Based on a Named-Entity-Aware Model and Phoneme Similarity Estimation29 May 2023 0 repositories listed
-
RASR2: The RWTH ASR Toolkit for Generic Sequence-to-sequence Speech Recognition28 May 2023 0 repositories listed
-
A Comprehensive Overview and Comparative Analysis on Deep Learning Models: CNN, RNN, LSTM, GRU27 May 2023 0 repositories listed
-
2-bit Conformer quantization for automatic speech recognition26 May 2023 0 repositories listed
-
DisfluencyFixer: A tool to enhance Language Learning through Speech To Speech Disfluency Correction26 May 2023 0 repositories listed
-
Robustness of Multi-Source MT to Transcription Errors26 May 2023 0 repositories listed
-
ASR and Emotional Speech: A Word-Level Investigation of the Mutual Impact of Speech and Emotion Recognition25 May 2023 0 repositories listed
-
Improving Scheduled Sampling for Neural Transducer-based ASR25 May 2023 0 repositories listed
-
INTapt: Information-Theoretic Adversarial Prompt Tuning for Enhanced Non-Native Speech Recognition25 May 2023 0 repositories listed
-
Knowledge Distillation for Neural Transducer-based Target-Speaker ASR: Exploiting Parallel Mixture/Single-Talker Speech Data25 May 2023 0 repositories listed
-
Mixture-of-Expert Conformer for Streaming Multilingual ASR25 May 2023 0 repositories listed
-
Persistent Laplacian-enhanced Algorithm for Scarcely Labeled Data Classification25 May 2023 0 repositories listed
-
Svarah: Evaluating English ASR Systems on Indian Accents25 May 2023 0 repositories listed
-
Unified Modeling of Multi-Talker Overlapped Speech Recognition and Diarization with a Sidecar Separator25 May 2023 0 repositories listed
-
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation25 May 2023 0 repositories listed
-
Weakly-Supervised Speech Pre-training: A Case Study on Target Speech Recognition25 May 2023 0 repositories listed
-
Incorporating Ultrasound Tongue Images for Audio-Visual Speech Enhancement through Knowledge Distillation24 May 2023 0 repositories listed
-
InterFormer: Interactive Local and Global Features Fusion for Automatic Speech Recognition24 May 2023 0 repositories listed
-
Iteratively Improving Speech Recognition and Voice Conversion24 May 2023 0 repositories listed
-
Spoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM24 May 2023 0 repositories listed
-
RAND: Robustness Aware Norm Decay For Quantized Seq2seq Models24 May 2023 0 repositories listed
-
BA-SOT: Boundary-Aware Serialized Output Training for Multi-Talker ASR23 May 2023 0 repositories listed
-
Cross-lingual Knowledge Transfer and Iterative Pseudo-labeling for Low-Resource Speech Recognition with Transducers23 May 2023 0 repositories listed