Browse State-of-the-Art › Speech Recognition › Papers, page 25
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 25 of 65: papers 2,401 to 2,500 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Adapting Text-based Dialogue State Tracker for Spoken Dialogues29 Aug 2023 0 repositories listed
-
Neural approaches to spoken content embedding28 Aug 2023 0 repositories listed
-
The USTC-NERCSLIP Systems for the CHiME-7 DASR Challenge28 Aug 2023 0 repositories listed
-
Unsupervised Active Learning: Optimizing Labeling Cost-Effectiveness for Automatic Speech Recognition28 Aug 2023 0 repositories listed
-
Decoupled Structure for Improved Adaptability of End-to-End Models25 Aug 2023 0 repositories listed
-
24 Aug 2023 0 repositories listed
-
AdVerb: Visually Guided Audio Dereverberation23 Aug 2023 0 repositories listed
-
KinSPEAK: Improving speech recognition for Kinyarwanda via semi-supervised learning methods23 Aug 2023 0 repositories listed
-
Convoifilter: A case study of doing cocktail party speech recognition22 Aug 2023 0 repositories listed
-
Identifying depression-related topics in smartphone-collected free-response speech recordings using an automatic speech recognition system and a deep learning topic model22 Aug 2023 0 repositories listed
-
Improving Continuous Sign Language Recognition with Cross-Lingual Signs21 Aug 2023 0 repositories listed
-
TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition21 Aug 2023 0 repositories listed
-
20 Aug 2023 0 repositories listed
-
Indonesian Automatic Speech Recognition with XLSR-5320 Aug 2023 0 repositories listed
-
Accurate synthesis of Dysarthric Speech for ASR data augmentation16 Aug 2023 0 repositories listed
-
Radio2Text: Streaming Speech Recognition Using mmWave Radio Signals16 Aug 2023 0 repositories listed
-
AKVSR: Audio Knowledge Empowered Visual Speech Recognition by Compressing Audio Knowledge of a Pretrained Model15 Aug 2023 0 repositories listed
-
Improving CTC-AED model with integrated-CTC and auxiliary loss regularization15 Aug 2023 0 repositories listed
-
Cross-Attribute Matrix Factorization Model with Shared User Embedding14 Aug 2023 0 repositories listed
-
O-1: Self-training with Oracle and 1-best Hypothesis14 Aug 2023 0 repositories listed
-
Text Injection for Capitalization and Turn-Taking Prediction in Speech Models14 Aug 2023 0 repositories listed
-
Using Text Injection to Improve Recognition of Personal Identifiers in Speech14 Aug 2023 0 repositories listed
-
Alternative Pseudo-Labeling for Semi-Supervised Automatic Speech Recognition12 Aug 2023 0 repositories listed
-
Bilingual Streaming ASR with Grapheme units and Auxiliary Monolingual Loss11 Aug 2023 0 repositories listed
-
Improving Joint Speech-Text Representations Without Alignment11 Aug 2023 0 repositories listed
-
11 Aug 2023 0 repositories listed Syntology 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 1 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
A Novel Self-training Approach for Low-resource Speech Recognition10 Aug 2023 0 repositories listed
-
A Novel Method for improving accuracy in neural network by reinstating traditional back propagation technique9 Aug 2023 0 repositories listed
-
FPGA Resource-aware Structured Pruning for Real-Time Neural Networks9 Aug 2023 0 repositories listed
-
TSSR: A Truncated and Signed Square Root Activation Function for Neural Networks9 Aug 2023 0 repositories listed
-
Unsupervised Out-of-Distribution Dialect Detection with Mahalanobis Distance9 Aug 2023 0 repositories listed
-
Comparative Analysis of the wav2vec 2.0 Feature Extractor8 Aug 2023 0 repositories listed
-
Boosting Chinese ASR Error Correction with Dynamic Error Scaling Mechanism7 Aug 2023 0 repositories listed
-
Dialogue Systems Can Generate Appropriate Responses without the Use of Question Marks? -- Investigation of the Effects of Question Marks on Dialogue Systems7 Aug 2023 0 repositories listed
-
A Critical Review of Physics-Informed Machine Learning Applications in Subsurface Energy Systems6 Aug 2023 0 repositories listed
-
ApproBiVT: Lead ASR Models to Generalize Better Using Approximated Bias-Variance Tradeoff Guided Early Stopping and Checkpoint Averaging5 Aug 2023 0 repositories listed
-
Speaker Diarization of Scripted Audiovisual Content4 Aug 2023 0 repositories listed
-
Federated Representation Learning for Automatic Speech Recognition3 Aug 2023 0 repositories listed
-
Careful Whisper -- leveraging advances in automatic speech recognition for robust and interpretable aphasia subtype classification2 Aug 2023 0 repositories listed
-
Improving grapheme-to-phoneme conversion by learning pronunciations from speech recordings31 Jul 2023 0 repositories listed
-
Pre-training End-to-end ASR Models with Augmented Speech Samples Queried by Text30 Jul 2023 0 repositories listed
-
UniBriVL: Robust Universal Representation and Generation of Audio Driven Diffusion Models29 Jul 2023 0 repositories listed
-
The timing bottleneck: Why timing and overlap are mission-critical for conversational user interfaces, speech recognition and dialogue systems28 Jul 2023 0 repositories listed
-
Cascaded Cross-Modal Transformer for Request and Complaint Detection27 Jul 2023 0 repositories listed
-
CIF-T: A Novel CIF-based Transducer Architecture for Automatic Speech Recognition26 Jul 2023 0 repositories listed
-
On-Device Speaker Anonymization of Acoustic Embeddings for ASR based onFlexible Location Gradient Reversal Layer25 Jul 2023 0 repositories listed
-
Boosting Punctuation Restoration with Data Generation and Reinforcement Learning24 Jul 2023 0 repositories listed
-
Integration of Frame- and Label-synchronous Beam Search for Streaming Encoder-decoder Speech Recognition24 Jul 2023 0 repositories listed
-
Robust Automatic Speech Recognition via WavAugment Guided Phoneme Adversarial Training24 Jul 2023 0 repositories listed
-
A meta learning scheme for fast accent domain expansion in Mandarin speech recognition23 Jul 2023 0 repositories listed
-
Exploring the Integration of Speech Separation and Recognition with Self-Supervised Learning Representation23 Jul 2023 0 repositories listed
-
Modality Confidence Aware Training for Robust End-to-End Spoken Language Understanding22 Jul 2023 0 repositories listed
-
Prompting Large Language Models with Speech Recognition Abilities21 Jul 2023 0 repositories listed
-
Globally Normalising the Transducer for Streaming Speech Recognition20 Jul 2023 0 repositories listed
-
Integrating Pretrained ASR and LM to Perform Sequence Generation for Spoken Language Understanding20 Jul 2023 0 repositories listed
-
MASR: Multi-label Aware Speech Representation20 Jul 2023 0 repositories listed
-
Transsion TSUP's speech recognition system for ASRU 2023 MADASR Challenge20 Jul 2023 0 repositories listed
-
Leveraging Visemes for Better Visual Speech Representation and Lip Reading19 Jul 2023 0 repositories listed
-
Model Adaptation for ASR in low-resource Indian Languages16 Jul 2023 0 repositories listed
-
Ed-Fed: A generic federated learning framework with resource-aware client selection for edge devices14 Jul 2023 0 repositories listed
-
On the Sensitivity of Deep Load Disaggregation to Adversarial Attacks14 Jul 2023 0 repositories listed
-
Replay to Remember: Continual Layer-Specific Fine-tuning for German Speech Recognition14 Jul 2023 0 repositories listed
-
Representation Learning With Hidden Unit Clustering For Low Resource Speech Applications14 Jul 2023 0 repositories listed
-
Towards Model-Size Agnostic, Compute-Free, Memorization-based Inference of Deep Learning14 Jul 2023 0 repositories listed
-
Towards spoken dialect identification of Irish14 Jul 2023 0 repositories listed
-
Exploring the Integration of Large Language Models into Automatic Speech Recognition Systems: An Empirical Study13 Jul 2023 0 repositories listed
-
Leveraging Pretrained ASR Encoders for Effective and Efficient End-to-End Speech Intent Classification and Slot Filling13 Jul 2023 0 repositories listed
-
Personalization for BERT-based Discriminative Speech Recognition Rescoring13 Jul 2023 0 repositories listed
-
Speech Diarization and ASR with GMM11 Jul 2023 0 repositories listed
-
SparseVSR: Lightweight and Noise Robust Visual Speech Recognition10 Jul 2023 0 repositories listed
-
Can Generative Large Language Models Perform ASR Error Correction?9 Jul 2023 0 repositories listed
-
Token-Level Serialized Output Training for Joint Streaming ASR and ST Leveraging Textual Alignments7 Jul 2023 0 repositories listed
-
Online Hybrid CTC/Attention End-to-End Automatic Speech Recognition Architecture5 Jul 2023 0 repositories listed
-
Transgressing the boundaries: towards a rigorous understanding of deep learning and its (non-)robustness5 Jul 2023 0 repositories listed
-
Using Data Augmentations and VTLN to Reduce Bias in Dutch End-to-End Speech Recognition Systems5 Jul 2023 0 repositories listed
-
Align With Purpose: Optimize Desired Properties in CTC Models with a General Plug-and-Play Framework4 Jul 2023 0 repositories listed
-
Boosting Norwegian Automatic Speech Recognition4 Jul 2023 0 repositories listed
-
Knowledge-Aware Audio-Grounded Generative Slot Filling for Limited Annotated Data4 Jul 2023 0 repositories listed
-
Transcribing Educational Videos Using Whisper: A preliminary study on using AI for transcribing educational videos4 Jul 2023 0 repositories listed
-
Multilingual Contextual Adapters To Improve Custom Word Recognition In Low-resource Languages3 Jul 2023 0 repositories listed
-
Conformer LLMs -- Convolution Augmented Large Language Models2 Jul 2023 0 repositories listed
-
Don't Stop Self-Supervision: Accent Adaptation of Speech Representations via Residual Adapters2 Jul 2023 0 repositories listed
-
Automatic Speech Recognition of Non-Native Child Speech for Language Learning Applications29 Jun 2023 0 repositories listed
-
Leveraging Cross-Utterance Context For ASR Decoding29 Jun 2023 0 repositories listed
-
Accelerating Transducers through Adjacent Token Merging28 Jun 2023 0 repositories listed
-
Prompting Large Language Models for Zero-Shot Domain Adaptation in Speech Recognition28 Jun 2023 0 repositories listed
-
A Survey on Deep Learning Hardware Accelerators for Heterogeneous HPC Platforms27 Jun 2023 0 repositories listed
-
Confidence-based Ensembles of End-to-End Speech Recognition Models27 Jun 2023 0 repositories listed
-
Hyper-parameter Adaptation of Conformer ASR Systems for Elderly and Dysarthric Speech Recognition27 Jun 2023 0 repositories listed
-
Large-scale unsupervised audio pre-training for video-to-speech synthesis27 Jun 2023 0 repositories listed
-
Reducing the gap between streaming and non-streaming Transducer-based ASR by adaptive two-stage knowledge distillation27 Jun 2023 0 repositories listed
-
Scaling Laws for Discriminative Speech Recognition Rescoring Models27 Jun 2023 0 repositories listed
-
Factorised Speaker-environment Adaptive Training of Conformer Speech Recognition Systems26 Jun 2023 0 repositories listed
-
Master-ASR: Achieving Multilingual Scalability and Low-Resource Adaptation in ASR with Modular Learning23 Jun 2023 0 repositories listed
-
Meta-Gating Framework for Fast and Continuous Resource Optimization in Dynamic Wireless Environments23 Jun 2023 0 repositories listed
-
The CHiME-7 DASR Challenge: Distant Meeting Transcription with Multiple Devices in Diverse Scenarios23 Jun 2023 0 repositories listed
-
Towards Effective and Compact Contextual Representation for Conformer Transducer Speech Recognition Systems23 Jun 2023 0 repositories listed
-
AudioPaLM: A Large Language Model That Can Speak and Listen22 Jun 2023 0 repositories listed
-
Exploring the Role of Audio in Video Captioning21 Jun 2023 0 repositories listed
-
Federated Self-Learning with Weak Supervision for Speech Recognition21 Jun 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.