Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 14
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 14 of 31: papers 1,301 to 1,400 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Speech Corpora Divergence Based Unsupervised Data Selection for ASR26 Feb 2023 0 repositories listed
-
Ensemble knowledge distillation of self-supervised speech models24 Feb 2023 0 repositories listed
-
Factual Consistency Oriented Speech Recognition24 Feb 2023 0 repositories listed
-
Evaluating Automatic Speech Recognition in an Incremental Setting23 Feb 2023 0 repositories listed
-
Improving Contextual Spelling Correction by External Acoustics Attention and Semantic Aware Data Augmentation22 Feb 2023 0 repositories listed
-
MADI: Inter-domain Matching and Intra-domain Discrimination for Cross-domain Speech Recognition22 Feb 2023 0 repositories listed
-
UML: A Universal Monolingual Output Layer for Multilingual ASR22 Feb 2023 0 repositories listed
-
Connecting Humanities and Social Sciences: Applying Language and Speech Technology to Online Panel Surveys21 Feb 2023 0 repositories listed
-
An ASR-free Fluency Scoring Approach with Self-Supervised Learning20 Feb 2023 0 repositories listed
-
Emphasizing Unseen Words: New Vocabulary Acquisition for End-to-End Speech Recognition20 Feb 2023 0 repositories listed
-
Speaker and Language Change Detection using Wav2vec2 and Whisper18 Feb 2023 0 repositories listed
-
Massively Multilingual Shallow Fusion with Large Language Models17 Feb 2023 0 repositories listed
-
Adaptable End-to-End ASR Models using Replaceable Internal LMs and Residual Softmax16 Feb 2023 0 repositories listed
-
16 Feb 2023 0 repositories listed
-
Speaker Change Detection for Transformer Transducer ASR16 Feb 2023 0 repositories listed
-
Stabilising and accelerating light gated recurrent units for automatic speech recognition16 Feb 2023 0 repositories listed
-
ASR Bundestag: A Large-Scale political debate dataset in German12 Feb 2023 0 repositories listed
-
PATCorrect: Non-autoregressive Phoneme-augmented Transformer for ASR Error Correction10 Feb 2023 0 repositories listed
-
Leveraging supplementary text data to kick-start automatic speech recognition system development with limited transcriptions9 Feb 2023 0 repositories listed
-
MAC: A unified framework boosting low resource automatic speech recognition5 Feb 2023 0 repositories listed
-
Improving Rare Words Recognition through Homophone Extension and Unified Writing for Low-resource Cantonese Speech Recognition2 Feb 2023 0 repositories listed
-
Fillers in Spoken Language Understanding: Computational and Psycholinguistic Perspectives25 Jan 2023 0 repositories listed
-
A Multi-Purpose Audio-Visual Corpus for Multi-Modal Persian Speech Recognition: the Arman-AV Dataset21 Jan 2023 0 repositories listed
-
Language Agnostic Data-Driven Inverse Text Normalization20 Jan 2023 0 repositories listed
-
From English to More Languages: Parameter-Efficient Model Reprogramming for Cross-Lingual Speech Recognition19 Jan 2023 0 repositories listed
-
BayesSpeech: A Bayesian Transformer Network for Automatic Speech Recognition16 Jan 2023 0 repositories listed
-
Multi-resolution location-based training for multi-channel continuous speech separation16 Jan 2023 0 repositories listed
-
Using Kaldi for Automatic Speech Recognition of Conversational Austrian German16 Jan 2023 0 repositories listed
-
Streaming Punctuation: A Novel Punctuation Technique Leveraging Bidirectional Context for Continuous Speech Recognition10 Jan 2023 0 repositories listed
-
Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition30 Dec 2022 0 repositories listed
-
Don't Be So Sure! Boosting ASR Decoding via Confidence Relaxation27 Dec 2022 0 repositories listed
-
Alignment Entropy Regularization22 Dec 2022 0 repositories listed
-
4D ASR: Joint modeling of CTC, Attention, Transducer, and Mask-Predict decoders21 Dec 2022 0 repositories listed
-
End-to-End Automatic Speech Recognition model for the Sudanese Dialect21 Dec 2022 0 repositories listed
-
Mu²SLAM: Multitask, Multilingual Speech and Language Models19 Dec 2022 0 repositories listed
-
Context-aware Fine-tuning of Self-supervised Speech Models16 Dec 2022 0 repositories listed
-
Fast Entropy-Based Methods of Word-Level Confidence Estimation for End-To-End Automatic Speech Recognition16 Dec 2022 0 repositories listed
-
Speech Aware Dialog System Technology Challenge (DSTC11)16 Dec 2022 0 repositories listed
-
Improving Fast-slow Encoder based Transducer with Streaming Deliberation15 Dec 2022 0 repositories listed
-
Disentangling Prosody Representations with Unsupervised Speech Reconstruction14 Dec 2022 0 repositories listed
-
Speech and Natural Language Processing Technologies for Pseudo-Pilot Simulator14 Dec 2022 0 repositories listed
-
End-to-End Speech Translation of Arabic to English Broadcast News11 Dec 2022 0 repositories listed
-
Improved Self-Supervised Multilingual Speech Representation Learning Combined with Auxiliary Language Information7 Dec 2022 0 repositories listed
-
Improved Speech Pre-Training with Supervision-Enhanced Acoustic Unit7 Dec 2022 0 repositories listed
-
Lattice-Free Sequence Discriminative Training for Phoneme-Based Neural Transducers7 Dec 2022 0 repositories listed
-
Progressive Multi-Scale Self-Supervised Learning for Speech Recognition7 Dec 2022 0 repositories listed
-
Unsupervised Fine-Tuning Data Selection for ASR Using Self-Supervised Speech Models3 Dec 2022 0 repositories listed
-
Continual Learning for On-Device Speech Recognition using Disentangled Conformers2 Dec 2022 0 repositories listed
-
Preliminary Study on SSCF-derived Polar Coordinate for ASR30 Nov 2022 0 repositories listed
-
Better Transcription of UK Supreme Court Hearings29 Nov 2022 0 repositories listed
-
Evaluating and reducing the distance between synthetic and real speech distributions29 Nov 2022 0 repositories listed
-
Neural Transducer Training: Reduced Memory Consumption with Sample-wise Computation29 Nov 2022 0 repositories listed
-
Inter-KD: Intermediate Knowledge Distillation for CTC-Based Automatic Speech Recognition28 Nov 2022 0 repositories listed
-
Bidirectional Representations for Low Resource Spoken Language Understanding24 Nov 2022 0 repositories listed
-
Multitask Learning for Low Resource Spoken Language Understanding24 Nov 2022 0 repositories listed
-
Device Directedness with Contextual Cues for Spoken Dialog Systems23 Nov 2022 0 repositories listed
-
22 Nov 2022 0 repositories listed
-
Complex-Valued Time-Frequency Self-Attention for Speech Dereverberation22 Nov 2022 0 repositories listed
-
SSCFormer: Push the Limit of Chunk-wise Conformer for Streaming ASR Using Sequentially Sampled Chunks and Chunked Causal Convolution21 Nov 2022 0 repositories listed
-
SpeechNet: Weakly Supervised, End-to-End Speech Recognition at Industrial Scale21 Nov 2022 0 repositories listed
-
Hey ASR System! Why Aren't You More Inclusive? Automatic Speech Recognition Systems' Bias and Proposed Bias Mitigation Techniques. A Literature Review17 Nov 2022 0 repositories listed
-
LongFNT: Long-form Speech Recognition with Factorized Neural Transducer17 Nov 2022 0 repositories listed
-
Unsupervised Model-based speaker adaptation of end-to-end lattice-free MMI model for speech recognition17 Nov 2022 0 repositories listed
-
Improving Speech Emotion Recognition with Unsupervised Speaking Style Transfer16 Nov 2022 0 repositories listed
-
On using the UA-Speech and TORGO databases to validate automatic dysarthric speech classification approaches16 Nov 2022 0 repositories listed
-
Introducing Semantics into Speech Encoders15 Nov 2022 0 repositories listed
-
Align, Write, Re-order: Explainable End-to-End Speech Translation via Operation Sequence Generation11 Nov 2022 0 repositories listed
-
Handling Trade-Offs in Speech Separation with Sparsely-Gated Mixture of Experts11 Nov 2022 0 repositories listed
-
A Study on the Integration of Pre-trained SSL, ASR, LM and SLU Models for Spoken Language Understanding10 Nov 2022 0 repositories listed
-
Adaptive Multi-Corpora Language Model Training for Speech Recognition9 Nov 2022 0 repositories listed
-
Improving Noisy Student Training on Non-target Domain Data for Automatic Speech Recognition9 Nov 2022 0 repositories listed
-
End-to-End Evaluation of a Spoken Dialogue System for Learning Basic Mathematics7 Nov 2022 0 repositories listed
-
Streaming, fast and accurate on-device Inverse Text Normalization for Automatic Speech Recognition7 Nov 2022 0 repositories listed
-
Bridging Speech and Textual Pre-trained Models with Unsupervised ASR6 Nov 2022 0 repositories listed
-
Evaluation of Automated Speech Recognition Systems for Conversational Speech: A Linguistic Perspective5 Nov 2022 0 repositories listed
-
LAMASSU: Streaming Language-Agnostic Multilingual Speech Recognition and Translation Using Neural Transducers5 Nov 2022 0 repositories listed
-
Biased Self-supervised learning for ASR4 Nov 2022 0 repositories listed
-
Resource-Efficient Transfer Learning From Speech Foundation Model Using Hierarchical Feature Fusion4 Nov 2022 0 repositories listed
-
Stutter-TTS: Controlled Synthesis and Improved Recognition of Stuttered Speech4 Nov 2022 0 repositories listed
-
H_eval: A new hybrid evaluation metric for automatic speech recognition tasks3 Nov 2022 0 repositories listed
-
Leveraging Domain Features for Detecting Adversarial Attacks Against Deep Speech Recognition in Noise3 Nov 2022 0 repositories listed
-
Phonetic-assisted Multi-Target Units Modeling for Improving Conformer-Transducer ASR system3 Nov 2022 0 repositories listed
-
Probing Statistical Representations For End-To-End ASR3 Nov 2022 0 repositories listed
-
Streaming Audio-Visual Speech Recognition with Alignment Regularization3 Nov 2022 0 repositories listed
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder2 Nov 2022 0 repositories listed
-
Monolingual Recognizers Fusion for Code-switching Speech Recognition2 Nov 2022 0 repositories listed
-
More Speaking or More Speakers?2 Nov 2022 0 repositories listed
-
Towards Zero-Shot Code-Switched Speech Recognition2 Nov 2022 0 repositories listed
-
A Comparative Study on Multichannel Speaker-Attributed Automatic Speech Recognition in Multi-party Meetings1 Nov 2022 0 repositories listed
-
A Preliminary Study on Automated Speaking Assessment of English as a Second Language (ESL) Students1 Nov 2022 0 repositories listed
-
Adapting self-supervised models to multi-talker speech recognition using speaker embeddings1 Nov 2022 0 repositories listed
-
Mandarin-English Code-Switching Speech Recognition System for Specific Domain1 Nov 2022 0 repositories listed
-
Unified End-to-End Speech Recognition and Endpointing for Fast and Efficient Speech Systems1 Nov 2022 0 repositories listed
-
An analysis of degenerating speech due to progressive dysarthria on ASR performance31 Oct 2022 0 repositories listed
-
Audio-Visual Speech Enhancement and Separation by Utilizing Multi-Modal Self-Supervised Embeddings31 Oct 2022 0 repositories listed
-
DiaCorrect: End-to-end error correction for speaker diarization31 Oct 2022 0 repositories listed
-
FusionFormer: Fusing Operations in Transformer for Efficient Streaming Speech Recognition31 Oct 2022 0 repositories listed
-
Structured State Space Decoder for Speech Recognition and Synthesis31 Oct 2022 0 repositories listed
-
DuDe: Dual-Decoder Multilingual ASR for Indian Languages using Common Label Set30 Oct 2022 0 repositories listed
-
Phonemic Representation and Transcription for Speech to Text Applications for Under-resourced Indigenous African Languages: The Case of Kiswahili29 Oct 2022 0 repositories listed