Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 20
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 20 of 32: papers 1,901 to 2,000 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
End-to-End Multi-speaker ASR with Independent Vector Analysis1 Apr 2022 0 repositories listed
-
Alternate Intermediate Conditioning with Syllable-level and Character-level Targets for Japanese ASR1 Apr 2022 0 repositories listed
-
Multi-task RNN-T with Semantic Decoder for Streamable Spoken Language Understanding1 Apr 2022 0 repositories listed
-
Probing Speech Emotion Recognition Transformers for Linguistic Knowledge1 Apr 2022 0 repositories listed
-
Text-To-Speech Data Augmentation for Low Resource Speech Recognition1 Apr 2022 0 repositories listed
-
Zero-Shot Cross-lingual Aphasia Detection using Automatic Speech Recognition1 Apr 2022 0 repositories listed
-
A Comparative Study on Speaker-attributed Automatic Speech Recognition in Multi-party Meetings31 Mar 2022 0 repositories listed
-
Analyzing the factors affecting usefulness of Self-Supervised Pre-trained Representations for Speech Recognition31 Mar 2022 0 repositories listed
-
Effectiveness of text to speech pseudo labels for forced alignment and cross lingual pretrained models for low resource speech recognition31 Mar 2022 0 repositories listed
-
Importance of Different Temporal Modulations of Speech: A Tale of Two Perspectives31 Mar 2022 0 repositories listed
-
Memory-Efficient Training of RNN-Transducer with Sampled Softmax31 Mar 2022 0 repositories listed
-
31 Mar 2022 0 repositories listed
-
Code Switched and Code Mixed Speech Recognition for Indic languages30 Mar 2022 0 repositories listed
-
Improving Speech Recognition for Indic Languages using Language Model30 Mar 2022 0 repositories listed
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages?30 Mar 2022 0 repositories listed
-
Dynamic Latency for CTC-Based Streaming Automatic Speech Recognition With Emformer29 Mar 2022 0 repositories listed
-
Frequency-Directional Attention Model for Multilingual Automatic Speech Recognition29 Mar 2022 0 repositories listed
-
Improving Generalization of Deep Neural Network Acoustic Models with Length Perturbation and N-best Based Label Smoothing29 Mar 2022 0 repositories listed
-
Mel Frequency Spectral Domain Defenses against Adversarial Attacks on Speech Recognition Systems29 Mar 2022 0 repositories listed
-
Short-Term Word-Learning in a Dynamically Changing Environment29 Mar 2022 0 repositories listed
-
Impact of Dataset on Acoustic Models for Automatic Speech Recognition25 Mar 2022 0 repositories listed
-
Computing Optimal Location of Microphone for Improved Speech Recognition24 Mar 2022 0 repositories listed
-
Disentangleing Content and Fine-grained Prosody Information via Hybrid ASR Bottleneck Features for Voice Conversion24 Mar 2022 0 repositories listed
-
Lahjoita puhetta -- a large-scale corpus of spoken Finnish with some benchmarks24 Mar 2022 0 repositories listed
-
A Text-to-Speech Pipeline, Evaluation Methodology, and Initial Fine-Tuning Results for Child Speech Synthesis22 Mar 2022 0 repositories listed
-
Building Robust Spoken Language Understanding by Cross Attention between Phoneme Sequence and ASR Hypothesis22 Mar 2022 0 repositories listed
-
Pseudo Label Is Better Than Human Label22 Mar 2022 0 repositories listed
-
Exploiting Cross Domain Acoustic-to-articulatory Inverted Features For Disordered Speech Recognition19 Mar 2022 0 repositories listed
-
Representative Subset Selection for Efficient Fine-Tuning in Self-Supervised Speech Recognition18 Mar 2022 0 repositories listed
-
Prediction of speech intelligibility with DNN-based performance measures17 Mar 2022 0 repositories listed
-
Whither the Priors for (Vocal) Interactivity?16 Mar 2022 0 repositories listed
-
Spectral Modification Based Data Augmentation For Improving End-to-End ASR For Children's Speech13 Mar 2022 0 repositories listed
-
Transformer-based Streaming ASR with Cumulative Attention11 Mar 2022 0 repositories listed
-
Attacks as Defenses: Designing Robust Audio CAPTCHAs Using Attacks on Automatic Speech Recognition Systems10 Mar 2022 0 repositories listed
-
A practical framework for multi-domain speech recognition and an instance sampling method to neural language modeling9 Mar 2022 0 repositories listed
-
Which French speech recognition system for assistant robots?4 Mar 2022 0 repositories listed
-
A Conformer Based Acoustic Model for Robust Automatic Speech Recognition1 Mar 2022 0 repositories listed
-
Extended Graph Temporal Classification for Multi-Speaker End-to-End ASR1 Mar 2022 0 repositories listed
-
Measuring the Impact of Individual Domain Factors in Self-Supervised Pre-Training1 Mar 2022 0 repositories listed
-
Integrating Text Inputs For Training and Adapting RNN Transducer ASR Models26 Feb 2022 0 repositories listed
-
A Survey of Multilingual Models for Automatic Speech Recognition25 Feb 2022 0 repositories listed
-
Language technology practitioners as language managers: arbitrating data bias and predictive bias in ASR25 Feb 2022 0 repositories listed
-
Ask2Mask: Guided Data Selection for Masked Speech Modeling24 Feb 2022 0 repositories listed
-
Towards Better Meta-Initialization with Task Augmentation for Kindergarten-aged Speech Recognition24 Feb 2022 0 repositories listed
-
Differentially Private Speaker Anonymization23 Feb 2022 0 repositories listed
-
Korean Tokenization for Beam Search Rescoring in Speech Recognition22 Feb 2022 0 repositories listed
-
VADOI:Voice-Activity-Detection Overlapping Inference For End-to-end Long-form Speech Recognition22 Feb 2022 0 repositories listed
-
r-G2P: Evaluating and Enhancing Robustness of Grapheme to Phoneme Conversion by Controlled noise introducing and Contextual information incorporation21 Feb 2022 0 repositories listed
-
Speaker Adaptation Using Spectro-Temporal Deep Features for Dysarthric and Elderly Speech Recognition21 Feb 2022 0 repositories listed
-
19 Feb 2022 0 repositories listed
-
Domain Adaptation of low-resource Target-Domain models using well-trained ASR Conformer Models18 Feb 2022 0 repositories listed
-
'Beach' to 'Bitch': Inadvertent Unsafe Transcription of Kids' Content on YouTube17 Feb 2022 0 repositories listed
-
Mitigating Closed-model Adversarial Examples with Bayesian Neural Modeling for Enhanced End-to-End Speech Recognition17 Feb 2022 0 repositories listed
-
MLP-ASR: Sequence-length agnostic all-MLP architectures for speech recognition17 Feb 2022 0 repositories listed
-
Conversational Speech Recognition By Learning Conversation-level Characteristics16 Feb 2022 0 repositories listed
-
Knowledge Transfer from Large-scale Pretrained Language Models to End-to-end Speech Recognizers16 Feb 2022 0 repositories listed
-
Multi-style Training for South African Call Centre Audio15 Feb 2022 0 repositories listed
-
Saving RNN Computations with a Neuron-Level Fuzzy Memoization Scheme14 Feb 2022 0 repositories listed
-
Multimodal Depression Classification Using Articulatory Coordination Features And Hierarchical Attention Based Text Embeddings13 Feb 2022 0 repositories listed
-
A two-step approach to leverage contextual data: speech recognition in air-traffic communications8 Feb 2022 0 repositories listed
-
Enhancing ASR for Stuttered Speech with Limited Data Using Detect and Pass8 Feb 2022 0 repositories listed
-
Polyphonic pitch detection with convolutional recurrent neural networks4 Feb 2022 0 repositories listed
-
The CUHK-TENCENT speaker diarization system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge4 Feb 2022 0 repositories listed
-
Joint Speech Recognition and Audio Captioning3 Feb 2022 0 repositories listed
-
The RoyalFlush System of Speech Recognition for M2MeT Challenge3 Feb 2022 0 repositories listed
-
ASR-Aware End-to-end Neural Diarization2 Feb 2022 0 repositories listed
-
Error Correction in ASR using Sequence-to-Sequence Models2 Feb 2022 0 repositories listed
-
RescoreBERT: Discriminative Speech Recognition Rescoring with BERT2 Feb 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Feb 2022 0 repositories listed
-
Language Dependencies in Adversarial Attacks on Speech Recognition Systems1 Feb 2022 0 repositories listed
-
Visualizing Automatic Speech Recognition -- Means for a Better Understanding?1 Feb 2022 0 repositories listed
-
Reducing language context confusion for end-to-end code-switching automatic speech recognition28 Jan 2022 0 repositories listed
-
Sentiment-Aware Automatic Speech Recognition pre-training for enhanced Speech Emotion Recognition27 Jan 2022 0 repositories listed
-
Synthesizing Dysarthric Speech Using Multi-talker TTS for Dysarthric Speech Recognition27 Jan 2022 0 repositories listed
-
On the Effectiveness of Pinyin-Character Dual-Decoding for End-to-End Mandarin Chinese ASR26 Jan 2022 0 repositories listed
-
26 Jan 2022 0 repositories listed
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models25 Jan 2022 0 repositories listed
-
Improving the fusion of acoustic and text representations in RNN-T25 Jan 2022 0 repositories listed
-
Run-and-back stitch search: novel block synchronous decoding for streaming encoder-decoder ASR25 Jan 2022 0 repositories listed
-
Transformer-Based Video Front-Ends for Audio-Visual Speech Recognition for Single and Multi-Person Video25 Jan 2022 0 repositories listed
-
A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition22 Jan 2022 0 repositories listed
-
How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR18 Jan 2022 0 repositories listed
-
Human and Automatic Speech Recognition Performance on German Oral History Interviews18 Jan 2022 0 repositories listed
-
DUAL: Textless Spoken Question Answering with Speech Discrete Unit Adaptive Learning16 Jan 2022 0 repositories listed
-
RED-ACE: Robust Error Detection for ASR using Confidence Embeddings16 Jan 2022 0 repositories listed
-
Recent Progress in the CUHK Dysarthric Speech Recognition System15 Jan 2022 0 repositories listed
-
Learning to Enhance or Not: Neural Network-Based Switching of Enhanced and Observed Signals for Overlapping Speech Recognition11 Jan 2022 0 repositories listed
-
A Likelihood Ratio based Domain Adaptation Method for E2E Models10 Jan 2022 0 repositories listed
-
Cross-Modal ASR Post-Processing System for Error Correction and Utterance Rejection10 Jan 2022 0 repositories listed
-
Speech-to-SQL: Towards Speech-driven SQL Query Generation From Natural Language Question4 Jan 2022 0 repositories listed
-
Tencent-MVSE: A Large-Scale Benchmark Dataset for Multi-Modal Video Similarity Evaluation1 Jan 2022 0 repositories listed
-
Multi-Dialect Arabic Speech Recognition25 Dec 2021 0 repositories listed
-
Multi-Variant Consistency based Self-supervised Learning for Robust Automatic Speech Recognition23 Dec 2021 0 repositories listed
-
Voice Quality and Pitch Features in Transformer-Based Speech Recognition21 Dec 2021 0 repositories listed
-
Integrating Knowledge in End-to-End Automatic Speech Recognition for Mandarin-English Code-Switching19 Dec 2021 0 repositories listed
-
Multi-turn RNN-T for streaming recognition of multi-party speech19 Dec 2021 0 repositories listed
-
Prompt Tuning GPT-2 language model for parameter-efficient domain adaptation of ASR systems16 Dec 2021 0 repositories listed
-
Improving Hybrid CTC/Attention End-to-end Speech Recognition with Pretrained Acoustic and Language Model14 Dec 2021 0 repositories listed
-
Real-Time Neural Voice Camouflage14 Dec 2021 0 repositories listed
-
Robustifying automatic speech recognition by extracting slowly varying features14 Dec 2021 0 repositories listed