Browse State-of-the-Art › Automatic Speech Recognition (ASR) › Papers, page 18
Automatic Speech Recognition (ASR)
Papers archive 2025-07-28
archive papers tagged: 3,012 · with a code link: 622 · where Syntology ran a sample: 77 (64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (77 of 3,012 tagged: 64 with a run with no instrument failure, 13 where every run was a failure of Syntology's instrument)
Page 18 of 31: papers 1,701 to 1,800 of 3,012, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Polyphonic pitch detection with convolutional recurrent neural networks4 Feb 2022 0 repositories listed
-
The CUHK-TENCENT speaker diarization system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge4 Feb 2022 0 repositories listed
-
Joint Speech Recognition and Audio Captioning3 Feb 2022 0 repositories listed
-
The RoyalFlush System of Speech Recognition for M2MeT Challenge3 Feb 2022 0 repositories listed
-
ASR-Aware End-to-end Neural Diarization2 Feb 2022 0 repositories listed
-
Error Correction in ASR using Sequence-to-Sequence Models2 Feb 2022 0 repositories listed
-
RescoreBERT: Discriminative Speech Recognition Rescoring with BERT2 Feb 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Feb 2022 0 repositories listed
-
Language Dependencies in Adversarial Attacks on Speech Recognition Systems1 Feb 2022 0 repositories listed
-
Visualizing Automatic Speech Recognition -- Means for a Better Understanding?1 Feb 2022 0 repositories listed
-
Reducing language context confusion for end-to-end code-switching automatic speech recognition28 Jan 2022 0 repositories listed
-
Sentiment-Aware Automatic Speech Recognition pre-training for enhanced Speech Emotion Recognition27 Jan 2022 0 repositories listed
-
Synthesizing Dysarthric Speech Using Multi-talker TTS for Dysarthric Speech Recognition27 Jan 2022 0 repositories listed
-
On the Effectiveness of Pinyin-Character Dual-Decoding for End-to-End Mandarin Chinese ASR26 Jan 2022 0 repositories listed
-
26 Jan 2022 0 repositories listed
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models25 Jan 2022 0 repositories listed
-
Improving the fusion of acoustic and text representations in RNN-T25 Jan 2022 0 repositories listed
-
Run-and-back stitch search: novel block synchronous decoding for streaming encoder-decoder ASR25 Jan 2022 0 repositories listed
-
Transformer-Based Video Front-Ends for Audio-Visual Speech Recognition for Single and Multi-Person Video25 Jan 2022 0 repositories listed
-
A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition22 Jan 2022 0 repositories listed
-
How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR18 Jan 2022 0 repositories listed
-
Human and Automatic Speech Recognition Performance on German Oral History Interviews18 Jan 2022 0 repositories listed
-
DUAL: Textless Spoken Question Answering with Speech Discrete Unit Adaptive Learning16 Jan 2022 0 repositories listed
-
RED-ACE: Robust Error Detection for ASR using Confidence Embeddings16 Jan 2022 0 repositories listed
-
Recent Progress in the CUHK Dysarthric Speech Recognition System15 Jan 2022 0 repositories listed
-
Learning to Enhance or Not: Neural Network-Based Switching of Enhanced and Observed Signals for Overlapping Speech Recognition11 Jan 2022 0 repositories listed
-
A Likelihood Ratio based Domain Adaptation Method for E2E Models10 Jan 2022 0 repositories listed
-
Cross-Modal ASR Post-Processing System for Error Correction and Utterance Rejection10 Jan 2022 0 repositories listed
-
Speech-to-SQL: Towards Speech-driven SQL Query Generation From Natural Language Question4 Jan 2022 0 repositories listed
-
Tencent-MVSE: A Large-Scale Benchmark Dataset for Multi-Modal Video Similarity Evaluation1 Jan 2022 0 repositories listed
-
Multi-Dialect Arabic Speech Recognition25 Dec 2021 0 repositories listed
-
Multi-Variant Consistency based Self-supervised Learning for Robust Automatic Speech Recognition23 Dec 2021 0 repositories listed
-
Voice Quality and Pitch Features in Transformer-Based Speech Recognition21 Dec 2021 0 repositories listed
-
Integrating Knowledge in End-to-End Automatic Speech Recognition for Mandarin-English Code-Switching19 Dec 2021 0 repositories listed
-
Multi-turn RNN-T for streaming recognition of multi-party speech19 Dec 2021 0 repositories listed
-
Prompt Tuning GPT-2 language model for parameter-efficient domain adaptation of ASR systems16 Dec 2021 0 repositories listed
-
Improving Hybrid CTC/Attention End-to-end Speech Recognition with Pretrained Acoustic and Language Model14 Dec 2021 0 repositories listed
-
Real-Time Neural Voice Camouflage14 Dec 2021 0 repositories listed
-
Robustifying automatic speech recognition by extracting slowly varying features14 Dec 2021 0 repositories listed
-
PM-MMUT: Boosted Phone-Mask Data Augmentation using Multi-Modeling Unit Training for Phonetic-Reduction-Robust E2E Speech Recognition13 Dec 2021 0 repositories listed
-
Improving Code-switching Language Modeling with Artificially Generated Texts using Cycle-consistent Adversarial Networks12 Dec 2021 0 repositories listed
-
Improving Speech Recognition on Noisy Speech via Speech Enhancement with Multi-Discriminators CycleGAN12 Dec 2021 0 repositories listed
-
Building a great multi-lingual teacher with sparsely-gated mixture of experts for speech recognition10 Dec 2021 0 repositories listed
-
Directed Speech Separation for Automatic Speech Recognition of Long Form Conversational Speech10 Dec 2021 0 repositories listed
-
Revisiting the Boundary between ASR and NLU in the Age of Conversational Dialog Systems10 Dec 2021 0 repositories listed
-
Sequence-level self-learning with multiple hypotheses10 Dec 2021 0 repositories listed
-
A study on native American English speech recognition by Indian listeners with varying word familiarity level8 Dec 2021 0 repositories listed
-
BBS-KWS:The Mandarin Keyword Spotting System Won the Video Keyword Wakeup Challenge3 Dec 2021 0 repositories listed
-
Catch Me If You Can: Blackbox Adversarial Attacks on Automatic Speech Recognition using Frequency Masking3 Dec 2021 0 repositories listed
-
A higher order Minkowski loss for improved prediction ability of acoustic model in ASR2 Dec 2021 0 repositories listed
-
A Mixture of Expert Based Deep Neural Network for Improved ASR2 Dec 2021 0 repositories listed
-
Loss Landscape Dependent Self-Adjusting Learning Rates in Decentralized Stochastic Gradient Descent2 Dec 2021 0 repositories listed
-
An Experiment on Speech-to-Text Translation Systems for Manipuri to English on Low Resource Setting1 Dec 2021 0 repositories listed
-
An Investigation of Hybrid architectures for Low Resource Multilingual Speech Recognition system in Indian context1 Dec 2021 0 repositories listed
-
Analysis of Manipuri Tones in ManiTo: A Tonal Contrast Database1 Dec 2021 0 repositories listed
-
IE-CPS Lexicon: An Automatic Speech Recognition Oriented Indian-English Pronunciation Dictionary1 Dec 2021 0 repositories listed
-
Improve Sinhala Speech Recognition Through e2e LF-MMI Model1 Dec 2021 0 repositories listed
-
Predicting lexical skills from oral reading with acoustic measures1 Dec 2021 0 repositories listed
-
Speech-T: Transducer for Text to Speech and Beyond1 Dec 2021 0 repositories listed
-
29 Nov 2021 0 repositories listed
-
Effect of noise suppression losses on speech distortion and ASR performance23 Nov 2021 0 repositories listed
-
Multi-Channel Multi-Speaker ASR Using 3D Spatial Feature22 Nov 2021 0 repositories listed
-
Capitalization and Punctuation Restoration: a Survey21 Nov 2021 0 repositories listed
-
Deep Spoken Keyword Spotting: An Overview20 Nov 2021 0 repositories listed
-
Switching Independent Vector Analysis and Its Extension to Blind and Spatially Guided Convolutional Beamforming Algorithms20 Nov 2021 0 repositories listed
-
Lattention: Lattice-attention in ASR rescoring19 Nov 2021 0 repositories listed
-
A Conformer-based ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement and Speech Separation18 Nov 2021 0 repositories listed
-
Towards Measuring Fairness in Speech Recognition: Casual Conversations Dataset Transcriptions18 Nov 2021 0 repositories listed
-
A Novel End-to-End CAPT System for L2 Children Learners16 Nov 2021 0 repositories listed
-
Heterogeneous Language Model Optimization in Automatic Speech Recognition16 Nov 2021 0 repositories listed
-
Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations16 Nov 2021 0 repositories listed
-
On Spoken Language Understanding Systems for Low Resourced Languages16 Nov 2021 0 repositories listed
-
Progressive Down-Sampling for Acoustic Encoding16 Nov 2021 0 repositories listed
-
Speech-to-SQL Parsing: Error Correction with Multi-modal Representations16 Nov 2021 0 repositories listed
-
Two Front-Ends, One Model : Fusing Heterogeneous Speech Features for Low Resource ASR with Multilingual Pre-Training16 Nov 2021 0 repositories listed
-
Who Are We Talking About? Handling Person Names in Speech Translation16 Nov 2021 0 repositories listed
-
Attention based end to end Speech Recognition for Voice Search in Hindi and English15 Nov 2021 0 repositories listed
-
Prediction of Listener Perception of Argumentative Speech in a Crowdsourced Dataset Using (Psycho-)Linguistic and Fluency Features13 Nov 2021 0 repositories listed
-
Self-Normalized Importance Sampling for Neural Language Modeling11 Nov 2021 0 repositories listed
-
Scaling ASR Improves Zero and Few Shot Learning10 Nov 2021 0 repositories listed
-
Privacy attacks for automatic speech recognition acoustic models in a federated learning framework6 Nov 2021 0 repositories listed
-
Conformer-based Hybrid ASR System for Switchboard Dataset5 Nov 2021 0 repositories listed
-
Context-Aware Transformer Transducer for Speech Recognition5 Nov 2021 0 repositories listed
-
Effective Cross-Utterance Language Modeling for Conversational Speech Recognition5 Nov 2021 0 repositories listed
-
4 Nov 2021 0 repositories listed
-
Speech recognition for air traffic control via feature learning and end-to-end training4 Nov 2021 0 repositories listed
-
STC speaker recognition systems for the NIST SRE 20213 Nov 2021 0 repositories listed
-
Recent Advances in End-to-End Automatic Speech Recognition2 Nov 2021 0 repositories listed
-
Collaborative Data Relabeling for Robust and Diverse Voice Apps Recommendation in Intelligent Personal Assistants1 Nov 2021 0 repositories listed
-
Comprehensive Punctuation Restoration for English and Polish1 Nov 2021 0 repositories listed
-
Indic Languages Automatic Speech Recognition using Meta-Learning Approach1 Nov 2021 0 repositories listed
-
Sequence Transduction with Graph-based Supervision1 Nov 2021 0 repositories listed
-
SNRi Target Training for Joint Speech Enhancement and Recognition1 Nov 2021 0 repositories listed
-
Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English1 Nov 2021 0 repositories listed
-
Voice Query Auto Completion1 Nov 2021 0 repositories listed
-
Cross-attention conformer for context modeling in speech enhancement for ASR30 Oct 2021 0 repositories listed
-
Speaker conditioning of acoustic models using affine transformation for multi-speaker speech recognition30 Oct 2021 0 repositories listed
-
Fusing ASR Outputs in Joint Training for Speech Emotion Recognition29 Oct 2021 0 repositories listed
-
Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction28 Oct 2021 0 repositories listed
-
Asynchronous Decentralized Distributed Training of Acoustic Models21 Oct 2021 0 repositories listed