Browse State-of-the-Art › speech-recognition › Papers, page 34
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 34 of 58: papers 3,301 to 3,400 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Likelihood Ratio based Domain Adaptation Method for E2E Models10 Jan 2022 0 repositories listed
-
Cross-Modal ASR Post-Processing System for Error Correction and Utterance Rejection10 Jan 2022 0 repositories listed
-
Two-Pass End-to-End ASR Model Compression8 Jan 2022 0 repositories listed
-
Textual Data Augmentation for Arabic-English Code-Switching Speech Recognition7 Jan 2022 0 repositories listed
-
Speech-to-SQL: Towards Speech-driven SQL Query Generation From Natural Language Question4 Jan 2022 0 repositories listed
-
Robust Natural Language Processing: Recent Advances, Challenges, and Future Directions3 Jan 2022 0 repositories listed
-
Tencent-MVSE: A Large-Scale Benchmark Dataset for Multi-Modal Video Similarity Evaluation1 Jan 2022 0 repositories listed
-
Temporal Attention Augmented Transformer Hawkes Process29 Dec 2021 0 repositories listed
-
Bridging the Gap: Using Deep Acoustic Representations to Learn Grounded Language from Percepts and Raw Speech27 Dec 2021 0 repositories listed
-
Multi-Dialect Arabic Speech Recognition25 Dec 2021 0 repositories listed
-
Multi-Variant Consistency based Self-supervised Learning for Robust Automatic Speech Recognition23 Dec 2021 0 repositories listed
-
TOD-DA: Towards Boosting the Robustness of Task-oriented Dialogue Modeling on Spoken Conversations23 Dec 2021 0 repositories listed
-
VoiceMoji: A Novel On-Device Pipeline for Seamless Emoji Insertion in Dictation22 Dec 2021 0 repositories listed
-
Voice Quality and Pitch Features in Transformer-Based Speech Recognition21 Dec 2021 0 repositories listed
-
Load-balanced Gather-scatter Patterns for Sparse Deep Neural Networks20 Dec 2021 0 repositories listed
-
Integrating Knowledge in End-to-End Automatic Speech Recognition for Mandarin-English Code-Switching19 Dec 2021 0 repositories listed
-
Investigation of Densely Connected Convolutional Networks with Domain Adversarial Learning for Noise Robust Speech Recognition19 Dec 2021 0 repositories listed
-
Multi-turn RNN-T for streaming recognition of multi-party speech19 Dec 2021 0 repositories listed
-
A singular Riemannian geometry approach to Deep Neural Networks I. Theoretical foundations17 Dec 2021 0 repositories listed
-
Prompt Tuning GPT-2 language model for parameter-efficient domain adaptation of ASR systems16 Dec 2021 0 repositories listed
-
Improving Hybrid CTC/Attention End-to-end Speech Recognition with Pretrained Acoustic and Language Model14 Dec 2021 0 repositories listed
-
Real-Time Neural Voice Camouflage14 Dec 2021 0 repositories listed
-
Robustifying automatic speech recognition by extracting slowly varying features14 Dec 2021 0 repositories listed
-
PM-MMUT: Boosted Phone-Mask Data Augmentation using Multi-Modeling Unit Training for Phonetic-Reduction-Robust E2E Speech Recognition13 Dec 2021 0 repositories listed
-
Improving Code-switching Language Modeling with Artificially Generated Texts using Cycle-consistent Adversarial Networks12 Dec 2021 0 repositories listed
-
Improving Speech Recognition on Noisy Speech via Speech Enhancement with Multi-Discriminators CycleGAN12 Dec 2021 0 repositories listed
-
Building a great multi-lingual teacher with sparsely-gated mixture of experts for speech recognition10 Dec 2021 0 repositories listed
-
Directed Speech Separation for Automatic Speech Recognition of Long Form Conversational Speech10 Dec 2021 0 repositories listed
-
Revisiting the Boundary between ASR and NLU in the Age of Conversational Dialog Systems10 Dec 2021 0 repositories listed
-
Sequence-level self-learning with multiple hypotheses10 Dec 2021 0 repositories listed
-
Are E2E ASR models ready for an industrial usage?9 Dec 2021 0 repositories listed
-
LipSound2: Self-Supervised Pre-Training for Lip-to-Speech Reconstruction and Lip Reading9 Dec 2021 0 repositories listed
-
A study on native American English speech recognition by Indian listeners with varying word familiarity level8 Dec 2021 0 repositories listed
-
BBS-KWS:The Mandarin Keyword Spotting System Won the Video Keyword Wakeup Challenge3 Dec 2021 0 repositories listed
-
Catch Me If You Can: Blackbox Adversarial Attacks on Automatic Speech Recognition using Frequency Masking3 Dec 2021 0 repositories listed
-
A higher order Minkowski loss for improved prediction ability of acoustic model in ASR2 Dec 2021 0 repositories listed
-
A Mixture of Expert Based Deep Neural Network for Improved ASR2 Dec 2021 0 repositories listed
-
Loss Landscape Dependent Self-Adjusting Learning Rates in Decentralized Stochastic Gradient Descent2 Dec 2021 0 repositories listed
-
An End-to-End Speech Recognition for the Nepali Language1 Dec 2021 0 repositories listed
-
An Experiment on Speech-to-Text Translation Systems for Manipuri to English on Low Resource Setting1 Dec 2021 0 repositories listed
-
An Investigation of Hybrid architectures for Low Resource Multilingual Speech Recognition system in Indian context1 Dec 2021 0 repositories listed
-
Analysis of Manipuri Tones in ManiTo: A Tonal Contrast Database1 Dec 2021 0 repositories listed
-
IE-CPS Lexicon: An Automatic Speech Recognition Oriented Indian-English Pronunciation Dictionary1 Dec 2021 0 repositories listed
-
Impact of Microphone position Measurement Error on Multi Channel Distant Speech Recognition & Intelligibility1 Dec 2021 0 repositories listed
-
Improve Sinhala Speech Recognition Through e2e LF-MMI Model1 Dec 2021 0 repositories listed
-
Phone Based Keyword Spotting for Transcribing Very Low Resource Languages1 Dec 2021 0 repositories listed
-
Predicting lexical skills from oral reading with acoustic measures1 Dec 2021 0 repositories listed
-
Speech-T: Transducer for Text to Speech and Beyond1 Dec 2021 0 repositories listed
-
29 Nov 2021 0 repositories listed
-
Joint Modeling of Code-Switched and Monolingual ASR via Conditional Factorization29 Nov 2021 0 repositories listed
-
Mixed Precision Low-bit Quantization of Neural Network Language Models for Speech Recognition29 Nov 2021 0 repositories listed
-
Mixed Precision of Quantization of Transformer Language Models for Speech Recognition29 Nov 2021 0 repositories listed
-
Effect of noise suppression losses on speech distortion and ASR performance23 Nov 2021 0 repositories listed
-
Guided-TTS: A Diffusion Model for Text-to-Speech via Classifier Guidance23 Nov 2021 0 repositories listed
-
SpeechMoE2: Mixture-of-Experts Model with Improved Routing23 Nov 2021 0 repositories listed
-
Multi-Channel Multi-Speaker ASR Using 3D Spatial Feature22 Nov 2021 0 repositories listed
-
Capitalization and Punctuation Restoration: a Survey21 Nov 2021 0 repositories listed
-
Deep Spoken Keyword Spotting: An Overview20 Nov 2021 0 repositories listed
-
Switching Independent Vector Analysis and Its Extension to Blind and Spatially Guided Convolutional Beamforming Algorithms20 Nov 2021 0 repositories listed
-
A comparison of streaming models and data augmentation methods for robust speech recognition19 Nov 2021 0 repositories listed
-
Lattention: Lattice-attention in ASR rescoring19 Nov 2021 0 repositories listed
-
Semi-supervised transfer learning for language expansion of end-to-end speech recognition models to low-resource languages19 Nov 2021 0 repositories listed
-
A Conformer-based ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement and Speech Separation18 Nov 2021 0 repositories listed
-
Towards Measuring Fairness in Speech Recognition: Casual Conversations Dataset Transcriptions18 Nov 2021 0 repositories listed
-
Subject Enveloped Deep Sample Fuzzy Ensemble Learning Algorithm of Parkinson's Speech Data17 Nov 2021 0 repositories listed
-
17 Nov 2021 0 repositories listed
-
A Novel End-to-End CAPT System for L2 Children Learners16 Nov 2021 0 repositories listed
-
Heterogeneous Language Model Optimization in Automatic Speech Recognition16 Nov 2021 0 repositories listed
-
Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations16 Nov 2021 0 repositories listed
-
Leveraging Uni-Modal Self-Supervised Learning for Multimodal Audio-visual Speech Recognition16 Nov 2021 0 repositories listed
-
Modeling speech recognition and synthesis simultaneously: Encoding and decoding lexical and sublexical semantic information into speech with no access to speech data16 Nov 2021 0 repositories listed
-
On Spoken Language Understanding Systems for Low Resourced Languages16 Nov 2021 0 repositories listed
-
Progressive Down-Sampling for Acoustic Encoding16 Nov 2021 0 repositories listed
-
Speech Synthesis for Low Resource Languages using Transliteration Enabled Transfer Learning16 Nov 2021 0 repositories listed
-
Speech-to-SQL Parsing: Error Correction with Multi-modal Representations16 Nov 2021 0 repositories listed
-
Two Front-Ends, One Model : Fusing Heterogeneous Speech Features for Low Resource ASR with Multilingual Pre-Training16 Nov 2021 0 repositories listed
-
Unified Speech-Text Pre-training for Speech Translation and Recognition16 Nov 2021 0 repositories listed
-
Unsupervised Speech Enhancement with speech recognition embedding and disentanglement losses16 Nov 2021 0 repositories listed
-
Who Are We Talking About? Handling Person Names in Speech Translation16 Nov 2021 0 repositories listed
-
Attention based end to end Speech Recognition for Voice Search in Hindi and English15 Nov 2021 0 repositories listed
-
Analysis of Data Augmentation Methods for Low-Resource Maltese ASR15 Nov 2021 0 repositories listed
-
Joint Unsupervised and Supervised Training for Multilingual ASR15 Nov 2021 0 repositories listed
-
Binary classification of spoken words with passive phononic metamaterials14 Nov 2021 0 repositories listed
-
Prediction of Listener Perception of Argumentative Speech in a Crowdsourced Dataset Using (Psycho-)Linguistic and Fluency Features13 Nov 2021 0 repositories listed
-
A Convolutional Neural Network Based Approach to Recognize Bangla Spoken Digits from Speech Signal12 Nov 2021 0 repositories listed
-
Can neural networks predict dynamics they have never seen?12 Nov 2021 0 repositories listed
-
Self-Normalized Importance Sampling for Neural Language Modeling11 Nov 2021 0 repositories listed
-
Scaling ASR Improves Zero and Few Shot Learning10 Nov 2021 0 repositories listed
-
Joint Neural AEC and Beamforming with Double-Talk Detection9 Nov 2021 0 repositories listed
-
Retrieving Speaker Information from Personalized Acoustic Models for Speech Recognition7 Nov 2021 0 repositories listed
-
Privacy attacks for automatic speech recognition acoustic models in a federated learning framework6 Nov 2021 0 repositories listed
-
Conformer-based Hybrid ASR System for Switchboard Dataset5 Nov 2021 0 repositories listed
-
Context-Aware Transformer Transducer for Speech Recognition5 Nov 2021 0 repositories listed
-
Effective Cross-Utterance Language Modeling for Conversational Speech Recognition5 Nov 2021 0 repositories listed
-
Learning of Time-Frequency Attention Mechanism for Automatic Modulation Recognition5 Nov 2021 0 repositories listed
-
Oracle Teacher: Leveraging Target Information for Better Knowledge Distillation of CTC Models5 Nov 2021 0 repositories listed
-
4 Nov 2021 0 repositories listed
-
Speech recognition for air traffic control via feature learning and end-to-end training4 Nov 2021 0 repositories listed
-
Voice Conversion Can Improve ASR in Very Low-Resource Settings4 Nov 2021 0 repositories listed
-
STC speaker recognition systems for the NIST SRE 20213 Nov 2021 0 repositories listed