Browse State-of-the-Art › speech-recognition › Papers, page 33
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 33 of 58: papers 3,201 to 3,300 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Streaming parallel transducer beam search with fast-slow cascaded encoders29 Mar 2022 0 repositories listed
-
On-the-Fly Feature Based Rapid Speaker Adaptation for Dysarthric and Elderly Speech Recognition28 Mar 2022 0 repositories listed
-
Chain-based Discriminative Autoencoders for Speech Recognition25 Mar 2022 0 repositories listed
-
Impact of Dataset on Acoustic Models for Automatic Speech Recognition25 Mar 2022 0 repositories listed
-
Computing Optimal Location of Microphone for Improved Speech Recognition24 Mar 2022 0 repositories listed
-
Disentangleing Content and Fine-grained Prosody Information via Hybrid ASR Bottleneck Features for Voice Conversion24 Mar 2022 0 repositories listed
-
Lahjoita puhetta -- a large-scale corpus of spoken Finnish with some benchmarks24 Mar 2022 0 repositories listed
-
A Text-to-Speech Pipeline, Evaluation Methodology, and Initial Fine-Tuning Results for Child Speech Synthesis22 Mar 2022 0 repositories listed
-
Building Robust Spoken Language Understanding by Cross Attention between Phoneme Sequence and ASR Hypothesis22 Mar 2022 0 repositories listed
-
Modeling speech recognition and synthesis simultaneously: Encoding and decoding lexical and sublexical semantic information into speech with no direct access to speech data22 Mar 2022 0 repositories listed
-
Pseudo Label Is Better Than Human Label22 Mar 2022 0 repositories listed
-
Enhancing Speech Recognition Decoding via Layer Aggregation21 Mar 2022 0 repositories listed
-
XTREME-S: Evaluating Cross-lingual Speech Representations21 Mar 2022 0 repositories listed
-
Exploiting Cross Domain Acoustic-to-articulatory Inverted Features For Disordered Speech Recognition19 Mar 2022 0 repositories listed
-
Similarity and Content-based Phonetic Self Attention for Speech Recognition19 Mar 2022 0 repositories listed
-
Representative Subset Selection for Efficient Fine-Tuning in Self-Supervised Speech Recognition18 Mar 2022 0 repositories listed
-
Prediction of speech intelligibility with DNN-based performance measures17 Mar 2022 0 repositories listed
-
Whither the Priors for (Vocal) Interactivity?16 Mar 2022 0 repositories listed
-
Privacy-Preserving Speech Representation Learning using Vector Quantization15 Mar 2022 0 repositories listed
-
Interpretable Dysarthric Speaker Adaptation based on Optimal-Transport14 Mar 2022 0 repositories listed
-
Spectral Modification Based Data Augmentation For Improving End-to-End ASR For Children's Speech13 Mar 2022 0 repositories listed
-
Transformer-based Streaming ASR with Cumulative Attention11 Mar 2022 0 repositories listed
-
Attacks as Defenses: Designing Robust Audio CAPTCHAs Using Attacks on Automatic Speech Recognition Systems10 Mar 2022 0 repositories listed
-
A practical framework for multi-domain speech recognition and an instance sampling method to neural language modeling9 Mar 2022 0 repositories listed
-
Language Adaptive Cross-lingual Speech Representation Learning with Sparse Sharing Sub-networks9 Mar 2022 0 repositories listed
-
Sentence-Select: Large-Scale Language Model Data Selection for Rare-Word Speech Recognition9 Mar 2022 0 repositories listed
-
Which French speech recognition system for assistant robots?4 Mar 2022 0 repositories listed
-
A Conformer Based Acoustic Model for Robust Automatic Speech Recognition1 Mar 2022 0 repositories listed
-
Extended Graph Temporal Classification for Multi-Speaker End-to-End ASR1 Mar 2022 0 repositories listed
-
Measuring the Impact of Individual Domain Factors in Self-Supervised Pre-Training1 Mar 2022 0 repositories listed
-
Integrating Text Inputs For Training and Adapting RNN Transducer ASR Models26 Feb 2022 0 repositories listed
-
A Survey of Multilingual Models for Automatic Speech Recognition25 Feb 2022 0 repositories listed
-
Language technology practitioners as language managers: arbitrating data bias and predictive bias in ASR25 Feb 2022 0 repositories listed
-
Ask2Mask: Guided Data Selection for Masked Speech Modeling24 Feb 2022 0 repositories listed
-
Closing the Gap between Single-User and Multi-User VoiceFilter-Lite24 Feb 2022 0 repositories listed
-
Towards Better Meta-Initialization with Task Augmentation for Kindergarten-aged Speech Recognition24 Feb 2022 0 repositories listed
-
Differentially Private Speaker Anonymization23 Feb 2022 0 repositories listed
-
Adversarial Attacks on Speech Recognition Systems for Mission-Critical Applications: A Survey22 Feb 2022 0 repositories listed
-
Korean Tokenization for Beam Search Rescoring in Speech Recognition22 Feb 2022 0 repositories listed
-
VADOI:Voice-Activity-Detection Overlapping Inference For End-to-end Long-form Speech Recognition22 Feb 2022 0 repositories listed
-
r-G2P: Evaluating and Enhancing Robustness of Grapheme to Phoneme Conversion by Controlled noise introducing and Contextual information incorporation21 Feb 2022 0 repositories listed
-
Speaker Adaptation Using Spectro-Temporal Deep Features for Dysarthric and Elderly Speech Recognition21 Feb 2022 0 repositories listed
-
The PCG-AIID System for L3DAS22 Challenge: MIMO and MISO convolutional recurrent Network for Multi Channel Speech Enhancement and Speech Recognition21 Feb 2022 0 repositories listed
-
19 Feb 2022 0 repositories listed
-
Domain Adaptation of low-resource Target-Domain models using well-trained ASR Conformer Models18 Feb 2022 0 repositories listed
-
End-to-end contextual asr based on posterior distribution adaptation for hybrid ctc/attention system18 Feb 2022 0 repositories listed
-
'Beach' to 'Bitch': Inadvertent Unsafe Transcription of Kids' Content on YouTube17 Feb 2022 0 repositories listed
-
Curriculum optimization for low-resource speech recognition17 Feb 2022 0 repositories listed
-
Mitigating Closed-model Adversarial Examples with Bayesian Neural Modeling for Enhanced End-to-End Speech Recognition17 Feb 2022 0 repositories listed
-
MLP-ASR: Sequence-length agnostic all-MLP architectures for speech recognition17 Feb 2022 0 repositories listed
-
Conversational Speech Recognition By Learning Conversation-level Characteristics16 Feb 2022 0 repositories listed
-
Knowledge Transfer from Large-scale Pretrained Language Models to End-to-end Speech Recognizers16 Feb 2022 0 repositories listed
-
Learning Contextually Fused Audio-visual Representations for Audio-visual Speech Recognition15 Feb 2022 0 repositories listed
-
Multi-style Training for South African Call Centre Audio15 Feb 2022 0 repositories listed
-
Saving RNN Computations with a Neuron-Level Fuzzy Memoization Scheme14 Feb 2022 0 repositories listed
-
Vau da muntanialas: Energy-efficient multi-die scalable acceleration of RNN inference14 Feb 2022 0 repositories listed
-
Multimodal Depression Classification Using Articulatory Coordination Features And Hierarchical Attention Based Text Embeddings13 Feb 2022 0 repositories listed
-
USTED: Improving ASR with a Unified Speech and Text Encoder-Decoder12 Feb 2022 0 repositories listed
-
Ultra-low Power Always-on Intelligent and Connected SNN-based System for Multimedia IoT-enabled Applications10 Feb 2022 0 repositories listed
-
The Volcspeech system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge9 Feb 2022 0 repositories listed
-
A two-step approach to leverage contextual data: speech recognition in air-traffic communications8 Feb 2022 0 repositories listed
-
Enhancing ASR for Stuttered Speech with Limited Data Using Detect and Pass8 Feb 2022 0 repositories listed
-
T-NGA: Temporal Network Grafting Algorithm for Learning to Process Spiking Audio Sensor Events7 Feb 2022 0 repositories listed
-
Polyphonic pitch detection with convolutional recurrent neural networks4 Feb 2022 0 repositories listed
-
The CUHK-TENCENT speaker diarization system for the ICASSP 2022 multi-channel multi-party meeting transcription challenge4 Feb 2022 0 repositories listed
-
Joint Speech Recognition and Audio Captioning3 Feb 2022 0 repositories listed
-
The RoyalFlush System of Speech Recognition for M2MeT Challenge3 Feb 2022 0 repositories listed
-
ASR-Aware End-to-end Neural Diarization2 Feb 2022 0 repositories listed
-
Error Correction in ASR using Sequence-to-Sequence Models2 Feb 2022 0 repositories listed
-
RescoreBERT: Discriminative Speech Recognition Rescoring with BERT2 Feb 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Feb 2022 0 repositories listed
-
Language Dependencies in Adversarial Attacks on Speech Recognition Systems1 Feb 2022 0 repositories listed
-
Visualizing Automatic Speech Recognition -- Means for a Better Understanding?1 Feb 2022 0 repositories listed
-
Improving End-to-End Models for Set Prediction in Spoken Language Understanding28 Jan 2022 0 repositories listed
-
Neural-FST Class Language Model for End-to-End Speech Recognition28 Jan 2022 0 repositories listed
-
Reducing language context confusion for end-to-end code-switching automatic speech recognition28 Jan 2022 0 repositories listed
-
Sentiment-Aware Automatic Speech Recognition pre-training for enhanced Speech Emotion Recognition27 Jan 2022 0 repositories listed
-
Synthesizing Dysarthric Speech Using Multi-talker TTS for Dysarthric Speech Recognition27 Jan 2022 0 repositories listed
-
On the Effectiveness of Pinyin-Character Dual-Decoding for End-to-End Mandarin Chinese ASR26 Jan 2022 0 repositories listed
-
26 Jan 2022 0 repositories listed
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models25 Jan 2022 0 repositories listed
-
Improving the fusion of acoustic and text representations in RNN-T25 Jan 2022 0 repositories listed
-
Run-and-back stitch search: novel block synchronous decoding for streaming encoder-decoder ASR25 Jan 2022 0 repositories listed
-
Transformer-Based Video Front-Ends for Audio-Visual Speech Recognition for Single and Multi-Person Video25 Jan 2022 0 repositories listed
-
Data and knowledge-driven approaches for multilingual training to improve the performance of speech recognition systems of Indian languages24 Jan 2022 0 repositories listed
-
Endpoint Detection for Streaming End-to-End Multi-talker ASR24 Jan 2022 0 repositories listed
-
Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition24 Jan 2022 0 repositories listed
-
PickNet: Real-Time Channel Selection for Ad Hoc Microphone Arrays24 Jan 2022 0 repositories listed
-
Variational Auto-Encoder Based Variability Encoding for Dysarthric Speech Recognition24 Jan 2022 0 repositories listed
-
A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition22 Jan 2022 0 repositories listed
-
Enabling Deep Learning on Edge Devices through Filter Pruning and Knowledge Transfer22 Jan 2022 0 repositories listed
-
How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR18 Jan 2022 0 repositories listed
-
Human and Automatic Speech Recognition Performance on German Oral History Interviews18 Jan 2022 0 repositories listed
-
DUAL: Textless Spoken Question Answering with Speech Discrete Unit Adaptive Learning16 Jan 2022 0 repositories listed
-
RED-ACE: Robust Error Detection for ASR using Confidence Embeddings16 Jan 2022 0 repositories listed
-
Recent Progress in the CUHK Dysarthric Speech Recognition System15 Jan 2022 0 repositories listed
-
Investigation of Data Augmentation Techniques for Disordered Speech Recognition14 Jan 2022 0 repositories listed
-
Spectro-Temporal Deep Features for Disordered Speech Assessment and Recognition14 Jan 2022 0 repositories listed
-
The Effectiveness of Time Stretching for Enhancing Dysarthric Speech for Improved Dysarthric Speech Recognition13 Jan 2022 0 repositories listed
-
Learning to Enhance or Not: Neural Network-Based Switching of Enhanced and Observed Signals for Overlapping Speech Recognition11 Jan 2022 0 repositories listed