Browse State-of-the-Art › speech-recognition › Papers, page 32
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 32 of 58: papers 3,101 to 3,200 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
On monoaural speech enhancement for automatic recognition of real noisy speech using mixture invariant training3 May 2022 0 repositories listed
-
A Meeting Transcription System for an Ad-Hoc Acoustic Sensor Network2 May 2022 0 repositories listed
-
A Novel Speech-Driven Lip-Sync Model with CNN and LSTM2 May 2022 0 repositories listed
-
Bilingual End-to-End ASR with Byte-Level Subwords1 May 2022 0 repositories listed
-
CMU’s IWSLT 2022 Dialect Speech Translation System1 May 2022 0 repositories listed
-
Corpus Development of Kiswahili Speech Recognition Test and Evaluation sets, Preemptively Mitigating Demographic Bias Through Collaboration with Linguists1 May 2022 0 repositories listed
-
Discourse on ASR Measurement: Introducing the ARPOCA Assessment Tool1 May 2022 0 repositories listed
-
Enhancing Documentation of Hupa with Automatic Speech Recognition1 May 2022 0 repositories listed
-
Findings of the Shared Task on Speech Recognition for Vulnerable Individuals in Tamil1 May 2022 0 repositories listed
-
Fine-tuning pre-trained models for Automatic Speech Recognition, experiments on a fieldwork corpus of Japhug (Trans-Himalayan family)1 May 2022 0 repositories listed
-
JHU IWSLT 2022 Dialect Speech Translation System Description1 May 2022 0 repositories listed
-
MTL-SLT: Multi-Task Learning for Spoken Language Tasks1 May 2022 0 repositories listed
-
Multimodal fusion via cortical network inspired losses1 May 2022 0 repositories listed
-
NVIDIA NeMo Offline Speech Translation Systems for IWSLT 20221 May 2022 0 repositories listed
-
Phoneme transcription of endangered languages: an evaluation of recent ASR architectures in the single speaker scenario1 May 2022 0 repositories listed
-
Self-supervised Semantic-driven Phoneme Discovery for Zero-resource Speech Recognition1 May 2022 0 repositories listed
-
SSNCSE_NLP@LT-EDI-ACL2022: Speech Recognition for Vulnerable Individuals in Tamil using pre-trained XLSR models1 May 2022 0 repositories listed
-
SUH_ASR@LT-EDI-ACL2022: Transformer based Approach for Speech Recognition for Vulnerable Individuals in Tamil1 May 2022 0 repositories listed
-
The HW-TSC’s Offline Speech Translation System for IWSLT 2022 Evaluation1 May 2022 0 repositories listed
-
The Xiaomi Text-to-Text Simultaneous Speech Translation System for IWSLT 20221 May 2022 0 repositories listed
-
Why does Self-Supervised Learning for Speech Recognition Benefit Speaker Recognition?27 Apr 2022 0 repositories listed
-
Mask scalar prediction for improving robust automatic speech recognition26 Apr 2022 0 repositories listed
-
Cleanformer: A multichannel array configuration-invariant neural enhancement frontend for ASR in smart speakers25 Apr 2022 0 repositories listed
-
Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications25 Apr 2022 0 repositories listed
-
Supervised Attention in Sequence-to-Sequence Models for Speech Recognition25 Apr 2022 0 repositories listed
-
Improved far-field speech recognition using Joint Variational Autoencoder24 Apr 2022 0 repositories listed
-
E2E Segmenter: Joint Segmenting and Decoding for Long-Form ASR22 Apr 2022 0 repositories listed
-
Efficient Training of Neural Transducer for Speech Recognition22 Apr 2022 0 repositories listed
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment22 Apr 2022 0 repositories listed
-
Robustness Testing of Data and Knowledge Driven Anomaly Detection in Cyber-Physical Systems20 Apr 2022 0 repositories listed
-
An Investigation of Monotonic Transducers for Large-Scale Automatic Speech Recognition19 Apr 2022 0 repositories listed
-
Blockwise Streaming Transformer for Spoken Language Understanding and Simultaneous Speech Translation19 Apr 2022 0 repositories listed
-
Disappeared Command: Spoofing Attack On Automatic Speech Recognition Systems with Sound Masking19 Apr 2022 0 repositories listed
-
Automated speech tools for helping communities process restricted-access corpora for language revival efforts15 Apr 2022 0 repositories listed
-
Lombard Effect for Bilingual Speakers in Cantonese and English: importance of spectro-temporal features14 Apr 2022 0 repositories listed
-
A Unified Cascaded Encoder ASR Model for Dynamic Model Sizes13 Apr 2022 0 repositories listed
-
Self-critical Sequence Training for Automatic Speech Recognition13 Apr 2022 0 repositories listed
-
Study of Indian English Pronunciation Variabilities relative to Received Pronunciation13 Apr 2022 0 repositories listed
-
ASR in German: A Detailed Error Analysis12 Apr 2022 0 repositories listed
-
CorrectSpeech: A Fully Automated System for Speech Correction and Accent Reduction12 Apr 2022 0 repositories listed
-
Building an ASR Error Robust Spoken Virtual Patient System in a Highly Class-Imbalanced Scenario Without Speech Data11 Apr 2022 0 repositories listed
-
Multistream neural architectures for cued-speech recognition using a pre-trained visual feature extractor and constrained CTC decoding11 Apr 2022 0 repositories listed
-
Unified Speech-Text Pre-training for Speech Translation and Recognition11 Apr 2022 0 repositories listed
-
Deep Embeddings for Robust User-Based Amateur Vocal Percussion Classification10 Apr 2022 0 repositories listed
-
Adding Connectionist Temporal Summarization into Conformer to Improve Its Decoder Efficiency For Speech Recognition8 Apr 2022 0 repositories listed
-
Auditory-Based Data Augmentation for End-to-End Automatic Speech Recognition8 Apr 2022 0 repositories listed
-
Defense against Adversarial Attacks on Hybrid Speech Recognition using Joint Adversarial Fine-tuning with Denoiser8 Apr 2022 0 repositories listed
-
Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition8 Apr 2022 0 repositories listed
-
Detecting Dysfluencies in Stuttering Therapy Using wav2vec 2.07 Apr 2022 0 repositories listed
-
Enabling All In-Edge Deep Learning: A Literature Review7 Apr 2022 0 repositories listed
-
MAESTRO: Matched Speech Text Representations through Modality Matching7 Apr 2022 0 repositories listed
-
Three-Module Modeling For End-to-End Spoken Language Understanding Using Pre-trained DNN-HMM-Based Acoustic-Phonetic Model7 Apr 2022 0 repositories listed
-
A survey on recently proposed activation functions for Deep Learning6 Apr 2022 0 repositories listed
-
A Wav2vec2-Based Experimental Study on Self-Supervised Learning Methods to Improve Child Speech Recognition6 Apr 2022 0 repositories listed
-
Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation6 Apr 2022 0 repositories listed
-
Simple and Effective Unsupervised Speech Synthesis6 Apr 2022 0 repositories listed
-
Successes and critical failures of neural networks in capturing human-like speech recognition6 Apr 2022 0 repositories listed
-
A Complementary Joint Training Approach Using Unpaired Speech and Text for Low-Resource Automatic Speech Recognition5 Apr 2022 0 repositories listed
-
Audio-visual multi-channel speech separation, dereverberation and recognition5 Apr 2022 0 repositories listed
-
Disentangled Speech Representation Learning Based on Factorized Hierarchical Variational Autoencoder with Self-Supervised Objective5 Apr 2022 0 repositories listed
-
Hear No Evil: Towards Adversarial Robustness of Automatic Speech Recognition via Multi-Task Learning5 Apr 2022 0 repositories listed
-
Unsupervised Data Selection via Discrete Speech Representation for ASR5 Apr 2022 0 repositories listed
-
A Study of Gender Impact in Self-supervised Models for Speech-to-Text Systems4 Apr 2022 0 repositories listed
-
An Analysis of Semantically-Aligned Speech-Text Embeddings4 Apr 2022 0 repositories listed
-
Cross-lingual Self-Supervised Speech Representations for Improved Dysarthric Speech Recognition4 Apr 2022 0 repositories listed
-
Deliberation Model for On-Device Spoken Language Understanding4 Apr 2022 0 repositories listed
-
Self-Supervised Speech Representations Preserve Speech Characteristics while Anonymizing Voices4 Apr 2022 0 repositories listed
-
Deep Speech Based End-to-End Automated Speech Recognition (ASR) for Indian-English Accents3 Apr 2022 0 repositories listed
-
End-to-end model for named entity recognition from speech without paired training data2 Apr 2022 0 repositories listed
-
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation2 Apr 2022 0 repositories listed
-
Speaker adaptation for Wav2vec2 based dysarthric ASR2 Apr 2022 0 repositories listed
-
End-to-End Integration of Speech Recognition, Speech Enhancement, and Self-Supervised Learning Representation1 Apr 2022 0 repositories listed
-
End-to-End Multi-speaker ASR with Independent Vector Analysis1 Apr 2022 0 repositories listed
-
End-to-end multi-talker audio-visual ASR using an active speaker attention module1 Apr 2022 0 repositories listed
-
Filter-based Discriminative Autoencoders for Children Speech Recognition1 Apr 2022 0 repositories listed
-
InterAug: Augmenting Noisy Intermediate Predictions for CTC-based ASR1 Apr 2022 0 repositories listed
-
Alternate Intermediate Conditioning with Syllable-level and Character-level Targets for Japanese ASR1 Apr 2022 0 repositories listed
-
Multi-task RNN-T with Semantic Decoder for Streamable Spoken Language Understanding1 Apr 2022 0 repositories listed
-
Multiple Confidence Gates For Joint Training Of SE And ASR1 Apr 2022 0 repositories listed
-
Probing Speech Emotion Recognition Transformers for Linguistic Knowledge1 Apr 2022 0 repositories listed
-
Text-To-Speech Data Augmentation for Low Resource Speech Recognition1 Apr 2022 0 repositories listed
-
Zero-Shot Cross-lingual Aphasia Detection using Automatic Speech Recognition1 Apr 2022 0 repositories listed
-
A Comparative Study on Speaker-attributed Automatic Speech Recognition in Multi-party Meetings31 Mar 2022 0 repositories listed
-
An Empirical Study of Language Model Integration for Transducer based Speech Recognition31 Mar 2022 0 repositories listed
-
Analyzing the factors affecting usefulness of Self-Supervised Pre-trained Representations for Speech Recognition31 Mar 2022 0 repositories listed
-
Effectiveness of text to speech pseudo labels for forced alignment and cross lingual pretrained models for low resource speech recognition31 Mar 2022 0 repositories listed
-
Exploiting Single-Channel Speech for Multi-Channel End-to-End Speech Recognition: A Comparative Study31 Mar 2022 0 repositories listed
-
Importance of Different Temporal Modulations of Speech: A Tale of Two Perspectives31 Mar 2022 0 repositories listed
-
Improving Language Identification of Accented Speech31 Mar 2022 0 repositories listed
-
Memory-Efficient Training of RNN-Transducer with Sampled Softmax31 Mar 2022 0 repositories listed
-
31 Mar 2022 0 repositories listed
-
Code Switched and Code Mixed Speech Recognition for Indic languages30 Mar 2022 0 repositories listed
-
Improving Speech Recognition for Indic Languages using Language Model30 Mar 2022 0 repositories listed
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages?30 Mar 2022 0 repositories listed
-
Dynamic Latency for CTC-Based Streaming Automatic Speech Recognition With Emformer29 Mar 2022 0 repositories listed
-
Frequency-Directional Attention Model for Multilingual Automatic Speech Recognition29 Mar 2022 0 repositories listed
-
Improving Generalization of Deep Neural Network Acoustic Models with Length Perturbation and N-best Based Label Smoothing29 Mar 2022 0 repositories listed
-
Mel Frequency Spectral Domain Defenses against Adversarial Attacks on Speech Recognition Systems29 Mar 2022 0 repositories listed
-
Noise-robust Speech Recognition with 10 Minutes Unparalleled In-domain Data29 Mar 2022 0 repositories listed
-
Short-Term Word-Learning in a Dynamically Changing Environment29 Mar 2022 0 repositories listed