Browse State-of-the-Art › speech-recognition › Papers, page 38
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 38 of 58: papers 3,701 to 3,800 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Leveraging Pre-trained Language Model for Speech Sentiment Analysis11 Jun 2021 0 repositories listed
-
Balanced End-to-End Monolingual pre-training for Low-Resourced Indic Languages Code-Switching Speech Recognition10 Jun 2021 0 repositories listed
-
PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition10 Jun 2021 0 repositories listed
-
10 Jun 2021 0 repositories listed
-
U2++: Unified Two-pass Bidirectional End-to-end Model for Speech Recognition10 Jun 2021 0 repositories listed
-
A Comparative Study on Neural Architectures and Training Methods for Japanese Speech Recognition9 Jun 2021 0 repositories listed
-
Unsupervised Automatic Speech Recognition: A Review9 Jun 2021 0 repositories listed
-
Raw Waveform Encoder with Multi-Scale Globally Attentive Locally Recurrent Networks for End-to-End Speech Recognition8 Jun 2021 0 repositories listed
-
8 Jun 2021 0 repositories listed
-
Data Augmentation Methods for End-to-end Speech Recognition on Distant-Talk Scenarios7 Jun 2021 0 repositories listed
-
Human Listening and Live Captioning: Multi-Task Training for Speech Enhancement5 Jun 2021 0 repositories listed
-
Do You Listen with One or Two Microphones? A Unified ASR Model for Single and Multi-Channel Audio4 Jun 2021 0 repositories listed
-
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition4 Jun 2021 0 repositories listed
-
A Discussion On the Validity of Manifold Learning3 Jun 2021 0 repositories listed
-
Semantic-WER: A Unified Metric for the Evaluation of ASR Transcript for End Usability3 Jun 2021 0 repositories listed
-
Dual Script E2E framework for Multilingual and Code-Switching ASR2 Jun 2021 0 repositories listed
-
Improving low-resource ASR performance with untranscribed out-of-domain data2 Jun 2021 0 repositories listed
-
Should We Always Separate?: Switching Between Enhanced and Observed Signals for Overlapping Speech Recognition2 Jun 2021 0 repositories listed
-
2020福爾摩沙臺語語音辨識比賽之初步實驗 (A Preliminary Study of Formosa Speech Recognition Challenge 2020 – Taiwanese ASR)1 Jun 2021 0 repositories listed
-
A Neural Acoustic Echo Canceller Optimized Using An Automatic Speech Recognizer And Large Scale Synthetic Data1 Jun 2021 0 repositories listed
-
Developing ASR for Indonesian-English Bilingual Language Teaching1 Jun 2021 0 repositories listed
-
End-to-end ASR to jointly predict transcriptions and linguistic annotations1 Jun 2021 0 repositories listed
-
End-to-End Automatic Speech Recognition: Its Impact on the Workflowin Documenting Yoloxóchitl Mixtec1 Jun 2021 0 repositories listed
-
Evaluating Automatic Speech Recognition Quality and Its Impact on Counselor Utterance Coding1 Jun 2021 0 repositories listed
-
Highland Puebla Nahuatl Speech Translation Corpus for Endangered Language Documentation1 Jun 2021 0 repositories listed
-
Language ID Prediction from Speech Using Self-Attentive Pooling1 Jun 2021 0 repositories listed
-
Multilingual Speech Translation with Unified Transformer: Huawei Noah's Ark Lab at IWSLT 20211 Jun 2021 0 repositories listed
-
NSYSU-MITLab團隊於福爾摩沙語音辨識競賽2020之語音辨識系統 (NSYSU-MITLab Speech Recognition System for Formosa Speech Recognition Challenge 2020)1 Jun 2021 0 repositories listed
-
Bangla Natural Language Processing: A Comprehensive Analysis of Classical, Machine Learning, and Deep Learning Based Methods31 May 2021 0 repositories listed
-
Fine-grained Generalization Analysis of Structured Output Prediction31 May 2021 0 repositories listed
-
Low-Resource Spoken Language Identification Using Self-Attentive Pooling and Deep 1D Time-Channel Separable Convolutions31 May 2021 0 repositories listed
-
Towards One Model to Rule All: Multilingual Strategy for Dialectal Code-Switching Arabic ASR31 May 2021 0 repositories listed
-
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition26 May 2021 0 repositories listed
-
Training Speech Enhancement Systems with Noisy Speech Datasets26 May 2021 0 repositories listed
-
A Streaming End-to-End Framework For Spoken Language Understanding20 May 2021 0 repositories listed
-
Mondegreen: A Post-Processing Solution to Speech Recognition Error Correction for Voice Search Queries20 May 2021 0 repositories listed
-
LiSTra, Automatic Speech Translation: English to Lingala case study16 May 2021 0 repositories listed
-
Listen with Intent: Improving Speech Recognition with Audio-to-Intent Front-End14 May 2021 0 repositories listed
-
Streaming Transformer for Hardware Efficient Voice Trigger Detection and False Trigger Mitigation14 May 2021 0 repositories listed
-
Exploring CTC Based End-to-End Techniques for Myanmar Speech Recognition13 May 2021 0 repositories listed
-
Attention-based Neural Beamforming Layers for Multi-channel Speech Recognition12 May 2021 0 repositories listed
-
Stacked Acoustic-and-Textual Encoding: Integrating the Pre-trained Models into Speech Translation Encoders12 May 2021 0 repositories listed
-
StutterNet: Stuttering Detection Using Time Delay Neural Network12 May 2021 0 repositories listed
-
10 May 2021 0 repositories listed
-
What shall we do with an hour of data? Speech recognition for the un- and under-served languages of Common Voice10 May 2021 0 repositories listed
-
English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System9 May 2021 0 repositories listed
-
Latency-Controlled Neural Architecture Search for Streaming Speech Recognition8 May 2021 0 repositories listed
-
Robustness of end-to-end Automatic Speech Recognition Models -- A Case Study using Mozilla DeepSpeech8 May 2021 0 repositories listed
-
Efficient Weight factorization for Multilingual Speech Recognition7 May 2021 0 repositories listed
-
Challenges and Obstacles Towards Deploying Deep Learning Models on Mobile Devices6 May 2021 0 repositories listed
-
Accent Recognition with Hybrid Phonetic Features5 May 2021 0 repositories listed
-
Performance Evaluation of Deep Convolutional Maxout Neural Network in Speech Recognition4 May 2021 0 repositories listed
-
Streaming end-to-end speech recognition with jointly trained neural feature enhancement4 May 2021 0 repositories listed
-
Hybrid Intelligence3 May 2021 0 repositories listed
-
3 May 2021 0 repositories listed
-
Quantifying and Maximizing the Benefits of Back-End Noise Adaption on Attention-Based Speech Recognition Models3 May 2021 0 repositories listed
-
Searchable Hidden Intermediates for End-to-End Models of Decomposable Sequence Tasks2 May 2021 0 repositories listed
-
Spectral modification for recognition of children’s speech undermismatched conditions1 May 2021 0 repositories listed
-
Deformable TDNN with adaptive receptive fields for speech recognition30 Apr 2021 0 repositories listed
-
Using Transformers to Provide Teachers with Personalized Feedback on their Classroom Discourse: The TalkMoves Application29 Apr 2021 0 repositories listed
-
Personalized Keyphrase Detection using Speaker and Environment Information28 Apr 2021 0 repositories listed
-
On Addressing Practical Challenges for RNN-Transducer27 Apr 2021 0 repositories listed
-
Head-synchronous Decoding for Transformer-based Streaming ASR26 Apr 2021 0 repositories listed
-
Multi-Task Learning for End-to-End ASR Word and Utterance Confidence with Deletion Prediction26 Apr 2021 0 repositories listed
-
Semantic Data Augmentation for End-to-End Mandarin Speech Recognition26 Apr 2021 0 repositories listed
-
Bridging the gap between streaming and non-streaming ASR systems bydistilling ensembles of CTC and RNN-T models25 Apr 2021 0 repositories listed
-
Quantization of Deep Neural Networks for Accurate Edge Computing25 Apr 2021 0 repositories listed
-
Scalable End-to-End RF Classification: A Case Study on Undersized Dataset Regularization by Convolutional-MST25 Apr 2021 0 repositories listed
-
Language ID Prediction from Speech Using Self-Attentive Pooling and 1D-Convolutions24 Apr 2021 0 repositories listed
-
Fast Text-Only Domain Adaptation of RNN-Transducer Prediction Network22 Apr 2021 0 repositories listed
-
Accented Speech Recognition: A Survey21 Apr 2021 0 repositories listed
-
Discriminative Self-training for Punctuation Prediction21 Apr 2021 0 repositories listed
-
Disfluency Detection with Unlabeled Data and Small BERT Models21 Apr 2021 0 repositories listed
-
Label-Synchronous Speech-to-Text Alignment for ASR Using Forward and Backward Transformers21 Apr 2021 0 repositories listed
-
On Sampling-Based Training Criteria for Neural Language Modeling21 Apr 2021 0 repositories listed
-
Pre-training for Spoken Language Understanding with Joint Textual and Phonetic Representation Learning21 Apr 2021 0 repositories listed
-
Scene-aware Far-field Automatic Speech Recognition21 Apr 2021 0 repositories listed
-
On the Impact of Word Error Rate on Acoustic-Linguistic Speech Emotion Recognition: An Update for the Deep Learning Era20 Apr 2021 0 repositories listed
-
Acoustic Data-Driven Subword Modeling for End-to-End Speech Recognition19 Apr 2021 0 repositories listed
-
Advanced Long-context End-to-end Speech Recognition Using Context-expanded Transformers19 Apr 2021 0 repositories listed
-
Fusing information streams in end-to-end audio-visual speech recognition19 Apr 2021 0 repositories listed
-
Learning on Hardware: A Tutorial on Neural Network Accelerators and Co-Processors19 Apr 2021 0 repositories listed
-
Best Practices for Noise-Based Augmentation to Improve the Performance of Deployable Speech-Based Emotion Recognition Systems18 Apr 2021 0 repositories listed
-
MIMO Self-attentive RNN Beamformer for Multi-speaker Speech Separation17 Apr 2021 0 repositories listed
-
Multilingual and Cross-Lingual Intent Detection from Spoken Data17 Apr 2021 0 repositories listed
-
Bridging the Gap Between Clean Data Training and Real-World Inference for Spoken Language Understanding13 Apr 2021 0 repositories listed
-
Equivalence of Segmental and Neural Transducer Modeling: A Proof of Concept13 Apr 2021 0 repositories listed
-
Experiments of ASR-based mispronunciation detection for children and adult English learners13 Apr 2021 0 repositories listed
-
Source and Target Bidirectional Knowledge Distillation for End-to-end Speech Translation13 Apr 2021 0 repositories listed
-
Comparing the Benefit of Synthetic Training Data for Various Automatic Speech Recognition Architectures12 Apr 2021 0 repositories listed
-
Improved Conformer-based End-to-End Speech Recognition Using Neural Architecture Search12 Apr 2021 0 repositories listed
-
Innovative Bert-based Reranking Language Models for Speech Recognition11 Apr 2021 0 repositories listed
-
Non-autoregressive Transformer-based End-to-end ASR using BERT10 Apr 2021 0 repositories listed
-
Accented Speech Recognition Inspired by Human Perception9 Apr 2021 0 repositories listed
-
On Architectures and Training for Raw Waveform Feature Extraction in ASR9 Apr 2021 0 repositories listed
-
Language model fusion for streaming end to end speech recognition9 Apr 2021 0 repositories listed
-
Lookup-Table Recurrent Language Models for Long Tail Speech Recognition9 Apr 2021 0 repositories listed
-
The NTNU Taiwanese ASR System for Formosa Speech Recognition Challenge 20209 Apr 2021 0 repositories listed
-
8 Apr 2021 0 repositories listed
-
Contextual Semi-Supervised Learning: An Approach To Leverage Air-Surveillance and Untranscribed ATC Data in ASR Systems8 Apr 2021 0 repositories listed