Browse State-of-the-Art › Speech Recognition › Papers, page 42
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 42 of 65: papers 4,101 to 4,200 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Panoramic Survey of Natural Language Processing in the Arab World25 Nov 2020 0 repositories listed
-
Bootstrap an end-to-end ASR system by multilingual training, transfer learning, text-to-text mapping and synthetic audio25 Nov 2020 0 repositories listed
-
A Review of Recent Advances of Binary Neural Networks for Edge Computing24 Nov 2020 0 repositories listed
-
Adam^+: A Stochastic Method with Adaptive Variance Reduction24 Nov 2020 0 repositories listed
-
Synth2Aug: Cross-domain speaker recognition with TTS synthesized speech24 Nov 2020 0 repositories listed
-
Multi-task Language Modeling for Improving Speech Recognition of Rare Words23 Nov 2020 0 repositories listed
-
Streaming Multi-speaker ASR with RNN-T23 Nov 2020 0 repositories listed
-
Using Synthetic Audio to Improve The Recognition of Out-Of-Vocabulary Words in End-To-End ASR Systems23 Nov 2020 0 repositories listed
-
Deep Learning in EEG: Advance of the Last Ten-Year Critical Period22 Nov 2020 0 repositories listed
-
Improving RNN-T ASR Accuracy Using Context Audio20 Nov 2020 0 repositories listed
-
TaL: a synchronised multi-speaker corpus of ultrasound tongue imaging, audio, and lip videos19 Nov 2020 0 repositories listed
-
Cascade RNN-Transducer: Syllable Based Streaming On-device Mandarin Speech Recognition with a Syllable-to-Character Converter17 Nov 2020 0 repositories listed
-
Empowering Things with Intelligence: A Survey of the Progress, Challenges, and Opportunities in Artificial Intelligence of Things17 Nov 2020 0 repositories listed
-
Refining Automatic Speech Recognition System for older adults17 Nov 2020 0 repositories listed
-
Audio-visual Multi-channel Integration and Recognition of Overlapped Speech16 Nov 2020 0 repositories listed
-
Deep Shallow Fusion for RNN-T Personalization16 Nov 2020 0 repositories listed
-
Improving Speech Enhancement Performance by Leveraging Contextual Broad Phonetic Class Information15 Nov 2020 0 repositories listed
-
11 TeraFLOPs per second photonic convolutional accelerator for deep learning optical neural networks14 Nov 2020 0 repositories listed
-
Self-supervised reinforcement learning for speaker localisation with the iCub humanoid robot12 Nov 2020 0 repositories listed
-
FAT: Training Neural Networks for Reliable Inference Under Hardware Faults11 Nov 2020 0 repositories listed
-
On End-to-end Multi-channel Time Domain Speech Separation in Reverberant Environments11 Nov 2020 0 repositories listed
-
Towards Semi-Supervised Semantics Understanding from Speech11 Nov 2020 0 repositories listed
-
Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS10 Nov 2020 0 repositories listed
-
Benchmarking LF-MMI, CTC and RNN-T Criteria for Streaming ASR9 Nov 2020 0 repositories listed
-
Efficient End-to-End Speech Recognition Using Performers in Conformers9 Nov 2020 0 repositories listed
-
Gated Recurrent Fusion with Joint Training Framework for Robust End-to-End Speech Recognition9 Nov 2020 0 repositories listed
-
Neural Architecture Search with an Efficient Multiobjective Evolutionary Framework9 Nov 2020 0 repositories listed
-
Personalized Query Rewriting in Conversational AI Agents9 Nov 2020 0 repositories listed
-
On the Usefulness of Self-Attention for Automatic Speech Recognition with Transformers8 Nov 2020 0 repositories listed
-
Acoustics Based Intent Recognition Using Discovered Phonetic Units for Low Resource Languages7 Nov 2020 0 repositories listed
-
ESPnet-se: end-to-end speech enhancement and separation toolkit designed for asr integration7 Nov 2020 0 repositories listed
-
Resource-Constrained Federated Learning with Heterogeneous Labels and Models6 Nov 2020 0 repositories listed
-
Alignment Restricted Streaming Recurrent Neural Network Transducer5 Nov 2020 0 repositories listed
-
Multi-Accent Adaptation based on Gate Mechanism5 Nov 2020 0 repositories listed
-
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework4 Nov 2020 0 repositories listed
-
Cross-Lingual Machine Speech Chain for Javanese, Sundanese, Balinese, and Bataks Speech Recognition and Synthesis4 Nov 2020 0 repositories listed
-
Data Augmentation for End-to-end Code-switching Speech Recognition4 Nov 2020 0 repositories listed
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
Paralinguistic Privacy Protection at the Edge4 Nov 2020 0 repositories listed
-
Sequence-to-Sequence Learning via Attention Transfer for Incremental Speech Recognition4 Nov 2020 0 repositories listed
-
Dynamic latency speech recognition with asynchronous revision3 Nov 2020 0 repositories listed
-
Improving RNN transducer with normalized jointer network3 Nov 2020 0 repositories listed
-
Integration of speech separation, diarization, and recognition for multi-speaker meetings: System description, comparison, and analysis3 Nov 2020 0 repositories listed
-
Internal Language Model Estimation for Domain-Adaptive End-to-End Speech Recognition3 Nov 2020 0 repositories listed
-
Learning Explicit Prosody Models and Deep Speaker Embeddings for Atypical Voice Conversion3 Nov 2020 0 repositories listed
-
Streaming Attention-Based Models with Augmented Memory for End-to-End Speech Recognition3 Nov 2020 0 repositories listed
-
Unsupervised Pattern Discovery from Thematic Speech Archives Based on Multilingual Bottleneck Features3 Nov 2020 0 repositories listed
-
Warped Language Models for Noise Robust Language Understanding3 Nov 2020 0 repositories listed
-
DNN-Based Semantic Model for Rescoring N-best Speech Recognition List2 Nov 2020 0 repositories listed
-
Focus on the present: a regularization method for the ASR source-target attention layer2 Nov 2020 0 repositories listed
-
Multitask Learning and Joint Optimization for Transformer-RNN-Transducer Speech Recognition2 Nov 2020 0 repositories listed
-
SapAugment: Learning A Sample Adaptive Policy for Data Augmentation2 Nov 2020 0 repositories listed
-
Effectively pretraining a speech translation decoder with Machine Translation data1 Nov 2020 0 repositories listed
-
ELITR: European Live Translator1 Nov 2020 0 repositories listed
-
Impact of ASR on Alzheimer’s Disease Detection: All Errors are Equal, but Deletions are More Equal than Others1 Nov 2020 0 repositories listed
-
Improving End-to-End Bangla Speech Recognition with Semi-supervised Training1 Nov 2020 0 repositories listed
-
May I Ask Who’s Calling? Named Entity Recognition on Call Center Transcripts for Privacy Law Compliance1 Nov 2020 0 repositories listed
-
Simultaneous Translation1 Nov 2020 0 repositories listed
-
Directional ASR: A New Paradigm for E2E Multi-Speaker Speech Recognition with Source Localization30 Oct 2020 0 repositories listed
-
Phoneme Based Neural Transducer for Large Vocabulary Speech Recognition30 Oct 2020 0 repositories listed
-
Streaming Simultaneous Speech Translation with Augmented Memory Transformer30 Oct 2020 0 repositories listed
-
May I Ask Who's Calling? Named Entity Recognition on Call Center Transcripts for Privacy Law Compliance29 Oct 2020 0 repositories listed
-
Semi-Supervised Speech Recognition via Graph-based Temporal Classification29 Oct 2020 0 repositories listed
-
Training Speech Recognition Models with Federated Learning: A Quality/Cost Framework29 Oct 2020 0 repositories listed
-
CASS-NAT: CTC Alignment-based Single Step Non-autoregressive Transformer for Speech Recognition28 Oct 2020 0 repositories listed
-
Decoupling Pronunciation and Language for End-to-end Code-switching Automatic Speech Recognition28 Oct 2020 0 repositories listed
-
Fusion Models for Improved Visual Captioning28 Oct 2020 0 repositories listed
-
INT8 Winograd Acceleration for Conv1D Equipped ASR Models Deployed on Mobile Devices28 Oct 2020 0 repositories listed
-
Non-Autoregressive Transformer ASR with CTC-Enhanced Decoder Input28 Oct 2020 0 repositories listed
-
One In A Hundred: Select The Best Predicted Sequence from Numerous Candidates for Streaming Speech Recognition28 Oct 2020 0 repositories listed
-
Cascaded encoders for unifying streaming and non-streaming ASR27 Oct 2020 0 repositories listed
-
Effective Decoder Masking for Transformer Based End-to-End Speech Recognition27 Oct 2020 0 repositories listed
-
Emotion recognition by fusing time synchronous and time asynchronous representations27 Oct 2020 0 repositories listed
-
End-to-End Far-Field Speech Recognition with Unified Dereverberation and Beamforming27 Oct 2020 0 repositories listed
-
Multitask Training with Text Data for End-to-End Speech Recognition27 Oct 2020 0 repositories listed
-
Transformer in action: a comparative study of transformer-based acoustic models for large scale speech recognition applications27 Oct 2020 0 repositories listed
-
Improved Mask-CTC for Non-Autoregressive End-to-End ASR26 Oct 2020 0 repositories listed
-
Improved Neural Language Model Fusion for Streaming Recurrent Neural Network Transducer26 Oct 2020 0 repositories listed
-
25 Oct 2020 0 repositories listed
-
Align-Refine: Non-Autoregressive Speech Recognition via Iterative Realignment24 Oct 2020 0 repositories listed
-
Auxiliary Sequence Labeling Tasks for Disfluency Detection24 Oct 2020 0 repositories listed
-
Improving Noise Robustness of an End-to-End Neural Model for Automatic Speech Recognition23 Oct 2020 0 repositories listed
-
On Minimum Word Error Rate Training of the Hybrid Autoregressive Transducer23 Oct 2020 0 repositories listed
-
Multilingual Approach to Joint Speech and Accent Recognition with DNN-HMM Framework22 Oct 2020 0 repositories listed
-
Developing Real-time Streaming Transformer Transducer for Speech Recognition on Large-scale Dataset22 Oct 2020 0 repositories listed
-
Improving Streaming Automatic Speech Recognition With Non-Streaming Model Distillation On Unsupervised Data22 Oct 2020 0 repositories listed
-
MAM: Masked Acoustic Modeling for End-to-End Speech-to-Text Translation22 Oct 2020 0 repositories listed
-
SlimIPL: Language-Model-Free Iterative Pseudo-Labeling22 Oct 2020 0 repositories listed
-
A General Multi-Task Learning Framework to Leverage Text Data for Speech to Text Tasks21 Oct 2020 0 repositories listed
-
Cascaded Models With Cyclic Feedback For Direct Speech Translation21 Oct 2020 0 repositories listed
-
Knowledge Distillation for Improved Accuracy in Spoken Question Answering21 Oct 2020 0 repositories listed
-
LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition21 Oct 2020 0 repositories listed
-
Sentence Boundary Augmentation For Neural Machine Translation Robustness21 Oct 2020 0 repositories listed
-
Knowledge Transfer for Efficient On-device False Trigger Mitigation20 Oct 2020 0 repositories listed
-
Replacing Human Audio with Synthetic Audio for On-device Unspoken Punctuation Prediction20 Oct 2020 0 repositories listed
-
Speaker Separation Using Speaker Inventories and Estimated Speech20 Oct 2020 0 repositories listed
-
Ensemble Chinese End-to-End Spoken Language Understanding for Abnormal Event Detection from audio stream19 Oct 2020 0 repositories listed
-
Reduce and Reconstruct: ASR for Low-Resource Phonetic Languages19 Oct 2020 0 repositories listed
-
Towards Data Distillation for End-to-end Spoken Conversational Question Answering18 Oct 2020 0 repositories listed
-
Studying the Similarity of COVID-19 Sounds based on Correlation Analysis of MFCC17 Oct 2020 0 repositories listed