Browse State-of-the-Art › speech-recognition › Papers, page 43
speech-recognition
Papers archive 2025-07-28
archive papers tagged: 5,715 · with a code link: 1,277 · where Syntology ran a sample: 162 (134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (162 of 5,715 tagged: 134 with a run with no instrument failure, 28 where every run was a failure of Syntology's instrument)
Page 43 of 58: papers 4,201 to 4,300 of 5,715, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Start-Before-End and End-to-End: Neural Speech Translation by AppTek and RWTH Aachen University1 Jul 2020 0 repositories listed
-
The AFRL IWSLT 2020 Systems: Work-From-Home Edition1 Jul 2020 0 repositories listed
-
Tigrinya Automatic Speech recognition with Morpheme based recognition units1 Jul 2020 0 repositories listed
-
Towards end-2-end learning for predicting behavior codes from spoken utterances in psychotherapy conversations1 Jul 2020 0 repositories listed
-
Towards Stream Translation: Adaptive Computation Time for Simultaneous Machine Translation1 Jul 2020 0 repositories listed
-
Towards Understanding ASR Error Correction for Medical Conversations1 Jul 2020 0 repositories listed
-
Multi-view Frequency LSTM: An Efficient Frontend for Automatic Speech Recognition30 Jun 2020 0 repositories listed
-
Neural Machine Translation for Multilingual Grapheme-to-Phoneme Conversion25 Jun 2020 0 repositories listed
-
25 Jun 2020 0 repositories listed
-
Streaming Transformer ASR with Blockwise Synchronous Inference25 Jun 2020 0 repositories listed
-
One Model to Pronounce Them All: Multilingual Grapheme-to-Phoneme Conversion With a Transformer Ensemble23 Jun 2020 0 repositories listed
-
Bayesian Neural Networks: An Introduction and Survey22 Jun 2020 0 repositories listed
-
Self-Supervised Representations Improve End-to-End Speech Translation22 Jun 2020 0 repositories listed
-
Deep Double-Side Learning Ensemble Model for Few-Shot Parkinson Speech Recognition20 Jun 2020 0 repositories listed
-
Boosting Active Learning for Speech Recognition with Noisy Pseudo-labeled Samples19 Jun 2020 0 repositories listed
-
Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers19 Jun 2020 0 repositories listed
-
End-to-End Code Switching Language Models for Automatic Speech Recognition16 Jun 2020 0 repositories listed
-
Quantization of Acoustic Model Parameters in Automatic Speech Recognition Framework16 Jun 2020 0 repositories listed
-
Towards Automated Assessment of Stuttering and Stuttering Therapy16 Jun 2020 0 repositories listed
-
Exploration of End-to-End ASR for OpenSTT -- Russian Open Speech-to-Text Dataset15 Jun 2020 0 repositories listed
-
Regularized Forward-Backward Decoder for Attention Models15 Jun 2020 0 repositories listed
-
The JHU Multi-Microphone Multi-Speaker ASR System for the CHiME-6 Challenge14 Jun 2020 0 repositories listed
-
UWSpeech: Speech to Speech Translation for Unwritten Languages14 Jun 2020 0 repositories listed
-
FrugalML: How to Use ML Prediction APIs More Accurately and Cheaply12 Jun 2020 0 repositories listed
-
"Notic My Speech" -- Blending Speech Patterns With Multimedia12 Jun 2020 0 repositories listed
-
Data Augmentation for Training Dialog Models Robust to Speech Recognition Errors10 Jun 2020 0 repositories listed
-
Improving Cross-Lingual Transfer Learning for End-to-End Speech Recognition with Speech Translation9 Jun 2020 0 repositories listed
-
Learning not to Discriminate: Task Agnostic Learning for Improving Monolingual and Code-switched Speech Recognition9 Jun 2020 0 repositories listed
-
On the Effectiveness of Neural Text Generation based Data Augmentation for Recognition of Morphologically Rich Speech9 Jun 2020 0 repositories listed
-
Characterizing the Weight Space for Different Learning Models4 Jun 2020 0 repositories listed
-
Contextual RNN-T For Open Domain ASR4 Jun 2020 0 repositories listed
-
Multi-talker ASR for an unknown number of sources: Joint training of source counting, separation and ASR4 Jun 2020 0 repositories listed
-
Self-Training for End-to-End Speech Translation3 Jun 2020 0 repositories listed
-
Transfer Learning for British Sign Language Modelling3 Jun 2020 0 repositories listed
-
Analyzing the Quality and Stability of a Streaming End-to-End On-Device Speech Recognizer2 Jun 2020 0 repositories listed
-
Detecting Audio Attacks on ASR Systems with Dropout Uncertainty2 Jun 2020 0 repositories listed
-
Dilated U-net based approach for multichannel speech enhancement from First-Order Ambisonics recordings2 Jun 2020 0 repositories listed
-
An Effective Contextual Language Modeling Framework for Speech Summarization with Augmented Features1 Jun 2020 0 repositories listed
-
Analyse de l'effet de la réverbération sur la reconnaissance automatique de la parole (Analyzing how reverberation affects Automatic Speech Recognition)1 Jun 2020 0 repositories listed
-
Constrained Variational Autoencoder for improving EEG based Speech Recognition Systems1 Jun 2020 0 repositories listed
-
Introduction d'informations sémantiques dans un système de reconnaissance de la parole (Despite spectacular advances in recent years, the Automatic Speech Recognition (ASR) systems still make mistakes, especially in noisy environments)1 Jun 2020 0 repositories listed
-
Learning to Recognize Code-switched Speech Without Forgetting Monolingual Speech Recognition1 Jun 2020 0 repositories listed
-
Reconnaissance automatique de la parole : génération des prononciations non natives pour l'enrichissement du lexique (In this study we propose a method for lexicon adaptation in order to improve the automatic speech recognition (ASR) of non-native speakers)1 Jun 2020 0 repositories listed
-
Streaming Language Identification using Combination of Acoustic Representations and ASR Hypotheses1 Jun 2020 0 repositories listed
-
Sur l'utilisation de la reconnaissance automatique de la parole pour l'aide au diagnostic différentiel entre la maladie de Parkinson et l'AMS (On using automatic speech recognition for the differential diagnosis of Parkinson's Disease and MSA This article presents a study regarding the contribution of automatic speech processing in the differential diagnosis between Parkinson's disease and MSA (Multi-System Atrophies))1 Jun 2020 0 repositories listed
-
Dynamic Masking for Improved Stability in Spoken Language Translation30 May 2020 0 repositories listed
-
Improving EEG based continuous speech recognition using GAN29 May 2020 0 repositories listed
-
Understanding effect of speech perception in EEG based speech recognition systems29 May 2020 0 repositories listed
-
Adversarial Attacks and Defense on Texts: A Survey28 May 2020 0 repositories listed
-
When Can Self-Attention Be Replaced by Feed Forward Layers?28 May 2020 0 repositories listed
-
Predicting Entity Popularity to Improve Spoken Entity Recognition by Virtual Assistants26 May 2020 0 repositories listed
-
An Audio-enriched BERT-based Framework for Spoken Multiple-choice Question Answering25 May 2020 0 repositories listed
-
An End-to-End Mispronunciation Detection System for L2 English Speech Leveraging Novel Anti-Phone Modeling25 May 2020 0 repositories listed
-
25 May 2020 0 repositories listed
-
21 May 2020 0 repositories listed
-
Large scale evaluation of importance maps in automatic speech recognition21 May 2020 0 repositories listed
-
Multistream CNN for Robust Acoustic Modeling21 May 2020 0 repositories listed
-
Simplified Self-Attention for Transformer-based End-to-End Speech Recognition21 May 2020 0 repositories listed
-
A Comparison of Label-Synchronous and Frame-Synchronous End-to-End Models for Speech Recognition20 May 2020 0 repositories listed
-
Early Stage LM Integration Using Local and Global Log-Linear Combination20 May 2020 0 repositories listed
-
Investigation of Large-Margin Softmax in Neural Language Modeling20 May 2020 0 repositories listed
-
Relative Positional Encoding for Speech Recognition and Direct Translation20 May 2020 0 repositories listed
-
Deep learning approaches for neural decoding: from CNNs to LSTMs and spikes to fMRI19 May 2020 0 repositories listed
-
Exploring Transformers for Large-Scale Speech Recognition19 May 2020 0 repositories listed
-
Improving Proper Noun Recognition in End-to-End ASR By Customization of the MWER Loss Criterion19 May 2020 0 repositories listed
-
An Effective End-to-End Modeling Approach for Mispronunciation Detection18 May 2020 0 repositories listed
-
Attention-based Transducer for Online Speech Recognition18 May 2020 0 repositories listed
-
Audio-visual Multi-channel Recognition of Overlapped Speech18 May 2020 0 repositories listed
-
Weak-Attention Suppression For Transformer Based Speech Recognition18 May 2020 0 repositories listed
-
Speech to Text Adaptation: Towards an Efficient Cross-Modal Distillation17 May 2020 0 repositories listed
-
A Deep Learning based Wearable Healthcare IoT Device for AI-enabled Hearing Assistance Automation16 May 2020 0 repositories listed
-
16 May 2020 0 repositories listed
-
Dynamic Sparsity Neural Networks for Automatic Speech Recognition16 May 2020 0 repositories listed
-
Large scale weakly and semi-supervised learning for low-resource video ASR16 May 2020 0 repositories listed
-
Reducing Spelling Inconsistencies in Code-Switching ASR using Contextualized CTC Loss16 May 2020 0 repositories listed
-
Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition16 May 2020 0 repositories listed
-
That Sounds Familiar: an Analysis of Phonetic Representations Transfer Across Languages16 May 2020 0 repositories listed
-
Context-Dependent Acoustic Modeling without Explicit Phone Clustering15 May 2020 0 repositories listed
-
Contextualizing ASR Lattice Rescoring with Hybrid Pointer Network Language Model15 May 2020 0 repositories listed
-
You Do Not Need More Data: Improving End-To-End Speech Recognition by Text-To-Speech Data Augmentation14 May 2020 0 repositories listed
-
DARTS-ASR: Differentiable Architecture Search for Multilingual Speech Recognition and Adaptation13 May 2020 0 repositories listed
-
Automatic Estimation of Intelligibility Measure for Consonants in Speech12 May 2020 0 repositories listed
-
DiscreTalk: Text-to-Speech as a Machine Translation Problem12 May 2020 0 repositories listed
-
Incremental Learning for End-to-End Automatic Speech Recognition11 May 2020 0 repositories listed
-
Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition11 May 2020 0 repositories listed
-
RNN-T Models Fail to Generalize to Out-of-Domain Audio: Causes and Solutions7 May 2020 0 repositories listed
-
The Perceptimatic English Benchmark for Speech Perception Models7 May 2020 0 repositories listed
-
Community Detection Clustering via Gumbel Softmax5 May 2020 0 repositories listed
-
End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training5 May 2020 0 repositories listed
-
Does Visual Self-Supervision Improve Learning of Speech Representations for Emotion Recognition?4 May 2020 0 repositories listed
-
Fast and Robust Unsupervised Contextual Biasing for Speech Recognition4 May 2020 0 repositories listed
-
Off-the-shelf deep learning is not enough: parsimony, Bayes and causality4 May 2020 0 repositories listed
-
A language score based output selection method for multilingual speech recognition2 May 2020 0 repositories listed
-
MultiQT: Multimodal Learning for Real-Time Question Tracking in Speech2 May 2020 0 repositories listed
-
A CLARIN Transcription Portal for Interview Data1 May 2020 0 repositories listed
-
Acoustic-Phonetic Approach for ASR of Less Resourced Languages Using Monolingual and Cross-Lingual Information1 May 2020 0 repositories listed
-
An Investigative Study of Multi-Modal Cross-Lingual Retrieval1 May 2020 0 repositories listed
-
Analysis of GlobalPhone and Ethiopian Languages Speech Corpora for Multilingual ASR1 May 2020 0 repositories listed
-
1 May 2020 0 repositories listed
-
ArzEn: A Speech Corpus for Code-switched Egyptian Arabic-English1 May 2020 0 repositories listed