Browse State-of-the-Art › Automatic Speech Recognition › Papers, page 28
Automatic Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 3,174 · with a code link: 677 · where Syntology ran a sample: 79 (62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (79 of 3,174 tagged: 62 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 28 of 32: papers 2,701 to 2,800 of 3,174, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Transformer-based Cascaded Multimodal Speech Translation29 Oct 2019 0 repositories listed
-
Improving sequence-to-sequence speech recognition training with on-the-fly data augmentation29 Oct 2019 0 repositories listed
-
Does Speech enhancement of publicly available data help build robust Speech Recognition Systems?29 Oct 2019 0 repositories listed
-
DFSMN-SAN with Persistent Memory Model for Automatic Speech Recognition28 Oct 2019 0 repositories listed
-
Unsupervised pre-training for sequence to sequence speech recognition28 Oct 2019 0 repositories listed
-
Meta Learning for End-to-End Low-Resource Speech Recognition26 Oct 2019 0 repositories listed
-
L2RS: A Learning-to-Rescore Mechanism for Automatic Speech Recognition25 Oct 2019 0 repositories listed
-
Towards Online End-to-end Transformer Automatic Speech Recognition25 Oct 2019 0 repositories listed
-
Pre-training in Deep Reinforcement Learning for Automatic Speech Recognition24 Oct 2019 0 repositories listed
-
Recognizing long-form speech using streaming end-to-end models24 Oct 2019 0 repositories listed
-
A practical two-stage training strategy for multi-stream end-to-end speech recognition23 Oct 2019 0 repositories listed
-
Analyzing ASR pretraining for low-resource speech-to-text translation23 Oct 2019 0 repositories listed
-
Correction of Automatic Speech Recognition with Transformer Sequence-to-sequence Model23 Oct 2019 0 repositories listed
-
RNN based Incremental Online Spoken Language Understanding23 Oct 2019 0 repositories listed
-
G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR22 Oct 2019 0 repositories listed
-
Robust Neural Machine Translation for Clean and Noisy Speech Transcripts22 Oct 2019 0 repositories listed
-
AeGAN: Time-Frequency Speech Denoising via Generative Adversarial Networks21 Oct 2019 0 repositories listed
-
Neuro-SERKET: Development of Integrative Cognitive System through the Composition of Deep Probabilistic Generative Models20 Oct 2019 0 repositories listed
-
Multi-Talker MVDR Beamforming Based on Extended Complex Gaussian Mixture Model17 Oct 2019 0 repositories listed
-
Lead2Gold: Towards exploiting the full potential of noisy transcriptions for speech recognition16 Oct 2019 0 repositories listed
-
Transformer ASR with Contextual Block Processing16 Oct 2019 0 repositories listed
-
Analyzing Large Receptive Field Convolutional Networks for Distant Speech Recognition15 Oct 2019 0 repositories listed
-
VAIS ASR: Building a conversational speech recognition system using language model combination12 Oct 2019 0 repositories listed
-
Hear "No Evil", See "Kenansville": Efficient and Transferable Black-Box Attacks on Speech Recognition and Voice Identification Systems11 Oct 2019 0 repositories listed
-
Query-by-example on-device keyword spotting11 Oct 2019 0 repositories listed
-
One-To-Many Multilingual End-to-end Speech Translation8 Oct 2019 0 repositories listed
-
A Case Study on Combining ASR and Visual Features for Generating Instructional Video Captions7 Oct 2019 0 repositories listed
-
Adapting a FrameNet Semantic Parser for Spoken Language Understanding Using Adversarial Learning7 Oct 2019 0 repositories listed
-
Modeling Confidence in Sequence-to-Sequence Models4 Oct 2019 0 repositories listed
-
Neural Zero-Inflated Quality Estimation Model For Automatic Speech Recognition System3 Oct 2019 0 repositories listed
-
From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition2 Oct 2019 0 repositories listed
-
使用生成對抗網路於強健式自動語音辨識的應用(Exploiting Generative Adversarial Network for Robustness Automatic Speech Recognition)1 Oct 2019 0 repositories listed
-
End-to-End Code-Switching ASR for Low-Resourced Language Pairs27 Sep 2019 0 repositories listed
-
Breaking the Data Barrier: Towards Robust Speech Translation via Adversarial Stability Training25 Sep 2019 0 repositories listed
-
Generating Robust Audio Adversarial Examples using Iterative Proportional Clipping25 Sep 2019 0 repositories listed
-
Improved Training Techniques for Online Neural Machine Translation25 Sep 2019 0 repositories listed
-
Understanding Semantics from Speech Through Pre-training24 Sep 2019 0 repositories listed
-
Improving OOV Detection and Resolution with External Language Models in Acoustic-to-Word ASR22 Sep 2019 0 repositories listed
-
Code-Switched Language Models Using Neural Based Synthetic Data from Parallel Sentences18 Sep 2019 0 repositories listed
-
Simultaneous Speech Recognition and Speaker Diarization for Monaural Dialogue Recordings with Target-Speaker Acoustic Models17 Sep 2019 0 repositories listed
-
An Investigation Into On-device Personalization of End-to-end Automatic Speech Recognition Models14 Sep 2019 0 repositories listed
-
Integrating Source-channel and Attention-based Sequence-to-sequence Models for Speech Recognition14 Sep 2019 0 repositories listed
-
Harnessing Indirect Training Data for End-to-End Automatic Speech Translation: Tricks of the Trade14 Sep 2019 0 repositories listed
-
Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model11 Sep 2019 0 repositories listed
-
Neural Network-Based Modeling of Phonetic Durations6 Sep 2019 0 repositories listed
-
Dialect-Specific Models for Automatic Speech Recognition of African American Vernacular English1 Sep 2019 0 repositories listed
-
Human-Informed Speakers and Interpreters Analysis in the WAW Corpus and an Automatic Method for Calculating Interpreters' Décalage1 Sep 2019 0 repositories listed
-
Motivations, challenges, and perspectives for the development of an Automatic Speech Recognition System for the under-resourced Ngiemboon Language1 Sep 2019 0 repositories listed
-
Semantic Language Model for Tunisian Dialect1 Sep 2019 0 repositories listed
-
Towards Accurate Text Verbalization for ASR Based on Audio Alignment1 Sep 2019 0 repositories listed
-
Deploying Technology to Save Endangered Languages23 Aug 2019 0 repositories listed
-
Gender Representation in French Broadcast Corpora and Its Impact on ASR Performance23 Aug 2019 0 repositories listed
-
Towards Better Understanding of Spontaneous Conversations: Overcoming Automatic Speech Recognition Errors With Intent Recognition21 Aug 2019 0 repositories listed
-
Two-Staged Acoustic Modeling Adaption for Robust Speech Recognition by the Example of German Oral History Interviews19 Aug 2019 0 repositories listed
-
End-to-End Multi-Speaker Speech Recognition using Speaker Embeddings and Transfer Learning13 Aug 2019 0 repositories listed
-
Unsupervised Stemming based Language Model for Telugu Broadcast News Transcription10 Aug 2019 0 repositories listed
-
Exploiting Cross-Lingual Speaker and Phonetic Diversity for Unsupervised Subword Modeling9 Aug 2019 0 repositories listed
-
Exploiting semi-supervised training through a dropout regularization in end-to-end speech recognition8 Aug 2019 0 repositories listed
-
Mitigating Noisy Inputs for Question Answering8 Aug 2019 0 repositories listed
-
Fast and Accurate Capitalization and Punctuation for Automatic Speech Recognition Using Transformer and Chunk Merging7 Aug 2019 0 repositories listed
-
An End-to-End Text-independent Speaker Verification Framework with a Keyword Adversarial Network6 Aug 2019 0 repositories listed
-
Practical Speech Recognition with HTK6 Aug 2019 0 repositories listed
-
Imperio: Robust Over-the-Air Adversarial Examples for Automatic Speech Recognition Systems5 Aug 2019 0 repositories listed
-
V2S attack: building DNN-based voice conversion from automatic speaker verification5 Aug 2019 0 repositories listed
-
A Speech Test Set of Practice Business Presentations with Additional Relevant Texts2 Aug 2019 0 repositories listed
-
DuTongChuan: Context-aware Translation Model for Simultaneous Interpreting30 Jul 2019 0 repositories listed
-
Correlation Distance Skip Connection Denoising Autoencoder (CDSK-DAE) for Speech Feature Enhancement26 Jul 2019 0 repositories listed
-
Hierarchical Sequence to Sequence Voice Conversion with Limited Data15 Jul 2019 0 repositories listed
-
Investigating Target Set Reduction for End-to-End Speech Recognition of Hindi-English Code-Switching Data15 Jul 2019 0 repositories listed
-
A Highly Efficient Distributed Deep Learning System For Automatic Speech Recognition10 Jul 2019 0 repositories listed
-
Acoustic Model Optimization Based On Evolutionary Stochastic Gradient Descent with Anchors for Automatic Speech Recognition10 Jul 2019 0 repositories listed
-
Large-Scale Mixed-Bandwidth Deep Neural Network Acoustic Modeling for Automatic Speech Recognition10 Jul 2019 0 repositories listed
-
Joint Speech Recognition and Speaker Diarization via Sequence Transduction9 Jul 2019 0 repositories listed
-
Teach an all-rounder with experts in different domains9 Jul 2019 0 repositories listed
-
ShrinkML: End-to-End ASR Model Compression Using Reinforcement Learning8 Jul 2019 0 repositories listed
-
Improved low-resource Somali speech recognition by semi-supervised acoustic and language model training6 Jul 2019 0 repositories listed
-
End-to-End Speech Recognition with High-Frame-Rate Features Extraction3 Jul 2019 0 repositories listed
-
2 Jul 2019 0 repositories listed
-
Latent Dirichlet Allocation Based Acoustic Data Selection for Automatic Speech Recognition2 Jul 2019 0 repositories listed
-
Scalable Multi Corpora Neural Language Models for ASR2 Jul 2019 0 repositories listed
-
Automated Cross-language Intelligibility Analysis of Parkinson's Disease Patients Using Speech Recognition Technologies1 Jul 2019 0 repositories listed
-
Comparison of Lattice-Free and Lattice-Based Sequence Discriminative Training Criteria for LVCSR1 Jul 2019 0 repositories listed
-
Analyzing Utility of Visual Context in Multimodal Speech Recognition Under Noisy Conditions30 Jun 2019 0 repositories listed
-
Auxiliary Interference Speaker Loss for Target-Speaker Speech Recognition26 Jun 2019 0 repositories listed
-
One Size Does Not Fit All: Quantifying and Exposing the Accuracy-Latency Trade-off in Machine Learning Cloud Service APIs via Tolerance Tiers26 Jun 2019 0 repositories listed
-
Multi-Span Acoustic Modelling using Raw Waveform Signals21 Jun 2019 0 repositories listed
-
Phoneme-Based Contextualization for Cross-Lingual Speech Recognition in End-to-End Models21 Jun 2019 0 repositories listed
-
Code-Switching Detection Using ASR-Generated Language Posteriors19 Jun 2019 0 repositories listed
-
Multi-Graph Decoding for Code-Switching ASR18 Jun 2019 0 repositories listed
-
Advancing Speech Recognition With No Speech Or With Noisy Speech17 Jun 2019 0 repositories listed
-
Adversarial Training for Multilingual Acoustic Modeling17 Jun 2019 0 repositories listed
-
Multi-Stream End-to-End Speech Recognition17 Jun 2019 0 repositories listed
-
Real to H-space Encoder for Speech Recognition17 Jun 2019 0 repositories listed
-
Cumulative Adaptation for BLSTM Acoustic Models14 Jun 2019 0 repositories listed
-
Learning Video Representations using Contrastive Bidirectional Transformer13 Jun 2019 0 repositories listed
-
Lattice Transformer for Speech Translation13 Jun 2019 0 repositories listed
-
Listening while Speaking and Visualizing: Improving ASR through Multimodal Chain3 Jun 2019 0 repositories listed
-
A user study to compare two conversational assistants designed for people with hearing impairments1 Jun 2019 0 repositories listed
-
Audio De-identification - a New Entity Recognition Task1 Jun 2019 0 repositories listed
-
1 Jun 2019 0 repositories listed