Browse State-of-the-Art › Speech Recognition › Papers, page 53
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 53 of 65: papers 5,201 to 5,300 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Automatic context window composition for distant speech recognition26 May 2018 0 repositories listed
-
Geometric Understanding of Deep Learning26 May 2018 0 repositories listed
-
Task-dependent modulation of the visual sensory thalamus assists visual-speech recognition24 May 2018 0 repositories listed
-
ASR-based Features for Emotion Recognition: A Transfer Learning Approach23 May 2018 0 repositories listed
-
Unsupervised Cross-Modal Alignment of Speech and Text Embedding Spaces18 May 2018 0 repositories listed
-
A Comparison of Modeling Units in Sequence-to-Sequence Speech Recognition with the Transformer on Mandarin Chinese16 May 2018 0 repositories listed
-
Composing Finite State Transducers on GPUs16 May 2018 0 repositories listed
-
A Purely End-to-end System for Multi-speaker Speech Recognition15 May 2018 0 repositories listed
-
Improved ASR for Under-Resourced Languages Through Multi-Task Learning with Acoustic Landmarks15 May 2018 0 repositories listed
-
Gradient-Leaks: Understanding and Controlling Deanonymization in Federated Learning15 May 2018 0 repositories listed
-
A comparable study of modeling units for end-to-end Mandarin speech recognition10 May 2018 0 repositories listed
-
Transfer Learning from Adult to Children for Speech Recognition: Evaluation, Analysis and Recommendations8 May 2018 0 repositories listed
-
Boosting Noise Robustness of Acoustic Model via Deep Adversarial Training2 May 2018 0 repositories listed
-
A Comparative Study of Extremely Low-Resource Transliteration of the World's Languages1 May 2018 0 repositories listed
-
A Leveled Reading Corpus of Modern Standard Arabic1 May 2018 0 repositories listed
-
A Real-life, French-accented Corpus of Air Traffic Control Communications1 May 2018 0 repositories listed
-
A Semi-autonomous System for Creating a Human-Machine Interaction Corpus in Virtual Reality: Application to the ACORFORMed System for Training Doctors to Break Bad News1 May 2018 0 repositories listed
-
A Vietnamese Dialog Act Corpus Based on ISO 24617-2 standard1 May 2018 0 repositories listed
-
A Web Service for Pre-segmenting Very Long Transcribed Speech Recordings1 May 2018 0 repositories listed
-
An Application for Building a Polish Telephone Speech Corpus1 May 2018 0 repositories listed
-
ASR for Documenting Acutely Under-Resourced Indigenous Languages1 May 2018 0 repositories listed
-
Building Open Javanese and Sundanese Corpora for Multilingual Text-to-Speech1 May 2018 0 repositories listed
-
Classification of Closely Related Sub-dialects of Arabic Using Support-Vector Machines1 May 2018 0 repositories listed
-
Collecting Code-Switched Data from Social Media1 May 2018 0 repositories listed
-
Collection and Analysis of Code-switch Egyptian Arabic-English Speech Corpus1 May 2018 0 repositories listed
-
Construction of English-French Multimodal Affective Conversational Corpus from TV Dramas1 May 2018 0 repositories listed
-
Contextual Dependencies in Time-Continuous Multidimensional Affect Recognition1 May 2018 0 repositories listed
-
CPJD Corpus: Crowdsourced Parallel Speech Corpus of Japanese Dialects1 May 2018 0 repositories listed
-
Creating Lithuanian and Latvian Speech Corpora from Inaccurately Annotated Web Data1 May 2018 0 repositories listed
-
DART: A Large Dataset of Dialectal Arabic Tweets1 May 2018 0 repositories listed
-
Data-Driven Pronunciation Modeling of Swiss German Dialectal Speech for Automatic Speech Recognition1 May 2018 0 repositories listed
-
Design and Development of Speech Corpora for Air Traffic Control Training1 May 2018 0 repositories listed
-
Discovering Canonical Indian English Accents: A Crowdsourcing-based Approach1 May 2018 0 repositories listed
-
Evaluation of Feature-Space Speaker Adaptation for End-to-End Acoustic Models1 May 2018 0 repositories listed
-
FARMI: A FrAmework for Recording Multi-Modal Interactions1 May 2018 0 repositories listed
-
From `Solved Problems' to New Challenges: A Report on LDC Activities1 May 2018 0 repositories listed
-
Improved Transcription and Indexing of Oral History Interviews for Digital Humanities Research1 May 2018 0 repositories listed
-
Increasing the Accessibility of Time-Aligned Speech Corpora with Spokes Mix1 May 2018 0 repositories listed
-
Matics Software Suite: New Tools for Evaluation and Data Exploration1 May 2018 0 repositories listed
-
MirasVoice: A bilingual (English-Persian) speech corpus1 May 2018 0 repositories listed
-
MOCCA: Measure of Confidence for Corpus Analysis - Automatic Reliability Check of Transcript and Automatic Segmentation1 May 2018 0 repositories listed
-
Multilingual Parallel Corpus for Global Communication Plan1 May 2018 0 repositories listed
-
Neural Caption Generation for News Images1 May 2018 0 repositories listed
-
Open ASR for Icelandic: Resources and a Baseline System1 May 2018 0 repositories listed
-
Parallel Corpora in Mboshi (Bantu C25, Congo-Brazzaville)1 May 2018 0 repositories listed
-
Phonetically Balanced Code-Mixed Speech Corpus for Hindi-English Automatic Speech Recognition1 May 2018 0 repositories listed
-
Pronunciation Variants and ASR of Colloquial Speech: A Case Study on Czech1 May 2018 0 repositories listed
-
Simulating ASR errors for training SLU systems1 May 2018 0 repositories listed
-
Speech Rate Calculations with Short Utterances: A Study from a Speech-to-Speech, Machine Translation Mediated Map Task1 May 2018 0 repositories listed
-
Text Normalization Infrastructure that Scales to Hundreds of Language Varieties1 May 2018 0 repositories listed
-
The WAW Corpus: The First Corpus of Interpreted Speeches and their Translations for English and Arabic1 May 2018 0 repositories listed
-
Towards an Automatic Assessment of Crowdsourced Data for NLU1 May 2018 0 repositories listed
-
Towards Processing of the Oral History Interviews and Related Printed Documents1 May 2018 0 repositories listed
-
Unified Guidelines and Resources for Arabic Dialect Orthography1 May 2018 0 repositories listed
-
Using Discourse Information for Education with a Spanish-Chinese Parallel Corpus1 May 2018 0 repositories listed
-
VAST: A Corpus of Video Annotation for Speech Technologies1 May 2018 0 repositories listed
-
Automatic Documentation of ICD Codes with Far-Field Speech Recognition30 Apr 2018 0 repositories listed
-
Investigations on End-to-End Audiovisual Fusion30 Apr 2018 0 repositories listed
-
Sparse Persistent RNNs: Squeezing Large Recurrent Networks On-Chip26 Apr 2018 0 repositories listed
-
End-to-End Multimodal Speech Recognition25 Apr 2018 0 repositories listed
-
Recent Progresses in Deep Learning based Acoustic Models (Updated)25 Apr 2018 0 repositories listed
-
An Information-Theoretic View for Deep Learning24 Apr 2018 0 repositories listed
-
Automatic speech recognition for launch control center communication using recurrent neural networks with data augmentation and custom language model24 Apr 2018 0 repositories listed
-
Multi-Head Decoder for End-to-End Speech Recognition22 Apr 2018 0 repositories listed
-
Neural Compatibility Modeling with Attentive Knowledge Distillation17 Apr 2018 0 repositories listed
-
Precise Detection of Speech Endpoints Dynamically: A Wavelet Convolution based approach17 Apr 2018 0 repositories listed
-
15 Apr 2018 0 repositories listed
-
Language Recognition using Time Delay Deep Neural Network13 Apr 2018 0 repositories listed
-
Global SNR Estimation of Speech Signals using Entropy and Uncertainty Estimates from Dropout Networks12 Apr 2018 0 repositories listed
-
Vision as an Interlingua: Learning Multilingual Semantic Embeddings of Untranscribed Speech9 Apr 2018 0 repositories listed
-
Joint Learning of Interactive Spoken Content Retrieval and Trainable User Simulator1 Apr 2018 0 repositories listed
-
ESPnet: End-to-End Speech Processing Toolkit30 Mar 2018 0 repositories listed
-
Towards Unsupervised Automatic Speech Recognition Trained by Unaligned Speech and Text only29 Mar 2018 0 repositories listed
-
Machine Speech Chain with One-shot Speaker Adaptation28 Mar 2018 0 repositories listed
-
28 Mar 2018 0 repositories listed
-
A Multi-Discriminator CycleGAN for Unsupervised Non-Parallel Speech Domain Adaptation27 Mar 2018 0 repositories listed
-
27 Mar 2018 0 repositories listed
-
Comprehending Real Numbers: Development of Bengali Real Number Speech Corpus27 Mar 2018 0 repositories listed
-
Multi-Modal Data Augmentation for End-to-End ASR27 Mar 2018 0 repositories listed
-
Student-Teacher Learning for BLSTM Mask-based Speech Enhancement27 Mar 2018 0 repositories listed
-
Clipping free attacks against artificial neural networks26 Mar 2018 0 repositories listed
-
Spectral feature mapping with mimic loss for robust speech recognition26 Mar 2018 0 repositories listed
-
Low-Resource Speech-to-Text Translation24 Mar 2018 0 repositories listed
-
Exploring the robustness of features and enhancement on speech recognition systems in highly-reverberant real environments23 Mar 2018 0 repositories listed
-
Highly-Reverberant Real Environment database: HRRE23 Mar 2018 0 repositories listed
-
Acoustic feature learning using cross-domain articulatory measurements19 Mar 2018 0 repositories listed
-
Speech to text and text to speech recognition systems-Areview17 Mar 2018 0 repositories listed
-
TBD: Benchmarking and Analyzing Deep Neural Network Training16 Mar 2018 0 repositories listed
-
Advancing Connectionist Temporal Classification With Attention Modeling15 Mar 2018 0 repositories listed
-
Resource aware design of a deep convolutional-recurrent neural network for speech recognition through audio-visual sensor fusion13 Mar 2018 0 repositories listed
-
From Nodes to Networks: Evolving Recurrent Neural Networks12 Mar 2018 0 repositories listed
-
Speech Recognition: Keyword Spotting Through Image Recognition10 Mar 2018 0 repositories listed
-
Extracting Domain Invariant Features by Unsupervised Learning for Robust Automatic Speech Recognition7 Mar 2018 0 repositories listed
-
On Modular Training of Neural Acoustics-to-Word Model for LVCSR3 Mar 2018 0 repositories listed
-
The History Began from AlexNet: A Comprehensive Survey on Deep Learning Approaches3 Mar 2018 0 repositories listed
-
Challenges in Speech Recognition and Translation of High-Value Low-Density Polysynthetic Languages1 Mar 2018 0 repositories listed
-
Evaluating Automatic Speech Recognition in Translation1 Mar 2018 0 repositories listed
-
Neural Network Methods for Natural Language Processing by Yoav Goldberg1 Mar 2018 0 repositories listed
-
On the Derivational Entropy of Left-to-Right Probabilistic Finite-State Automata and Hidden Markov Models1 Mar 2018 0 repositories listed
-
Portable Speech-to-Speech Translation on an Android Smartphone: The MFLTS System1 Mar 2018 0 repositories listed