Browse State-of-the-Art › Speech Recognition › Papers, page 47
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 47 of 65: papers 4,601 to 4,700 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Privacy-Preserving Adversarial Representation Learning in ASR: Reality or Illusion?12 Nov 2019 0 repositories listed
-
Data Efficient Direct Speech-to-Text Translation with Modality Agnostic Meta-Learning11 Nov 2019 0 repositories listed
-
Long-span language modeling for speech recognition11 Nov 2019 0 repositories listed
-
Evaluating Voice Conversion-based Privacy Protection against Informed Attackers10 Nov 2019 0 repositories listed
-
Multimodal Intelligence: Representation Learning, Information Fusion, and Applications10 Nov 2019 0 repositories listed
-
Listen and Fill in the Missing Letters: Non-Autoregressive Transformer for Speech Recognition10 Nov 2019 0 repositories listed
-
Speaker Adaptation for Attention-Based End-to-End Speech Recognition9 Nov 2019 0 repositories listed
-
8 Nov 2019 0 repositories listed
-
Investigation of Error Simulation Techniques for Learning Dialog Policies for Conversational Error Recovery8 Nov 2019 0 repositories listed
-
Boosting LSTM Performance Through Dynamic Precision Selection7 Nov 2019 0 repositories listed
-
A comparison of end-to-end models for long-form speech recognition6 Nov 2019 0 repositories listed
-
RNN-T For Latency Controlled ASR With Improved Beam Search5 Nov 2019 0 repositories listed
-
SHARP: An Adaptable, Energy-Efficient Accelerator for Recurrent Neural Network4 Nov 2019 0 repositories listed
-
Supervised level-wise pretraining for recurrent neural network initialization in multi-class classification4 Nov 2019 0 repositories listed
-
Chameleon: A Language Model Adaptation Toolkit for Automatic Speech Recognition of Conversational Speech1 Nov 2019 0 repositories listed
-
Data Augmentation for End-to-End Speech Translation: FBK@IWSLT ‘191 Nov 2019 0 repositories listed
-
Entity resolution for noisy ASR transcripts1 Nov 2019 0 repositories listed
-
Improving Generalization of Transformer for Speech Recognition with Parallel Schedule Sampling and Relative Positional Embedding1 Nov 2019 0 repositories listed
-
Long-distance Detection of Bioacoustic Events with Per-channel Energy Normalization1 Nov 2019 0 repositories listed
-
Multi-Task Modeling of Phonographic Languages: Translating Middle Egyptian Hieroglyphs1 Nov 2019 0 repositories listed
-
PyOpenDial: A Python-based Domain-Independent Toolkit for Developing Spoken Dialogue Systems with Probabilistic Rules1 Nov 2019 0 repositories listed
-
The IWSLT 2019 KIT Speech Translation System1 Nov 2019 0 repositories listed
-
A neural document language modeling framework for spoken document retrieval31 Oct 2019 0 repositories listed
-
Multi-scale Octave Convolutions for Robust Speech Recognition31 Oct 2019 0 repositories listed
-
Lightweight and Efficient End-to-End Speech Recognition Using Low-Rank Transformer30 Oct 2019 0 repositories listed
-
Transformer-based Cascaded Multimodal Speech Translation29 Oct 2019 0 repositories listed
-
Improving sequence-to-sequence speech recognition training with on-the-fly data augmentation29 Oct 2019 0 repositories listed
-
Does Speech enhancement of publicly available data help build robust Speech Recognition Systems?29 Oct 2019 0 repositories listed
-
LeanConvNets: Low-cost Yet Effective Convolutional Neural Networks29 Oct 2019 0 repositories listed
-
DFSMN-SAN with Persistent Memory Model for Automatic Speech Recognition28 Oct 2019 0 repositories listed
-
Towards Unsupervised Speech Recognition and Synthesis with Quantized Speech Representation Learning28 Oct 2019 0 repositories listed
-
Unsupervised pre-training for sequence to sequence speech recognition28 Oct 2019 0 repositories listed
-
Training ASR models by Generation of Contextual Information27 Oct 2019 0 repositories listed
-
Meta Learning for End-to-End Low-Resource Speech Recognition26 Oct 2019 0 repositories listed
-
L2RS: A Learning-to-Rescore Mechanism for Automatic Speech Recognition25 Oct 2019 0 repositories listed
-
25 Oct 2019 0 repositories listed
-
Towards Online End-to-end Transformer Automatic Speech Recognition25 Oct 2019 0 repositories listed
-
A Bayesian Approach to Recurrence in Neural Networks24 Oct 2019 0 repositories listed
-
An Empirical Study of Efficient ASR Rescoring with Transformers24 Oct 2019 0 repositories listed
-
Pre-training in Deep Reinforcement Learning for Automatic Speech Recognition24 Oct 2019 0 repositories listed
-
Recognizing long-form speech using streaming end-to-end models24 Oct 2019 0 repositories listed
-
A practical two-stage training strategy for multi-stream end-to-end speech recognition23 Oct 2019 0 repositories listed
-
Analyzing ASR pretraining for low-resource speech-to-text translation23 Oct 2019 0 repositories listed
-
Correction of Automatic Speech Recognition with Transformer Sequence-to-sequence Model23 Oct 2019 0 repositories listed
-
Efficient Dynamic WFST Decoding for Personalized Language Models23 Oct 2019 0 repositories listed
-
RNN based Incremental Online Spoken Language Understanding23 Oct 2019 0 repositories listed
-
G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR22 Oct 2019 0 repositories listed
-
Robust Neural Machine Translation for Clean and Noisy Speech Transcripts22 Oct 2019 0 repositories listed
-
22 Oct 2019 0 repositories listed
-
AeGAN: Time-Frequency Speech Denoising via Generative Adversarial Networks21 Oct 2019 0 repositories listed
-
Signal Combination for Language Identification21 Oct 2019 0 repositories listed
-
Neuro-SERKET: Development of Integrative Cognitive System through the Composition of Deep Probabilistic Generative Models20 Oct 2019 0 repositories listed
-
Predicting ice flow using machine learning20 Oct 2019 0 repositories listed
-
End-to-End Speech Recognition: A review for the French Language18 Oct 2019 0 repositories listed
-
Detecting Multiple Speech Disfluencies using a Deep Residual Network with Bidirectional Long Short-Term Memory17 Oct 2019 0 repositories listed
-
Multi-Talker MVDR Beamforming Based on Extended Complex Gaussian Mixture Model17 Oct 2019 0 repositories listed
-
Lead2Gold: Towards exploiting the full potential of noisy transcriptions for speech recognition16 Oct 2019 0 repositories listed
-
Transformer ASR with Contextual Block Processing16 Oct 2019 0 repositories listed
-
Analyzing Large Receptive Field Convolutional Networks for Distant Speech Recognition15 Oct 2019 0 repositories listed
-
MIMO-SPEECH: End-to-End Multi-Channel Multi-Speaker Speech Recognition15 Oct 2019 0 repositories listed
-
Transfer Learning for Algorithm Recommendation15 Oct 2019 0 repositories listed
-
A Research Platform for Multi-Robot Dialogue with Humans12 Oct 2019 0 repositories listed
-
VAIS ASR: Building a conversational speech recognition system using language model combination12 Oct 2019 0 repositories listed
-
Hear "No Evil", See "Kenansville": Efficient and Transferable Black-Box Attacks on Speech Recognition and Voice Identification Systems11 Oct 2019 0 repositories listed
-
Query-by-example on-device keyword spotting11 Oct 2019 0 repositories listed
-
One-To-Many Multilingual End-to-end Speech Translation8 Oct 2019 0 repositories listed
-
A Case Study on Combining ASR and Visual Features for Generating Instructional Video Captions7 Oct 2019 0 repositories listed
-
Adapting a FrameNet Semantic Parser for Spoken Language Understanding Using Adversarial Learning7 Oct 2019 0 repositories listed
-
Modeling Confidence in Sequence-to-Sequence Models4 Oct 2019 0 repositories listed
-
SNDCNN: Self-normalizing deep CNNs with scaled exponential linear units for speech recognition4 Oct 2019 0 repositories listed
-
Convolutional Neural Networks for Speech Controlled Prosthetic Hands3 Oct 2019 0 repositories listed
-
Neural Zero-Inflated Quality Estimation Model For Automatic Speech Recognition System3 Oct 2019 0 repositories listed
-
From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition2 Oct 2019 0 repositories listed
-
Additional Shared Decoder on Siamese Multi-view Encoders for Learning Acoustic Word Embeddings1 Oct 2019 0 repositories listed
-
Domain Expansion in DNN-based Acoustic Models for Robust Speech Recognition1 Oct 2019 0 repositories listed
-
室內遠距離語音辨識實驗(Experiments on In-House Far-Field Speech Recognition)1 Oct 2019 0 repositories listed
-
使用生成對抗網路於強健式自動語音辨識的應用(Exploiting Generative Adversarial Network for Robustness Automatic Speech Recognition)1 Oct 2019 0 repositories listed
-
1 Oct 2019 0 repositories listed
-
探究端對端語音辨識於發音檢測與診斷(Investigating on Computer-Assisted Pronunciation Training Leveraging End-to-End Speech Recognition Techniques)1 Oct 2019 0 repositories listed
-
Neural Hybrid Recommender: Recommendation needs collaboration29 Sep 2019 0 repositories listed
-
Self-Attention Transducers for End-to-End Speech Recognition28 Sep 2019 0 repositories listed
-
End-to-End Code-Switching ASR for Low-Resourced Language Pairs27 Sep 2019 0 repositories listed
-
Optimizing Speech Recognition For The Edge26 Sep 2019 0 repositories listed
-
AdaScale SGD: A Scale-Invariant Algorithm for Distributed Training25 Sep 2019 0 repositories listed
-
Breaking the Data Barrier: Towards Robust Speech Translation via Adversarial Stability Training25 Sep 2019 0 repositories listed
-
Generating Robust Audio Adversarial Examples using Iterative Proportional Clipping25 Sep 2019 0 repositories listed
-
Improved Training Techniques for Online Neural Machine Translation25 Sep 2019 0 repositories listed
-
Self-Supervised Speech Recognition via Local Prior Matching25 Sep 2019 0 repositories listed
-
Speech Recognition with Augmented Synthesized Speech25 Sep 2019 0 repositories listed
-
Top-down training for neural networks25 Sep 2019 0 repositories listed
-
Unsupervised Learning of Efficient and Robust Speech Representations25 Sep 2019 0 repositories listed
-
Understanding Semantics from Speech Through Pre-training24 Sep 2019 0 repositories listed
-
Improving OOV Detection and Resolution with External Language Models in Acoustic-to-Word ASR22 Sep 2019 0 repositories listed
-
Persian Signature Verification using Fully Convolutional Networks20 Sep 2019 0 repositories listed
-
A Comparison of Hybrid and End-to-End Models for Syllable Recognition19 Sep 2019 0 repositories listed
-
A Random Gossip BMUF Process for Neural Language Modeling19 Sep 2019 0 repositories listed
-
Self-Training for End-to-End Speech Recognition19 Sep 2019 0 repositories listed
-
Code-Switched Language Models Using Neural Based Synthetic Data from Parallel Sentences18 Sep 2019 0 repositories listed
-
Emotion Filtering at the Edge18 Sep 2019 0 repositories listed
-
Simultaneous Speech Recognition and Speaker Diarization for Monaural Dialogue Recordings with Target-Speaker Acoustic Models17 Sep 2019 0 repositories listed