Browse State-of-the-Art › Speech Recognition › Papers, page 38
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 38 of 65: papers 3,701 to 3,800 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
3 Aug 2021 0 repositories listed
-
Bifocal Neural ASR: Exploiting Keyword Spotting for Inference Optimization3 Aug 2021 0 repositories listed
-
Learning a Neural Diff for Speech Models3 Aug 2021 0 repositories listed
-
Adversarial Data Augmentation for Disordered Speech Recognition2 Aug 2021 0 repositories listed
-
Automatic recognition of suprasegmentals in speech2 Aug 2021 0 repositories listed
-
Decoupling recognition and transcription in Mandarin ASR2 Aug 2021 0 repositories listed
-
MOHAQ: Multi-Objective Hardware-Aware Quantization of Recurrent Neural Networks2 Aug 2021 0 repositories listed
-
The Role of Phonetic Units in Speech Emotion Recognition2 Aug 2021 0 repositories listed
-
User-Initiated Repetition-Based Recovery in Multi-Utterance Dialogue Systems2 Aug 2021 0 repositories listed
-
A Speech-enabled Fixed-phrase Translator for Healthcare Accessibility1 Aug 2021 0 repositories listed
-
Automatic generation of a 3D sign language avatar on AR glasses given 2D videos of human signers1 Aug 2021 0 repositories listed
-
Avengers, Ensemble! Benefits of ensembling in grapheme-to-phoneme prediction1 Aug 2021 0 repositories listed
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental1 Aug 2021 0 repositories listed
-
BTS: Back TranScription for Speech-to-Text Post-Processor using Text-to-Speech-to-Text1 Aug 2021 0 repositories listed
-
How Might We Create Better Benchmarks for Speech Recognition?1 Aug 2021 0 repositories listed
-
IMS’ Systems for the IWSLT 2021 Low-Resource Speech Translation Task1 Aug 2021 0 repositories listed
-
Interactive Reinforcement Learning for Table Balancing Robot1 Aug 2021 0 repositories listed
-
基于改进Conformer的新闻领域端到端语音识别(End-to-End Speech Recognition in News Field based on Conformer)1 Aug 2021 0 repositories listed
-
KIT’s IWSLT 2021 Offline Speech Translation System1 Aug 2021 0 repositories listed
-
Multilingual Speech Translation with Unified Transformer: Huawei Noah’s Ark Lab at IWSLT 20211 Aug 2021 0 repositories listed
-
On Knowledge Distillation for Translating Erroneous Speech Transcriptions1 Aug 2021 0 repositories listed
-
ON-TRAC’ systems for the IWSLT 2021 low-resource speech translation and multilingual speech translation shared tasks1 Aug 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus1 Aug 2021 0 repositories listed
-
Technology-Augmented Multilingual Communication Models: New Interaction Paradigms, Shifts in the Language Services Industry, and Implications for Training Programs1 Aug 2021 0 repositories listed
-
Without Further Ado: Direct and Simultaneous Speech Translation by AppTek in 20211 Aug 2021 0 repositories listed
-
ZJU’s IWSLT 2021 Speech Translation System1 Aug 2021 0 repositories listed
-
Can You Hear It? Backdoor Attacks via Ultrasonic Triggers30 Jul 2021 0 repositories listed
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition29 Jul 2021 0 repositories listed
-
An Adapter Based Pre-Training for Efficient and Scalable Self-Supervised Speech Representation Learning26 Jul 2021 0 repositories listed
-
Facetron: A Multi-speaker Face-to-Speech Model based on Cross-modal Latent Representations26 Jul 2021 0 repositories listed
-
23 Jul 2021 0 repositories listed
-
Using Deep Learning Techniques and Inferential Speech Statistics for AI Synthesised Speech Recognition23 Jul 2021 0 repositories listed
-
CarneliNet: Neural Mixture Model for Automatic Speech Recognition22 Jul 2021 0 repositories listed
-
Multitask-Based Joint Learning Approach To Robust ASR For Radio Communication Speech22 Jul 2021 0 repositories listed
-
Semantic Communications for Speech Recognition22 Jul 2021 0 repositories listed
-
On Prosody Modeling for ASR+TTS based Voice Conversion20 Jul 2021 0 repositories listed
-
Seed Words Based Data Selection for Language Model Adaptation20 Jul 2021 0 repositories listed
-
A baseline model for computationally inexpensive speech recognition for Kazakh using the Coqui STT framework19 Jul 2021 0 repositories listed
-
Multi-task Learning with Cross Attention for Keyword Spotting15 Jul 2021 0 repositories listed
-
VAD-free Streaming Hybrid CTC/Attention ASR for Unsegmented Recording15 Jul 2021 0 repositories listed
-
A Configurable Multilingual Model is All You Need to Recognize All Languages13 Jul 2021 0 repositories listed
-
Conformer-based End-to-end Speech Recognition With Rotary Position Embedding13 Jul 2021 0 repositories listed
-
The IWSLT 2021 BUT Speech Translation Systems13 Jul 2021 0 repositories listed
-
Zero-shot Speech Translation13 Jul 2021 0 repositories listed
-
Perceptual-based deep-learning denoiser as a defense against adversarial attacks on ASR systems12 Jul 2021 0 repositories listed
-
UniSpeech at scale: An Empirical Study of Pre-training Method on Large-Scale Speech Recognition Dataset12 Jul 2021 0 repositories listed
-
Loss Prediction: End-to-End Active Learning Approach For Speech Recognition9 Jul 2021 0 repositories listed
-
Noisy Training Improves E2E ASR for the Edge9 Jul 2021 0 repositories listed
-
On lattice-free boosted MMI training of HMM and CTC-based full-context ASR models9 Jul 2021 0 repositories listed
-
Representation Learning to Classify and Detect Adversarial Attacks against Speaker and Speech Recognition Systems9 Jul 2021 0 repositories listed
-
Improved Language Identification Through Cross-Lingual Self-Supervised Learning8 Jul 2021 0 repositories listed
-
End-to-End Rich Transcription-Style Automatic Speech Recognition with Semi-Supervised Learning7 Jul 2021 0 repositories listed
-
Improving Speech Recognition Accuracy of Local POI Using Geographical Models7 Jul 2021 0 repositories listed
-
A Comparative Study of Modular and Joint Approaches for Speaker-Attributed ASR on Monaural Long-Form Audio6 Jul 2021 0 repositories listed
-
Exploiting Single-Channel Speech For Multi-channel End-to-end Speech Recognition6 Jul 2021 0 repositories listed
-
Improving a neural network model by explanation-guided training for glioma classification based on MRI data5 Jul 2021 0 repositories listed
-
Investigation of Practical Aspects of Single Channel Speech Separation for ASR5 Jul 2021 0 repositories listed
-
Arabic Code-Switching Speech Recognition using Monolingual Data4 Jul 2021 0 repositories listed
-
Cross-Modal Transformer-Based Neural Correction Models for Automatic Speech Recognition4 Jul 2021 0 repositories listed
-
Unified Autoregressive Modeling for Joint End-to-End Multi-Talker Overlapped Speech Recognition and Speaker Attribute Estimation4 Jul 2021 0 repositories listed
-
Dual Causal/Non-Causal Self-Attention for Streaming End-to-End Speech Recognition2 Jul 2021 0 repositories listed
-
Multi-user VoiceFilter-Lite via Attentive Speaker Embedding2 Jul 2021 0 repositories listed
-
Supervised Contrastive Learning for Accented Speech Recognition2 Jul 2021 0 repositories listed
-
ESPnet-ST IWSLT 2021 Offline Speech Translation System1 Jul 2021 0 repositories listed
-
Improving Named Entity Recognition in Spoken Dialog Systems by Context and Speech Pattern Modeling1 Jul 2021 0 repositories listed
-
Interactive decoding of words from visual speech recognition models1 Jul 2021 0 repositories listed
-
Projection of Turn Completion in Incremental Spoken Dialogue Systems1 Jul 2021 0 repositories listed
-
SmarTerp: A CAI System to Support Simultaneous Interpreters in Real-Time1 Jul 2021 0 repositories listed
-
StableEmit: Selection Probability Discount for Reducing Emission Latency of Streaming Monotonic Attention ASR1 Jul 2021 0 repositories listed
-
Word-Free Spoken Language Understanding for Mandarin-Chinese1 Jul 2021 0 repositories listed
-
On joint training with interfaces for spoken language understanding30 Jun 2021 0 repositories listed
-
IMS' Systems for the IWSLT 2021 Low-Resource Speech Translation Task30 Jun 2021 0 repositories listed
-
Sequence-level Confidence Classifier for ASR Utterance Accuracy and Application to Acoustic Models30 Jun 2021 0 repositories listed
-
Rethinking End-to-End Evaluation of Decomposable Tasks: A Case Study on Spoken Language Understanding29 Jun 2021 0 repositories listed
-
On a novel training algorithm for sequence-to-sequence predictive recurrent networks27 Jun 2021 0 repositories listed
-
Use of Machine Learning Technique to maximize the signal over background for H →ττ27 Jun 2021 0 repositories listed
-
Building Intelligent Autonomous Navigation Agents25 Jun 2021 0 repositories listed
-
Lexical Access Model for Italian -- Modeling human speech processing: identification of words in running speech toward lexical access based on the detection of landmarks and other acoustic cues to features24 Jun 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus24 Jun 2021 0 repositories listed
-
Where are we in semantic concept extraction for Spoken Language Understanding?24 Jun 2021 0 repositories listed
-
Mixtures of Deep Neural Experts for Automated Speech Scoring23 Jun 2021 0 repositories listed
-
Zero-Shot Joint Modeling of Multiple Spoken-Text-Style Conversion Tasks using Switching Tokens23 Jun 2021 0 repositories listed
-
A Discriminative Entity-Aware Language Model for Virtual Assistants21 Jun 2021 0 repositories listed
-
How to Reach Real-Time AI on Consumer Devices? Solutions for Programmable and Custom Architectures21 Jun 2021 0 repositories listed
-
Pay Better Attention to Attention: Head Selection in Multilingual and Multi-Domain Sequence Modeling21 Jun 2021 0 repositories listed
-
Learning From the Master: Distilling Cross-Modal Advanced Knowledge for Lip Reading19 Jun 2021 0 repositories listed
-
An Improved Single Step Non-autoregressive Transformer for Automatic Speech Recognition18 Jun 2021 0 repositories listed
-
Analysis and Tuning of a Voice Assistant System for Dysfluent Speech18 Jun 2021 0 repositories listed
-
Low Resource German ASR with Untranscribed Data Spoken by Non-native Children -- INTERSPEECH 2021 Shared Task SPAPL System18 Jun 2021 0 repositories listed
-
On-Device Personalization of Automatic Speech Recognition Models for Disordered Speech18 Jun 2021 0 repositories listed
-
Layer Pruning on Demand with Intermediate CTC17 Jun 2021 0 repositories listed
-
Multi-mode Transformer Transducer with Stochastic Future Context17 Jun 2021 0 repositories listed
-
Best Practices for Noise-Based Augmentation to Improve the Performance of Deployable Speech-Based Emotion Recognition Systems16 Jun 2021 0 repositories listed
-
Collaborative Training of Acoustic Encoders for Speech Recognition16 Jun 2021 0 repositories listed
-
Topic Classification on Spoken Documents Using Deep Acoustic and Linguistic Features16 Jun 2021 0 repositories listed
-
A Study into Pre-training Strategies for Spoken Language Understanding on Dysarthric Speech15 Jun 2021 0 repositories listed
-
ASR Adaptation for E-commerce Chatbots using Cross-Utterance Context and Multi-Task Language Modeling15 Jun 2021 0 repositories listed
-
Dialectal Speech Recognition and Translation of Swiss German Speech to Standard German Text: Microsoft's Submission to SwissText 202115 Jun 2021 0 repositories listed
-
E2E-based Multi-task Learning Approach to Joint Speech and Accent Recognition15 Jun 2021 0 repositories listed
-
Multi-channel Opus compression for far-field automatic speech recognition with a fixed bitrate budget15 Jun 2021 0 repositories listed