Browse State-of-the-Art › Speech Recognition › Papers, page 37
Speech Recognition
Papers archive 2025-07-28
archive papers tagged: 6,433 · with a code link: 1,373 · where Syntology ran a sample: 196 (162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (196 of 6,433 tagged: 162 with a run with no instrument failure, 34 where every run was a failure of Syntology's instrument)
Page 37 of 65: papers 3,601 to 3,700 of 6,433, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Employing low-pass filtered temporal speech features for the training of ideal ratio mask in speech enhancement1 Oct 2021 0 repositories listed
-
Exploiting Low-Resource Code-Switching Data to Mandarin-English Speech Recognition Systems1 Oct 2021 0 repositories listed
-
Exploring the Integration of E2E ASR and Pronunciation Modeling for English Mispronunciation Detection1 Oct 2021 0 repositories listed
-
Improving Punctuation Restoration for Speech Transcripts via External Data1 Oct 2021 0 repositories listed
-
Incremental Layer-wise Self-Supervised Learning for Efficient Speech Domain Adaptation On Device1 Oct 2021 0 repositories listed
-
Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English with Transfer Learning1 Oct 2021 0 repositories listed
-
SpliceOut: A Simple and Efficient Audio Augmentation Method30 Sep 2021 0 repositories listed
-
Audio Lottery: Speech Recognition Made Ultra-Lightweight, Noise-Robust, and Transferable29 Sep 2021 0 repositories listed
-
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch29 Sep 2021 0 repositories listed
-
Conditioning Sequence-to-sequence Networks with Learned Activations29 Sep 2021 0 repositories listed
-
Demystifying Limited Adversarial Transferability in Automatic Speech Recognition Systems29 Sep 2021 0 repositories listed
-
EP-GAN: Unsupervised Federated Learning with Expectation-Propagation Prior GAN29 Sep 2021 0 repositories listed
-
Google Neural Network Models for Edge Devices: Analyzing and Mitigating Machine Learning Inference Bottlenecks29 Sep 2021 0 repositories listed
-
Learnability of convolutional neural networks for infinite dimensional input via mixed and anisotropic smoothness29 Sep 2021 0 repositories listed
-
MLP-based architecture with variable length input for automatic speech recognition29 Sep 2021 0 repositories listed
-
PhaseFool: Phase-oriented Audio Adversarial Examples via Energy Dissipation29 Sep 2021 0 repositories listed
-
Speech-MLP: a simple MLP architecture for speech processing29 Sep 2021 0 repositories listed
-
Synthesising Audio Adversarial Examples for Automatic Speech Recognition29 Sep 2021 0 repositories listed
-
Understanding the Role of Self Attention for Efficient Speech Recognition29 Sep 2021 0 repositories listed
-
W-CTC: a Connectionist Temporal Classification Loss with Wild Cards29 Sep 2021 0 repositories listed
-
Will a Blind Model Hear Better? Advanced Audiovisual Recognition System with Brain-Like Compensating and Gating29 Sep 2021 0 repositories listed
-
Neural Dependency Coding inspired Multimodal Fusion28 Sep 2021 0 repositories listed
-
Private Language Model Adaptation for Speech Recognition28 Sep 2021 0 repositories listed
-
Word-level confidence estimation for RNN transducers28 Sep 2021 0 repositories listed
-
27 Sep 2021 0 repositories listed
-
Challenges and Opportunities of Speech Recognition for Bengali Language27 Sep 2021 0 repositories listed
-
Topic Model Robustness to Automatic Speech Recognition Errors in Podcast Transcripts25 Sep 2021 0 repositories listed
-
Optimized Power Normalized Cepstral Coefficients towards Robust Deep Speaker Verification24 Sep 2021 0 repositories listed
-
A Lightweight dynamic filter for keyword spotting23 Sep 2021 0 repositories listed
-
Scenario Aware Speech Recognition: Advancements for Apollo Fearless Steps & CHiME-4 Corpora23 Sep 2021 0 repositories listed
-
Animal inspired Application of a Variant of Mel Spectrogram for Seismic Data Processing22 Sep 2021 0 repositories listed
-
Learning Domain Specific Language Models for Automatic Speech Recognition through Machine Translation21 Sep 2021 0 repositories listed
-
On the Difficulty of Segmenting Words with Attention21 Sep 2021 0 repositories listed
-
Audio-Visual Speech Recognition is Worth 32×32×8 Voxels20 Sep 2021 0 repositories listed
-
iRNN: Integer-only Recurrent Neural Network20 Sep 2021 0 repositories listed
-
MeetDot: Videoconferencing with Live Translation Captions20 Sep 2021 0 repositories listed
-
Model-Based Approach for Measuring the Fairness in ASR19 Sep 2021 0 repositories listed
-
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition19 Sep 2021 0 repositories listed
-
Dual-Encoder Architecture with Encoder Selection for Joint Close-Talk and Far-Talk Speech Recognition17 Sep 2021 0 repositories listed
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding17 Sep 2021 0 repositories listed
-
PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription16 Sep 2021 0 repositories listed
-
Utterance-level neural confidence measure for end-to-end children speech recognition16 Sep 2021 0 repositories listed
-
Improving Accent Identification and Accented Speech Recognition Under a Framework of Self-supervised Learning15 Sep 2021 0 repositories listed
-
Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning15 Sep 2021 0 repositories listed
-
LRWR: Large-Scale Benchmark for Lip Reading in Russian language14 Sep 2021 0 repositories listed
-
Non-autoregressive Transformer with Unified Bidirectional Decoder for Automatic Speech Recognition14 Sep 2021 0 repositories listed
-
Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech14 Sep 2021 0 repositories listed
-
Applications of Recurrent Neural Network for Biometric Authentication & Anomaly Detection13 Sep 2021 0 repositories listed
-
A Decidability-Based Loss Function12 Sep 2021 0 repositories listed
-
Unsupervised Domain Adaptation Schemes for Building ASR in Low-resource Languages12 Sep 2021 0 repositories listed
-
Large-vocabulary Audio-visual Speech Recognition in Noisy Environments10 Sep 2021 0 repositories listed
-
Remember the context! ASR slot error correction through memorization10 Sep 2021 0 repositories listed
-
Self-Attention Channel Combinator Frontend for End-to-End Multichannel Far-field Speech Recognition10 Sep 2021 0 repositories listed
-
DeepEMO: Deep Learning for Speech Emotion Recognition9 Sep 2021 0 repositories listed
-
A brief history of AI: how to prevent another winter (a critical review)3 Sep 2021 0 repositories listed
-
Using Topological Framework for the Design of Activation Function and Model Pruning in Deep Neural Networks3 Sep 2021 0 repositories listed
-
Coarse-To-Fine And Cross-Lingual ASR Transfer2 Sep 2021 0 repositories listed
-
Robustness of end-to-end Automatic Speech Recognition Models – A Case Study using Mozilla DeepSpeech1 Sep 2021 0 repositories listed
-
Tree-constrained Pointer Generator for End-to-end Contextual Speech Recognition1 Sep 2021 0 repositories listed
-
30 Aug 2021 0 repositories listed
-
30 Aug 2021 0 repositories listed
-
Multi-Channel Transformer Transducer for Speech Recognition30 Aug 2021 0 repositories listed
-
A Multimodal Framework for Video Ads Understanding29 Aug 2021 0 repositories listed
-
Investigations on Speech Recognition Systems for Low-Resource Dialectal Arabic-English Code-Switching Speech29 Aug 2021 0 repositories listed
-
Goal-driven text descriptions for images28 Aug 2021 0 repositories listed
-
4-bit Quantization of LSTM-based Speech Recognition Models27 Aug 2021 0 repositories listed
-
Exploring Retraining-Free Speech Recognition for Intra-sentential Code-Switching27 Aug 2021 0 repositories listed
-
Full Attention Bidirectional Deep Learning Structure for Single Channel Speech Enhancement27 Aug 2021 0 repositories listed
-
Grammar Based Speaker Role Identification for Air Traffic Control Speech Recognition27 Aug 2021 0 repositories listed
-
Improving callsign recognition with air-surveillance data in air-traffic communication27 Aug 2021 0 repositories listed
-
Injecting Text in Self-Supervised Speech Pretraining27 Aug 2021 0 repositories listed
-
Task-aware Warping Factors in Mask-based Speech Enhancement27 Aug 2021 0 repositories listed
-
Cross-domain Single-channel Speech Enhancement Model with Bi-projection Fusion Module for Noise-robust ASR26 Aug 2021 0 repositories listed
-
Position-Invariant Truecasing with a Word-and-Character Hierarchical Recurrent Neural Network26 Aug 2021 0 repositories listed
-
Graph Neural Networks: Methods, Applications, and Opportunities24 Aug 2021 0 repositories listed
-
Reducing Exposure Bias in Training Recurrent Neural Network Transducers24 Aug 2021 0 repositories listed
-
A Unified Transformer-based Framework for Duplex Text Normalization23 Aug 2021 0 repositories listed
-
Automatic Speech Recognition And Limited Vocabulary: A Survey23 Aug 2021 0 repositories listed
-
Subject Envelope based Multitype Reconstruction Algorithm of Speech Samples of Parkinson's Disease23 Aug 2021 0 repositories listed
-
A Dual-Decoder Conformer for Multilingual Speech Recognition22 Aug 2021 0 repositories listed
-
Generalizing RNN-Transducer to Out-Domain Audio via Sparse Self-Attention Layers22 Aug 2021 0 repositories listed
-
Multilingual Speech Recognition for Low-Resource Indian Languages using Multi-Task conformer22 Aug 2021 0 repositories listed
-
Hierarchical Summarization for Longform Spoken Dialog21 Aug 2021 0 repositories listed
-
A Multi-level Acoustic Feature Extraction Framework for Transformer Based End-to-End Speech Recognition18 Aug 2021 0 repositories listed
-
A Light-weight contextual spelling correction model for customizing transducer-based speech recognition systems17 Aug 2021 0 repositories listed
-
DEXTER: Deep Encoding of External Knowledge for Named Entity Recognition in Virtual Assistants15 Aug 2021 0 repositories listed
-
Multilingual training set selection for ASR in under-resourced Malian languages13 Aug 2021 0 repositories listed
-
Dereverberation of Autoregressive Envelopes for Far-field Speech Recognition12 Aug 2021 0 repositories listed
-
StarGAN-VC+ASR: StarGAN-based Non-Parallel Voice Conversion Regularized by Automatic Speech Recognition10 Aug 2021 0 repositories listed
-
The HW-TSC's Offline Speech Translation Systems for IWSLT 2021 Evaluation9 Aug 2021 0 repositories listed
-
Time-Frequency Localization Using Deep Convolutional Maxout Neural Network in Persian Speech Recognition9 Aug 2021 0 repositories listed
-
An empirical assessment of deep learning approaches to task-oriented dialog management7 Aug 2021 0 repositories listed
-
Spatio-Temporal Attention Mechanism and Knowledge Distillation for Lip Reading7 Aug 2021 0 repositories listed
-
Out-of-Domain Generalization from a Single Source: An Uncertainty Quantification Approach5 Aug 2021 0 repositories listed
-
Blind and neural network-guided convolutional beamformer for joint denoising, dereverberation, and source separation4 Aug 2021 0 repositories listed
-
Dyn-ASR: Compact, Multilingual Speech Recognition via Spoken Language and Accent Identification4 Aug 2021 0 repositories listed
-
Fast frequency modulation is encoded according to the listener expectations in the human subcortical auditory pathway4 Aug 2021 0 repositories listed
-
Improving Distinction between ASR Errors and Speech Disfluencies with Feature Space Interpolation4 Aug 2021 0 repositories listed
-
Spartus: A 9.4 TOp/s FPGA-based LSTM Accelerator Exploiting Spatio-Temporal Sparsity4 Aug 2021 0 repositories listed
-
Unsupervised Domain Adaptation in Speech Recognition using Phonetic Features4 Aug 2021 0 repositories listed