Browse State-of-the-Art › Speech Enhancement › Papers, page 7
Speech Enhancement
Papers archive 2025-07-28
archive papers tagged: 982 · with a code link: 280 · where Syntology ran a sample: 54 (45 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (54 of 982 tagged: 45 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)
Page 7 of 10: papers 601 to 700 of 982, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Time-Domain Speech Enhancement for Robust Automatic Speech Recognition24 Oct 2022 0 repositories listed
-
TridentSE: Guiding Speech Enhancement with 32 Global Tokens24 Oct 2022 0 repositories listed
-
Improved Normalizing Flow-Based Speech Enhancement using an All-pole Gammatone Filterbank for Conditional Input Representation21 Oct 2022 0 repositories listed
-
spatial-dccrn: dccrn equipped with frame-level angle feature and hybrid filtering for multi-channel speech enhancement17 Oct 2022 0 repositories listed
-
Accelerating RNN-based Speech Enhancement on a Multi-Core MCU with Mixed FP16-INT8 Post-Training Quantization14 Oct 2022 0 repositories listed
-
LeVoice ASR Systems for the ISCSLP 2022 Intelligent Cockpit Speech Recognition Challenge14 Oct 2022 0 repositories listed
-
Binaural Speech Enhancement Using STOI-Optimal Masks30 Sep 2022 0 repositories listed
-
Speech Enhancement Using Self-Supervised Pre-Trained Model and Vector Quantization28 Sep 2022 0 repositories listed
-
Speech Enhancement with Perceptually-motivated Optimization and Dual Transformations24 Sep 2022 0 repositories listed
-
GIST-AiTeR System for the Diarization Task of the 2022 VoxCeleb Speaker Recognition Challenge21 Sep 2022 0 repositories listed
-
A Universally-Deployable ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement, and Voice Separation14 Sep 2022 0 repositories listed
-
Multimodal Speech Enhancement Using Burst Propagation7 Sep 2022 0 repositories listed
-
22 Aug 2022 0 repositories listed
-
DNN-Free Low-Latency Adaptive Speech Enhancement Based on Frame-Online Beamforming Powered by Block-Online FastMNMF22 Jul 2022 0 repositories listed
-
Improving spatial cues for hearables using a parameterized binaural CDR estimator17 Jul 2022 0 repositories listed
-
Multi-channel target speech enhancement based on ERB-scaled spatial coherence features17 Jul 2022 0 repositories listed
-
15 Jul 2022 0 repositories listed
-
GLD-Net: Improving Monaural Speech Enhancement by Learning Global and Local Dependency Features with GLD Block30 Jun 2022 0 repositories listed
-
Improving Visual Speech Enhancement Network by Learning Audio-visual Affinity with Multi-head Attention30 Jun 2022 0 repositories listed
-
Challenges and Opportunities in Multi-device Speech Processing27 Jun 2022 0 repositories listed
-
SAQAM: Spatial Audio Quality Assessment Metric24 Jun 2022 0 repositories listed
-
Efficient Transformer-based Speech Enhancement Using Long Frames and STFT Magnitudes23 Jun 2022 0 repositories listed
-
Multi-channel end-to-end neural network for speech enhancement, source localization, and voice activity detection20 Jun 2022 0 repositories listed
-
19 Jun 2022 0 repositories listed
-
NASTAR: Noise Adaptive Speech Enhancement with Target-Conditional Resampling18 Jun 2022 0 repositories listed
-
EPG2S: Speech Generation and Speech Enhancement based on Electropalatography and Audio Signals using Multimodal Learning16 Jun 2022 0 repositories listed
-
To Dereverb Or Not to Dereverb? Perceptual Studies On Real-Time Dereverberation Targets16 Jun 2022 0 repositories listed
-
Canonical Cortical Graph Neural Networks and its Application for Speech Enhancement in Audio-Visual Hearing Aids6 Jun 2022 0 repositories listed
-
Far-Field Speaker Recognition Benchmark Derived From The DiPCo Corpus1 Jun 2022 0 repositories listed
-
Joint Training of Speech Enhancement and Self-supervised Model for Noise-robust ASR26 May 2022 0 repositories listed
-
NeuralEcho: A Self-Attentive Recurrent Neural Network For Unified Acoustic Echo Suppression And Speech Enhancement20 May 2022 0 repositories listed
-
Dictionary-Based Fusion of Contact and Acoustic Microphones for Wind Noise Reduction18 May 2022 0 repositories listed
-
Streaming Noise Context Aware Enhancement For Automatic Speech Recognition in Multi-Talker Environments17 May 2022 0 repositories listed
-
Task splitting for DNN-based acoustic echo and noise removal13 May 2022 0 repositories listed
-
A deep representation learning speech enhancement method using β-VAE11 May 2022 0 repositories listed
-
Generalized Fast Multichannel Nonnegative Matrix Factorization Based on Gaussian Scale Mixtures for Blind Source Separation11 May 2022 0 repositories listed
-
Speaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition9 May 2022 0 repositories listed
-
Acoustic echo suppression using a learning-based multi-frame minimum variance distortionless response filter7 May 2022 0 repositories listed
-
Improving Dual-Microphone Speech Enhancement by Learning Cross-Channel Features with Multi-Head Attention3 May 2022 0 repositories listed
-
On monoaural speech enhancement for automatic recognition of real noisy speech using mixture invariant training3 May 2022 0 repositories listed
-
A Meeting Transcription System for an Ad-Hoc Acoustic Sensor Network2 May 2022 0 repositories listed
-
Improved far-field speech recognition using Joint Variational Autoencoder24 Apr 2022 0 repositories listed
-
RadioSES: mmWave-Based Audioradio Speech Enhancement and Separation System14 Apr 2022 0 repositories listed
-
Listen only to me! How well can target speech extraction handle false alarms?11 Apr 2022 0 repositories listed
-
Expression-preserving face frontalization improves visually assisted speech processing6 Apr 2022 0 repositories listed
-
FFC-SE: Fast Fourier Convolution for Speech Enhancement6 Apr 2022 0 repositories listed
-
Audio-visual multi-channel speech separation, dereverberation and recognition5 Apr 2022 0 repositories listed
-
Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation2 Apr 2022 0 repositories listed
-
End-to-End Integration of Speech Recognition, Speech Enhancement, and Self-Supervised Learning Representation1 Apr 2022 0 repositories listed
-
Multiple Confidence Gates For Joint Training Of SE And ASR1 Apr 2022 0 repositories listed
-
SpecGrad: Diffusion Probabilistic Model based Neural Vocoder with Adaptive Noise Spectral Shaping31 Mar 2022 0 repositories listed
-
Phase-Aware Deep Speech Enhancement: It's All About The Frame Length30 Mar 2022 0 repositories listed
-
MetricGAN+/-: Increasing Robustness of Noise Reduction on Unseen Data23 Mar 2022 0 repositories listed
-
Joint Noise Reduction and Listening Enhancement for Full-End Speech Enhancement22 Mar 2022 0 repositories listed
-
FB-MSTCN: A Full-Band Single-Channel Speech Enhancement Method Based on Multi-Scale Temporal Convolutional Network15 Mar 2022 0 repositories listed
-
Investigating self-supervised learning for speech enhancement and separation15 Mar 2022 0 repositories listed
-
Integrating Statistical Uncertainty into Neural Network-Based Speech Enhancement4 Mar 2022 0 repositories listed
-
PercepNet+: A Phase and SNR Aware PercepNet for Real-Time Speech Enhancement4 Mar 2022 0 repositories listed
-
Phase Continuity: Learning Derivatives of Phase Spectrum for Speech Enhancement24 Feb 2022 0 repositories listed
-
Towards Low-distortion Multi-channel Speech Enhancement: The ESPNet-SE Submission to The L3DAS22 Challenge24 Feb 2022 0 repositories listed
-
The PCG-AIID System for L3DAS22 Challenge: MIMO and MISO convolutional recurrent Network for Multi Channel Speech Enhancement and Speech Recognition21 Feb 2022 0 repositories listed
-
EMGSE: Acoustic/EMG Fusion for Multimodal Speech Enhancement14 Feb 2022 0 repositories listed
-
Low-latency Monaural Speech Enhancement with Deep Filter-bank Equalizer14 Feb 2022 0 repositories listed
-
A Novel Speech Intelligibility Enhancement Model based on CanonicalCorrelation and Deep Learning11 Feb 2022 0 repositories listed
-
Royalflush Speaker Diarization System for ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge10 Feb 2022 0 repositories listed
-
Multimodal Audio-Visual Information Fusion using Canonical-Correlated Graph Neural Network for Energy-Efficient Speech Enhancement9 Feb 2022 0 repositories listed
-
A Speech Intelligibility Enhancement Model based on Canonical Correlation and Deep Learning for Hearing-Assistive Technologies8 Feb 2022 0 repositories listed
-
Optimization of a Real-Time Wavelet-Based Algorithm for Improving Speech Intelligibility5 Feb 2022 0 repositories listed
-
Joint Speech Recognition and Audio Captioning3 Feb 2022 0 repositories listed
-
The RoyalFlush System of Speech Recognition for M2MeT Challenge3 Feb 2022 0 repositories listed
-
The impact of removing head movements on audio-visual speech enhancement1 Feb 2022 0 repositories listed
-
A two-step backward compatible fullband speech enhancement system26 Jan 2022 0 repositories listed
-
A Bayesian Permutation training deep representation learning method for speech enhancement with variational autoencoder24 Jan 2022 0 repositories listed
-
End-to-End Neural Speech Coding for Real-Time Communications24 Jan 2022 0 repositories listed
-
How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR18 Jan 2022 0 repositories listed
-
Learning to Enhance or Not: Neural Network-Based Switching of Enhanced and Observed Signals for Overlapping Speech Recognition11 Jan 2022 0 repositories listed
-
Signal-Aware Direction-of-Arrival Estimation Using Attention Mechanisms3 Jan 2022 0 repositories listed
-
TFCN: Temporal-Frequential Convolutional Network for Single-Channel Speech Enhancement3 Jan 2022 0 repositories listed
-
Towards Robust Real-time Audio-Visual Speech Enhancement16 Dec 2021 0 repositories listed
-
Improving Speech Recognition on Noisy Speech via Speech Enhancement with Multi-Discriminators CycleGAN12 Dec 2021 0 repositories listed
-
Learning-based personal speech enhancement for teleconferencing by exploiting spatial-spectral features10 Dec 2021 0 repositories listed
-
A Training Framework for Stereo-Aware Speech Enhancement using Deep Neural Networks9 Dec 2021 0 repositories listed
-
Harmonic and non-Harmonic Based Noisy Reverberant Speech Enhancement in Time Domain9 Dec 2021 0 repositories listed
-
使用低通時序列語音特徵訓練理想比率遮罩法之語音強化 (Employing Low-Pass Filtered Temporal Speech Features for the Training of Ideal Ratio Mask in Speech Enhancement)1 Dec 2021 0 repositories listed
-
Dataset of Spatial Room Impulse Responses in a Variable Acoustics Room for Six Degrees-of-Freedom Rendering and Analysis23 Nov 2021 0 repositories listed
-
Effect of noise suppression losses on speech distortion and ASR performance23 Nov 2021 0 repositories listed
-
A Conformer-based ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement and Speech Separation18 Nov 2021 0 repositories listed
-
S-DCCRN: Super Wide Band DCCRN with learnable complex feature for speech enhancement16 Nov 2021 0 repositories listed
-
Unsupervised Speech Enhancement with speech recognition embedding and disentanglement losses16 Nov 2021 0 repositories listed
-
Joint Far- and Near-End Speech Intelligibility Enhancement based on the Approximated Speech Intelligibility Index15 Nov 2021 0 repositories listed
-
OSSEM: one-shot speaker adaptive speech enhancement using meta learning10 Nov 2021 0 repositories listed
-
8 Nov 2021 0 repositories listed
-
SEOFP-NET: Compression and Acceleration of Deep Neural Networks for Speech Enhancement Using Sign-Exponent-Only Floating-Points8 Nov 2021 0 repositories listed
-
Deep Noise Suppression Maximizing Non-Differentiable PESQ Mediated by a Non-Intrusive PESQNet6 Nov 2021 0 repositories listed
-
Weight, Block or Unit? Exploring Sparsity Tradeoffs for Speech Enhancement on Tiny Neural Accelerators3 Nov 2021 0 repositories listed
-
Reduction of Subjective Listening Effort for TV Broadcast Signals with Recurrent Neural Networks2 Nov 2021 0 repositories listed
-
SNRi Target Training for Joint Speech Enhancement and Recognition1 Nov 2021 0 repositories listed
-
Cross-attention conformer for context modeling in speech enhancement for ASR30 Oct 2021 0 repositories listed
-
Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction28 Oct 2021 0 repositories listed
-
Closing the Gap Between Time-Domain Multi-Channel Speech Enhancement on Real and Simulation Conditions27 Oct 2021 0 repositories listed