Browse State-of-the-Art › Speech Enhancement › Papers, page 5
Speech Enhancement
Papers archive 2025-07-28
archive papers tagged: 982 · with a code link: 280 · where Syntology ran a sample: 54 (45 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (54 of 982 tagged: 45 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)
Page 5 of 10: papers 401 to 500 of 982, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
The Effect of Training Dataset Size on Discriminative and Diffusion-Based Speech Enhancement Systems10 Jun 2024 0 repositories listed
-
Thunder : Unified Regression-Diffusion Speech Enhancement with a Single Reverse Step using Brownian Bridge10 Jun 2024 0 repositories listed
-
An Investigation of Noise Robustness for Flow-Matching-Based Zero-Shot TTS9 Jun 2024 0 repositories listed
-
URGENT Challenge: Universality, Robustness, and Generalizability For Speech Enhancement7 Jun 2024 0 repositories listed
-
Flexible Multichannel Speech Enhancement for Noise-Robust Frontend6 Jun 2024 0 repositories listed
-
Helsinki Speech Challenge 20246 Jun 2024 0 repositories listed
-
PLDNet: PLD-Guided Lightweight Deep Network Boosted by Efficient Attention for Handheld Dual-Microphone Speech Enhancement6 Jun 2024 0 repositories listed
-
Reference Channel Selection by Multi-Channel Masking for End-to-End Multi-Channel Speech Enhancement5 Jun 2024 0 repositories listed
-
5 Jun 2024 0 repositories listed
-
Speech enhancement deep-learning architecture for efficient edge processing27 May 2024 0 repositories listed
-
Non-autoregressive real-time Accent Conversion model with voice cloning21 May 2024 0 repositories listed
-
Building a Luganda Text-to-Speech Model From Crowdsourced Data16 May 2024 0 repositories listed
-
Monaural speech enhancement on drone via Adapter based transfer learning16 May 2024 0 repositories listed
-
Evaluating Speech Enhancement Systems Through Listening Effort13 May 2024 0 repositories listed
-
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios3 May 2024 0 repositories listed
-
TRAMBA: A Hybrid Transformer and Mamba Architecture for Practical Audio and Bone Conduction Speech Super Resolution and Enhancement on Mobile and Wearable Platforms2 May 2024 0 repositories listed
-
Deep low-latency joint speech transmission and enhancement over a gaussian channel30 Apr 2024 0 repositories listed
-
Rethinking Processing Distortions: Disentangling the Impact of Speech Enhancement Errors on Speech Recognition Performance23 Apr 2024 0 repositories listed
-
Exploring the Potential of Data-Driven Spatial Audio Enhancement Using a Single-Channel Model22 Apr 2024 0 repositories listed
-
TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition19 Apr 2024 0 repositories listed
-
Efficient High-Performance Bark-Scale Neural Network for Residual Echo and Noise Suppression8 Apr 2024 0 repositories listed
-
Artificial Intelligence for Cochlear Implants: Review of Strategies, Challenges, and Perspectives17 Mar 2024 0 repositories listed
-
SuperM2M: Supervised and Mixture-to-Mixture Co-Learning for Speech Enhancement and Noise-Robust ASR15 Mar 2024 0 repositories listed
-
A Closer Look at Wav2Vec2 Embeddings for On-Device Single-Channel Speech Enhancement3 Mar 2024 0 repositories listed
-
Exploration of Adapter for Noise Robust Automatic Speech Recognition28 Feb 2024 0 repositories listed
-
Audio-Visual Speech Enhancement in Noisy Environments via Emotion-Based Contextual Cues26 Feb 2024 0 repositories listed
-
SICRN: Advancing Speech Enhancement through State Space Model and Inplace Convolution Techniques22 Feb 2024 0 repositories listed
-
Mel-FullSubNet: Mel-Spectrogram Enhancement for Improving Both Speech Quality and ASR21 Feb 2024 0 repositories listed
-
Plugin Speech Enhancement: A Universal Speech Enhancement Framework Inspired by Dynamic Neural Network20 Feb 2024 0 repositories listed
-
SECP: A Speech Enhancement-Based Curation Pipeline For Scalable Acquisition Of Clean Speech19 Feb 2024 0 repositories listed
-
16 Feb 2024 0 repositories listed Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Diffusion Models for Audio Restoration15 Feb 2024 0 repositories listed
-
Overview of the L3DAS23 Challenge on Audio-Visual Extended Reality14 Feb 2024 0 repositories listed
-
Unrestricted Global Phase Bias-Aware Single-channel Speech Enhancement with Conformer-based Metric GAN13 Feb 2024 0 repositories listed
-
Array Geometry-Robust Attention-Based Neural Beamformer for Moving Speakers5 Feb 2024 0 repositories listed
-
1 Feb 2024 0 repositories listed
-
Real-time Stereo Speech Enhancement with Spatial-Cue Preservation based on Dual-Path Structure1 Feb 2024 0 repositories listed
-
SpeechComposer: Unifying Multiple Speech Tasks with Prompt Composition31 Jan 2024 0 repositories listed
-
A Two-Stage Framework in Cross-Spectrum Domain for Real-Time Speech Enhancement19 Jan 2024 0 repositories listed
-
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement18 Jan 2024 0 repositories listed
-
On Speech Pre-emphasis as a Simple and Inexpensive Method to Boost Speech Enhancement17 Jan 2024 0 repositories listed
-
Noise-robust zero-shot text-to-speech synthesis conditioned on self-supervised speech-representation model with adapters10 Jan 2024 0 repositories listed
-
FADI-AEC: Fast Score Based Diffusion Model Guided by Far-end Signal for Acoustic Echo Cancellation8 Jan 2024 0 repositories listed
-
A unified multichannel far-field speech recognition system: combining neural beamforming with attention based end-to-end model5 Jan 2024 0 repositories listed
-
Single-channel speech enhancement using learnable loss mixup20 Dec 2023 0 repositories listed
-
On real-time multi-stage speech enhancement systems19 Dec 2023 0 repositories listed
-
Attention-Driven Multichannel Speech Enhancement in Moving Sound Source Scenarios17 Dec 2023 0 repositories listed
-
A Deep Representation Learning-based Speech Enhancement Method Using Complex Convolution Recurrent Variational Autoencoder15 Dec 2023 0 repositories listed
-
SELM: Speech Enhancement Using Discrete Tokens and Language Models15 Dec 2023 0 repositories listed
-
Ultra Low Complexity Deep Learning Based Noise Suppression13 Dec 2023 0 repositories listed
-
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler5 Dec 2023 0 repositories listed
-
Head Orientation Estimation with Distributed Microphones Using Speech Radiation Patterns4 Dec 2023 0 repositories listed
-
SEFGAN: Harvesting the Power of Normalizing Flows and GANs for Efficient High-Quality Speech Enhancement4 Dec 2023 0 repositories listed
-
Subspace Hybrid MVDR Beamforming for Augmented Hearing30 Nov 2023 0 repositories listed
-
LC4SV: A Denoising Framework Learning to Compensate for Unseen Speaker Verification Models28 Nov 2023 0 repositories listed
-
CheapNET: Improving Light-weight speech enhancement network by projected loss function27 Nov 2023 0 repositories listed
-
Cooperative Dual Attention for Audio-Visual Speech Enhancement with Facial Cues24 Nov 2023 0 repositories listed
-
Sparsity-Driven EEG Channel Selection for Brain-Assisted Speech Enhancement22 Nov 2023 0 repositories listed
-
How does end-to-end speech recognition training impact speech enhancement artifacts?20 Nov 2023 0 repositories listed
-
SE Territory: Monaural Speech Enhancement Meets the Fixed Virtual Perceptual Space Mapping3 Nov 2023 0 repositories listed
-
DPATD: Dual-Phase Audio Transformer for Denoising30 Oct 2023 0 repositories listed
-
Scenario-Aware Audio-Visual TF-GridNet for Target Speech Extraction30 Oct 2023 0 repositories listed
-
Single channel speech enhancement by colored spectrograms26 Oct 2023 0 repositories listed
-
Generative Pre-training for Speech with Flow Matching25 Oct 2023 0 repositories listed
-
LC-TTFS: Towards Lossless Network Conversion for Spiking Neural Networks with TTFS Coding23 Oct 2023 0 repositories listed
-
Deep Beamforming for Speech Enhancement and Speaker Localization with an Array Response-Aware Loss Function19 Oct 2023 0 repositories listed
-
Real-time Speech Enhancement and Separation with a Unified Deep Neural Network for Single/Dual Talker Scenarios16 Oct 2023 0 repositories listed
-
A Single Speech Enhancement Model Unifying Dereverberation, Denoising, Speaker Counting, Separation, and Extraction12 Oct 2023 0 repositories listed
-
Magnitude-and-phase-aware Speech Enhancement with Parallel Sequence Modeling11 Oct 2023 0 repositories listed
-
Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms11 Oct 2023 0 repositories listed
-
VSANet: Real-time Speech Enhancement Based on Voice Activity Detection and Causal Spatial Attention11 Oct 2023 0 repositories listed
-
An experiment on an automated literature survey of data-driven speech enhancement methods10 Oct 2023 0 repositories listed
-
An Exploration of Task-decoupling on Two-stage Neural Post Filter for Real-time Personalized Acoustic Echo Cancellation7 Oct 2023 0 repositories listed
-
MBTFNet: Multi-Band Temporal-Frequency Neural Network For Singing Voice Enhancement6 Oct 2023 0 repositories listed
-
A Fused Deep Denoising Sound Coding Strategy for Bilateral Cochlear Implants2 Oct 2023 0 repositories listed
-
uSee: Unified Speech Enhancement and Editing with Conditional Diffusion Models2 Oct 2023 0 repositories listed
-
Toward Universal Speech Enhancement for Diverse Input Conditions29 Sep 2023 0 repositories listed
-
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study27 Sep 2023 0 repositories listed
-
Multichannel Voice Trigger Detection Based on Transform-average-concatenate27 Sep 2023 0 repositories listed
-
AutoPrep: An Automatic Preprocessing Framework for In-the-Wild Speech Data25 Sep 2023 0 repositories listed
-
DDTSE: Discriminative Diffusion Model for Target Speech Extraction25 Sep 2023 0 repositories listed
-
Speech enhancement with frequency domain auto-regressive modeling24 Sep 2023 0 repositories listed
-
A Multiscale Autoencoder (MSAE) Framework for End-to-End Neural Network Speech Enhancement21 Sep 2023 0 repositories listed
-
Deep Complex U-Net with Conformer for Audio-Visual Speech Enhancement20 Sep 2023 0 repositories listed
-
Joint Minimum Processing Beamforming and Near-end Listening Enhancement20 Sep 2023 0 repositories listed
-
Diffusion-based speech enhancement with a weighted generative-supervised learning loss19 Sep 2023 0 repositories listed
-
Exploring Speech Enhancement for Low-resource Speech Synthesis19 Sep 2023 0 repositories listed
-
Incorporating Ultrasound Tongue Images for Audio-Visual Speech Enhancement19 Sep 2023 0 repositories listed
-
Posterior sampling algorithms for unsupervised speech enhancement with recurrent variational autoencoder19 Sep 2023 0 repositories listed
-
Refining DNN-based Mask Estimation using CGMM-based EM Algorithm for Multi-channel Noise Reduction18 Sep 2023 0 repositories listed
-
Continuous Modeling of the Denoising Process for Speech Enhancement Based on Deep Learning17 Sep 2023 0 repositories listed
-
Unifying Robustness and Fidelity: A Comprehensive Study of Pretrained Generative Methods for Speech Enhancement in Adverse Conditions16 Sep 2023 0 repositories listed
-
Two-Step Knowledge Distillation for Tiny Speech Enhancement15 Sep 2023 0 repositories listed
-
AV2Wav: Diffusion-Based Re-synthesis from Continuous Self-supervised Features for Audio-Visual Speech Enhancement14 Sep 2023 0 repositories listed
-
Assessing the Generalization Gap of Learning-Based Speech Enhancement Systems in Noisy and Reverberant Environments12 Sep 2023 0 repositories listed
-
12 Sep 2023 0 repositories listed
-
Causal Signal-Based DCCRN with Overlapped-Frame Prediction for Online Speech Enhancement7 Sep 2023 0 repositories listed
-
Spiking Structured State Space Model for Monaural Speech Enhancement7 Sep 2023 0 repositories listed
-
Single-Channel Speech Enhancement with Deep Complex U-Networks and Probabilistic Latent Space Models4 Sep 2023 0 repositories listed
-
Noise robust speech emotion recognition with signal-to-noise ratio adapting speech enhancement3 Sep 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.