Datasets › VoiceBank + DEMAND

VoiceBank + DEMAND (Noisy speech database for training speech enhancement algorithms and TTS models)

Introduced in The Diverse Environments Multi-channel Acoustic Noise Database (DEMAND): A database of multichannel environmental noise recordings12 Sep 2016 archive 2025-07-28

VoiceBank+DEMAND is a noisy speech database for training speech enhancement algorithms and TTS models. The database was designed to train and test speech enhancement methods that operate at 48kHz. A more detailed description can be found in the paper associated with the database. Some of the noises were obtained from the Demand database, available here: http://parole.loria.fr/DEMAND/ . The speech database was obtained from the Voice Banking Corpus, available here: http://homepages.inf.ed.ac.uk/jyamagis/release/VCTK-Corpus.tar.gz .

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Speech Enhancement VoiceBank + DEMAND ROSE-CD(PESQ) PESQ (wb) 3.99 Robust One-step Speech Enhancement via Consistency Distillation LiangXu123/Robust-One-step-Speech-Enhancement-via-Consistency-Distillation-ROSE-CD- 42 Compare

Papers archive 2025-07-28

30 shown of 33 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 53. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Robust One-step Speech Enhancement via Consistency Distillation 1 2 8 Jul 2025 not harvested
PrimeK-Net: Multi-scale Spectral Learning via Group Prime-Kernel Convolutional Neural Networks for Single Channel Speech Enhancement 1 1 27 Feb 2025 not harvested
Let SSMs be ConvNets: State-space Modeling with Optimal Tensor Contractions 0 1 22 Jan 2025 not harvested
xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement 1 1 10 Jan 2025 not harvested
Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement 1 1 21 Dec 2024 not harvested
Dense-TSNet: Dense Connected Two-Stage Structure for Ultra-Lightweight Speech Enhancement 0 1 18 Sep 2024 not harvested
Investigating Training Objectives for Generative Speech Enhancement 1 1 16 Sep 2024 not harvested
aTENNuate: Optimized Real-time Speech Enhancement with Deep SSMs on Raw Audio 0 1 5 Sep 2024 not harvested
The PESQetarian: On the Relevance of Goodhart's Law for Speech Enhancement 0 1 5 Jun 2024 not harvested
An Investigation of Incorporating Mamba for Speech Enhancement 1 1 10 May 2024 ran 13 of 14 samples (1 unverified; 14 pointer-only for licence)
FSPEN: AN ULTRA-LIGHTWEIGHT NETWORK FOR REAL TIME SPEECH ENAHNCMENT 1 1 15 Apr 2024 not harvested
An Analysis of the Variance of Diffusion-based Speech Enhancement 0 1 1 Feb 2024 not harvested
ROSE: A Recognition-Oriented Speech Enhancement Framework in Air Traffic Control Using Multi-Objective Learning 1 1 11 Dec 2023 not harvested
Explicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech Enhancement 1 1 17 Aug 2023 ran 13 of 14 samples (1 unverified)
MetricGAN-OKD: Multi-Metric Optimization of MetricGAN via Online Knowledge Distillation for Speech Enhancement 1 2 24 Jul 2023 not harvested
DeepFilterNet: Perceptually Motivated Real-Time Speech Enhancement 1 1 14 May 2023 not harvested
D²Net: A Denoising and Dereverberation Network Based on Two-branch Encoder and Dual-path Transformer 0 1 21 Nov 2022 not harvested
SCP-GAN: Self-Correcting Discriminator Optimization for Training Consistency Preserving Metric GAN on Speech Enhancement Tasks 0 1 26 Oct 2022 not harvested
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement 2 1 22 Sep 2022 not harvested
Multi-View Attention Transfer for Efficient Speech Enhancement 0 1 22 Aug 2022 not harvested
Speech Enhancement and Dereverberation with Diffusion-based Generative Models 1 1 11 Aug 2022 not harvested
Boosting Self-Supervised Embeddings for Speech Enhancement 1 1 7 Apr 2022 ran 2 of 7 samples (5 unverified)
Perceptual Contrast Stretching on Target Feature for Speech Enhancement 1 1 31 Mar 2022 not harvested
MANNER: Multi-view Attention Network for Noise Erasure 1 1 4 Mar 2022 not harvested
MetricGAN+: An Improved Version of MetricGAN for Speech Enhancement 3 1 8 Apr 2021 ran 0 of 11 samples (11 unverified)
A Modulation-Domain Loss for Neural-Network-based Real-time Speech Enhancement 1 1 15 Feb 2021 not harvested
Monaural Speech Enhancement with Complex Convolutional Block Attention Module and Joint Time Frequency Losses 2 1 3 Feb 2021 not harvested
Improving Perceptual Quality by Phone-Fortified Perceptual Loss using Wasserstein Distance for Speech Enhancement 1 1 28 Oct 2020 ran 0 of 1 samples (1 unverified)
Perceptual Loss based Speech Denoising with an ensemble of Audio Pattern Recognition and Self-Supervised Models 1 1 22 Oct 2020 not harvested
Real Time Speech Enhancement in the Waveform Domain 3 2 23 Jun 2020 ran 1 of 2 samples (1 unverified; 2 pointer-only for licence)

The full list of 33 is in the JSON twin.

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

CC BY 4.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • VoiceBank + DEMAND

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections