Datasets › VoxForge

VoxForge

archive 2025-07-28

VoxForge is an open speech dataset that was set up to collect transcribed speech for use with Free and Open Source Speech Recognition Engines (on Linux, Windows and Mac). Image Source: http://www.voxforge.org/home

Benchmarks archive 2025-07-28

All 9 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Spoken language identification VoxForge European 2D ConvNet(MixUp=YES) Accuracy (%) 96.3 Spoken Language Identification using ConvNets — 5 Compare
Spoken language identification VoxForge Commonwealth 2D ConvNet(MixUp=YES) Accuracy (%) 95.4 Spoken Language Identification using ConvNets — 4 Compare
Spoken language identification VoxForge LEAF Accuracy 91.5 EfficientLEAF: A Faster LEarnable Audio Frontend of... cpjku/efficientleaf 3 Compare
Accented Speech Recognition VoxForge Indian Deep Speech 2 Percentage error 22.44 Deep Speech 2: End-to-End Speech Recognition in English... tensorflow/models +34 2 Compare
Accented Speech Recognition VoxForge Commonwealth Deep Speech 2 Percentage error 13.56 Deep Speech 2: End-to-End Speech Recognition in English... tensorflow/models +34 2 Compare
Accented Speech Recognition VoxForge European Deep Speech 2 Percentage error 17.55 Deep Speech 2: End-to-End Speech Recognition in English... tensorflow/models +34 2 Compare
Accented Speech Recognition VoxForge American-Canadian Deep Speech 2 Percentage error 7.55 Deep Speech 2: End-to-End Speech Recognition in English... tensorflow/models +34 2 Compare
Keyword Spotting VoxForge 1D-ConvNet Accuracy (%) 93.7 Spoken Language Identification using ConvNets — 2 Compare
Language Identification VoxForge ConformerG-P Accuracy 99.8 BigSSL: Exploring the Frontier of Large-Scale... — 1 Compare

Papers archive 2025-07-28

5 shown of 5 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 11. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
EfficientLEAF: A Faster LEarnable Audio Frontend of Questionable Use 1 3 12 Jul 2022 not harvested
BigSSL: Exploring the Frontier of Large-Scale Semi-Supervised Learning for Automatic Speech Recognition 0 1 27 Sep 2021 not harvested
Spoken Language Identification using ConvNets 0 11 9 Oct 2019 not harvested
Deep Speech 2: End-to-End Speech Recognition in English and Mandarin 35 4 8 Dec 2015 ran 2 of 39 samples (37 unverified)
Deep Speech: Scaling up end-to-end speech recognition 24 4 17 Dec 2014 ran 9 of 9 samples (0 unverified; 8 pointer-only for licence)

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • VoxForge American-Canadian
  • VoxForge Commonwealth
  • VoxForge European
  • VoxForge Indian
  • VoxForge

5 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections