Datasets › DIRHA

DIRHA (Distant-speech Interaction for Robust Home Applications)

Introduced by Mirco Ravanelli et al. in The DIRHA-English corpus and related tasks for distant-speech recognition in domestic environments1 Jan 2015 archive 2025-07-28

DIRHA-English is a multi-microphone database composed of real and simulated sequences of 1-minute. The overall corpus is composed of different types of sequences including: 1) Phonetically-rich sentences; 2) WSJ 5-k utterances; 3) WSJ 20-k utterances; 4) Conversational speech (also including keywords and commands). The sequences are available for both UK and US English at 48 kHz. The DIRHA-English dataset offers the possibility to work with a very large number of microphone channels, to use of microphone arrays having different characteristics and to work considering different speech recognition tasks (e.g., phone-loop, keyword spotting, ASR with small and very large language models).

Source: The DIRHA-English Corpus Image Source: https://arxiv.org/pdf/1710.02560v1.pdf

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Distant Speech Recognition DIRHA English WSJ Li-GRU Word Error Rate (WER) 23.9 The PyTorch-Kaldi Speech Recognition Toolkit mravanelli/pytorch-kaldi +10 3 Compare

Papers archive 2025-07-28

3 shown of 3 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 19. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Learning Problem-agnostic Speech Representations from Multiple Self-supervised Tasks 1 1 6 Apr 2019 not harvested
Interpretable Convolutional Filters with SincNet 1 1 23 Nov 2018 not harvested
The PyTorch-Kaldi Speech Recognition Toolkit 11 1 19 Nov 2018 ran 5 of 6 samples (1 unverified; 6 pointer-only for licence)

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • DIRHA English WSJ
  • DIRHA

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections