Datasets › DIRHA
DIRHA (Distant-speech Interaction for Robust Home Applications)
DIRHA-English is a multi-microphone database composed of real and simulated sequences of 1-minute. The overall corpus is composed of different types of sequences including: 1) Phonetically-rich sentences; 2) WSJ 5-k utterances; 3) WSJ 20-k utterances; 4) Conversational speech (also including keywords and commands). The sequences are available for both UK and US English at 48 kHz. The DIRHA-English dataset offers the possibility to work with a very large number of microphone channels, to use of microphone arrays having different characteristics and to work considering different speech recognition tasks (e.g., phone-loop, keyword spotting, ASR with small and very large language models).
Source: The DIRHA-English Corpus Image Source: https://arxiv.org/pdf/1710.02560v1.pdf
Benchmarks archive 2025-07-28
All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Distant Speech Recognition | DIRHA English WSJ | Li-GRU Word Error Rate (WER) 23.9 | The PyTorch-Kaldi Speech Recognition Toolkit | mravanelli/pytorch-kaldi +10 | 3 | Compare |
Papers archive 2025-07-28
3 shown of 3 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 19. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| Learning Problem-agnostic Speech Representations from Multiple Self-supervised Tasks | 1 | 1 | 6 Apr 2019 | not harvested |
| Interpretable Convolutional Filters with SincNet | 1 | 1 | 23 Nov 2018 | not harvested |
| The PyTorch-Kaldi Speech Recognition Toolkit | 11 | 1 | 19 Nov 2018 | ran 5 of 6 samples (1 unverified; 6 pointer-only for licence) |
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
Unknown
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- DIRHA English WSJ
- DIRHA
2 variant names, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections