Datasets › speechocean762
speechocean762
speechocean762 is an open-source speech corpus designed for pronunciation assessment use, consisting of 5000 English utterances from 250 non-native speakers, where half of the speakers are children. Five experts annotated each of the utterances at sentence-level, word-level and phoneme-level. This corpus is allowed to be used freely for commercial and non-commercial purposes. To avoid subjective bias, each expert scores independently under the same metric
Benchmarks archive 2025-07-28
All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Phone-level pronunciation scoring | speechocean762 | HierCB+ConPCO Pearson correlation coefficient (PCC) 0.701 | ConPCO: Preserving Phoneme Characteristics for Automatic... | — | 8 | Compare |
| Utterance-level pronounciation scoring | speechocean762 | 3MH Pearson correlation coefficient (PCC) 0.811 | A Hierarchical Context-aware Modeling Approach for... | — | 5 | Compare |
| Word-level pronunciation scoring | speechocean762 | 3MH Pearson correlation coefficient (PCC) 0.694 | A Hierarchical Context-aware Modeling Approach for... | — | 5 | Compare |
Papers archive 2025-07-28
7 shown of 7 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 14. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| ConPCO: Preserving Phoneme Characteristics for Automatic Pronunciation Assessment Leveraging Contrastive Ordinal Regularization | 0 | 3 | 5 Jun 2024 | not harvested |
| Fine-Tuning Self-Supervised Learning Models for End-to-End Pronunciation Scoring | 1 | 1 | 19 Sep 2023 | not harvested |
| A Hierarchical Context-aware Modeling Approach for Multi-aspect and Multi-granular Pronunciation Assessment | 0 | 3 | 29 May 2023 | not harvested |
| Hierarchical Pronunciation Assessment with Multi-Aspect Attention | 1 | 3 | 15 Nov 2022 | not harvested |
| SpeechBlender: Speech Augmentation Framework for Mispronunciation Data Generation | 0 | 1 | 2 Nov 2022 | not harvested |
| Transformer-Based Multi-Aspect Multi-Granularity Non-Native English Speaker Pronunciation Assessment | 1 | 6 | 6 May 2022 | not harvested |
| speechocean762: An Open-Source Non-native English Speech Corpus For Pronunciation Assessment | 2 | 1 | 3 Apr 2021 | not harvested |
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
Attribution 4.0 International (CC BY 4.0)
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- speechocean762
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections