Datasets › AISHELL-1

AISHELL-1

Introduced by Hui Bu et al. in AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline archive 2025-07-28

AISHELL-1 is a corpus for speech recognition research and building speech recognition systems for Mandarin.

Source: AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Speech Recognition AISHELL-1 FireRedASR-AED Word Error Rate (WER) 0.55 FireRedASR: Open-Source Industrial-Grade Mandarin Speech... fireredteam/fireredasr 18 Compare

Papers archive 2025-07-28

16 shown of 16 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 197. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration 1 1 24 Jan 2025 not harvested
CR-CTC: Consistency regularization on CTC for improved speech recognition 1 1 7 Oct 2024 not harvested
Lightweight Transducer Based on Frame-Level Criterion 1 2 5 Sep 2024 not harvested
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition 0 1 5 Jul 2024 not harvested
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models 2 1 14 Nov 2023 ran 5 of 7 samples (2 unverified; 7 pointer-only for licence)
Unimodal Aggregation for CTC-based Speech Recognition 1 1 15 Sep 2023 not harvested
BAT: Boundary aware transducer for memory-efficient and low-latency ASR 1 1 19 May 2023 not harvested
FunASR: A Fundamental End-to-End Speech Recognition Toolkit 1 2 18 May 2023 not harvested
Beyond Universal Transformer: block reusing with adaptor in Transformer for automatic speech recognition 0 1 23 Mar 2023 not harvested
Knowledge Transfer from Pre-trained Language Models to Cif-based Speech Recognizers via Hierarchical Distillation 2 1 30 Jan 2023 not harvested
MMSpeech: Multi-modal Multi-task Encoder-Decoder Pre-training for Speech Recognition 1 1 29 Nov 2022 not harvested
Improving Mandarin Speech Recogntion with Block-augmented Transformer 2 1 24 Jul 2022 not harvested
Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition 5 1 10 Dec 2020 not harvested
CAT: A CTC-CRF based ASR Toolkit Bridging the Hybrid and the End-to-end Approaches towards Data Efficiency and Low Latency 1 1 27 May 2020 not harvested
A Comparative Study on Transformer vs RNN in Speech Applications 2 1 13 Sep 2019 not harvested
End-to-end Speech Recognition with Adaptive Computation Steps 0 1 30 Aug 2018 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

Apache-2

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • AISHELL-1

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections