Datasets › CMRC

CMRC (Chinese Machine Reading Comprehension)

Introduced by Yiming Cui et al. in A Span-Extraction Dataset for Chinese Machine Reading Comprehension1 Jan 2019 archive 2025-07-28

CMRC is a dataset is annotated by human experts with near 20,000 questions as well as a challenging set which is composed of the questions that need reasoning over multiple clues.

Source: A Span-Extraction Dataset for Chinese Machine Reading Comprehension Image Source: https://www.aclweb.org/anthology/D19-1600.pdf

Benchmarks archive 2025-07-28

No leaderboard in the archive resolves to this dataset.

Papers archive 2025-07-28

No paper in the archive has a leaderboard row on this dataset; the archive counts 69 papers for it but never published that list.

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

CC-BY-SA-4.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • CMRC 2019 (Chinese Machine Reading Comprehension 2019)
  • CMRC 2019
  • CMRC 2018 (Chinese Machine Reading Comprehension 2018)
  • CMRC 2018
  • CMRC 2017
  • CMRC

6 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections