Datasets › CLUE

CLUE (Chinese Language Understanding Evaluation Benchmark)

Introduced by Liang Xu et al. in CLUE: A Chinese Language Understanding Evaluation Benchmark archive 2025-07-28

CLUE is a Chinese Language Understanding Evaluation benchmark. It consists of different NLU datasets. It is a community-driven project that brings together 9 tasks spanning several well-established single-sentence/sentence-pair classification tasks, as well as machine reading comprehension, all on original Chinese text.

Benchmarks archive 2025-07-28

All 8 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

2 shown of 2 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 99. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
GLM-130B: An Open Bilingual Pre-trained Model 9 14 5 Oct 2022 ran 5 of 21 samples (16 unverified)
XLNet: Generalized Autoregressive Pretraining for Language Understanding 27 1 19 Jun 2019 ran 10 of 24 samples (14 unverified; 3 pointer-only for licence)

Dataset loaders archive 2025-07-28

2 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • CLUE (CMRC2018)
  • CLUE (AFQMC)
  • CLUE (OCNLI_50K)
  • CLUE (DRCD)
  • CLUE (CMNLI)
  • CLUE (WSC1.1)
  • CLUE (C3)
  • ClueWeb09-B
  • CLUE

9 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections