Datasets › ReCoRD

ReCoRD

Introduced by Sheng Zhang et al. in ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension30 Oct 2018 archive 2025-07-28

Reading Comprehension with Commonsense Reasoning Dataset (ReCoRD) is a large-scale reading comprehension dataset which requires commonsense reasoning. ReCoRD consists of queries automatically generated from CNN/Daily Mail news articles; the answer to each query is a text span from a summarizing passage of the corresponding news. The goal of ReCoRD is to evaluate a machine's ability of commonsense reasoning in reading comprehension. ReCoRD is pronounced as [ˈrɛkərd].

Image Source: Zhang et al

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Common Sense Reasoning ReCoRD Turing NLR v5 XXL 5.4B (fine-tuned) EM 95.9 Toward Efficient Language Model Pretraining and... — 45 Compare

Papers archive 2025-07-28

20 shown of 20 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 111. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Integrating a Heterogeneous Graph with Entity-aware Self-attention using Relative Position Labels for Reading Comprehension Model 0 1 19 Jul 2023 not harvested
PaLM 2 Technical Report 1 3 17 May 2023 not harvested
BloombergGPT: A Large Language Model for Finance 2 4 30 Mar 2023 not harvested
LUKE-Graph: A Transformer-based Approach with Gated Relational Graph Attention for Cloze-style Reading Comprehension 0 1 12 Mar 2023 not harvested
Toward Efficient Language Model Pretraining and Downstream Adaptation via Self-Evolution: A Case Study on SuperGLUE 0 2 4 Dec 2022 not harvested
AlexaTM 20B: Few-Shot Learning Using a Large-Scale Multilingual Seq2Seq Model 1 1 2 Aug 2022 ran 1 of 1 samples (0 unverified)
N-Grammer: Augmenting Transformers with latent n-grams 2 1 13 Jul 2022 ran 0 of 6 samples (6 unverified)
Large Language Models are Zero-Shot Reasoners 4 1 24 May 2022 ran 0 of 4 samples (4 unverified; 1 pointer-only for licence)
PaLM: Scaling Language Modeling with Pathways 7 1 5 Apr 2022 ran 30 of 37 samples (7 unverified)
Efficient Language Modeling with Sparse all-MLP 0 5 14 Mar 2022 not harvested
ST-MoE: Designing Stable and Transferable Sparse Expert Models 3 2 17 Feb 2022 ran 5 of 5 samples (0 unverified; 5 pointer-only for licence)
KELM: Knowledge Enhanced Pre-Trained Language Representations with Message Passing on Hierarchical Relational Graphs 1 2 9 Sep 2021 ran 1 of 6 samples (5 unverified)
Finetuned Language Models Are Zero-Shot Learners 8 2 3 Sep 2021 ran 0 of 1 samples (1 unverified)
LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention 9 1 2 Oct 2020 ran 3 of 10 samples (7 unverified)
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 14 1 5 Jun 2020 ran 4 of 13 samples (9 unverified; 3 pointer-only for licence)
Language Models are Few-Shot Learners 67 1 28 May 2020 ran 15 of 65 samples (50 unverified; 4 pointer-only for licence)
Pingan Smart Health and SJTU at COIN - Shared Task: utilizing Pre-trained Language Models and Common-sense Knowledge in Machine Reading Tasks 0 1 1 Nov 2019 not harvested
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer 57 2 23 Oct 2019 ran 2 of 31 samples (29 unverified)
ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension 0 1 30 Oct 2018 not harvested
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 534 1 11 Oct 2018 ran 204 of 659 samples (455 unverified; 149 pointer-only for licence)

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

Custom

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • ReCoRD

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections