Datasets › MRPC

MRPC (Microsoft Research Paraphrase Corpus)

Introduced in Automatically Constructing a Corpus of Sentential Paraphrases archive 2025-07-28

Microsoft Research Paraphrase Corpus (MRPC) is a corpus consists of 5,801 sentence pairs collected from newswire articles. Each pair is labelled if it is a paraphrase or not by human annotators. The whole set is divided into a training subset (4,076 sentence pairs of which 2,753 are paraphrases) and a test subset (1,725 pairs of which 1,147 are paraphrases).

Source: Exploiting Semantic Annotations and Q-Learning for Constructing an Efficient Hierarchy/Graph Texts Organization Image Source: https://www.aclweb.org/anthology/I05-5002.pdf

Benchmarks archive 2025-07-28

All 4 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

30 shown of 36 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 786. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
SubRegWeigh: Effective and Efficient Annotation Weighing with Subword Regularization 1 1 10 Sep 2024 not harvested
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale 4 1 15 Aug 2022 ran 2 of 5 samples (3 unverified)
DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing 3 1 18 Nov 2021 ran 0 of 7 samples (7 unverified)
Differentiable Prompt Makes Pre-trained Language Models Better Few-shot Learners 4 1 30 Aug 2021 ran 3 of 4 samples (1 unverified)
AutoBERT-Zero: Evolving BERT Backbone from Scratch 0 2 15 Jul 2021 not harvested
Charformer: Fast Character Transformers via Gradient-based Subword Tokenization 2 1 23 Jun 2021 ran 7 of 10 samples (3 unverified)
FNet: Mixing Tokens with Fourier Transforms 12 1 9 May 2021 ran 2 of 2 samples (0 unverified; 1 pointer-only for licence)
Entailment as Few-Shot Learner 3 1 29 Apr 2021 ran 1 of 3 samples (2 unverified)
How to Train BERT with an Academic Budget 4 1 15 Apr 2021 not harvested
Nyströmformer: A Nyström-Based Algorithm for Approximating Self-Attention 10 1 7 Feb 2021 ran 1 of 2 samples (1 unverified; 1 pointer-only for licence)
CLEAR: Contrastive Learning for Sentence Representation 0 1 31 Dec 2020 not harvested
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning 2 2 22 Dec 2020 ran 1 of 1 samples (0 unverified; 1 pointer-only for licence)
RealFormer: Transformer Likes Residual Attention 5 1 21 Dec 2020 not harvested
A Statistical Framework for Low-bitwidth Training of Deep Neural Networks 2 1 27 Oct 2020 ran 1 of 4 samples (3 unverified; 1 pointer-only for licence)
Big Bird: Transformers for Longer Sequences 14 1 28 Jul 2020 ran 10 of 15 samples (5 unverified; 11 pointer-only for licence)
SqueezeBERT: What can computer vision teach NLP about efficient neural networks? 6 1 19 Jun 2020 ran 0 of 1 samples (1 unverified)
Synthesizer: Rethinking Self-Attention in Transformer Models 1 1 2 May 2020 ran 1 of 1 samples (0 unverified)
MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited Devices 7 1 6 Apr 2020 not harvested
Learning to Encode Position for Transformer with Continuous Dynamical Model 1 1 13 Mar 2020 ran 3 of 6 samples (3 unverified; 6 pointer-only for licence)
SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization 6 4 8 Nov 2019 ran 6 of 8 samples (2 unverified; 1 pointer-only for licence)
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer 57 5 23 Oct 2019 ran 2 of 31 samples (29 unverified)
Q8BERT: Quantized 8Bit BERT 5 1 14 Oct 2019 ran 3 of 11 samples (8 unverified; 3 pointer-only for licence)
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter 37 1 2 Oct 2019 ran 19 of 27 samples (8 unverified)
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations 48 1 26 Sep 2019 ran 46 of 126 samples (80 unverified; 22 pointer-only for licence)
TinyBERT: Distilling BERT for Natural Language Understanding 10 3 23 Sep 2019 ran 0 of 4 samples (4 unverified; 4 pointer-only for licence)
Q-BERT: Hessian Based Ultra Low Precision Quantization of BERT 0 1 12 Sep 2019 not harvested
StructBERT: Incorporating Language Structures into Pre-training for Deep Language Understanding 0 1 13 Aug 2019 not harvested
ERNIE 2.0: A Continual Pre-training Framework for Language Understanding 3 2 29 Jul 2019 ran 0 of 1 samples (1 unverified; 1 pointer-only for licence)
RoBERTa: A Robustly Optimized BERT Pretraining Approach 67 1 26 Jul 2019 ran 22 of 48 samples (26 unverified; 23 pointer-only for licence)
SpanBERT: Improving Pre-training by Representing and Predicting Spans 6 1 24 Jul 2019 ran 3 of 15 samples (12 unverified; 4 pointer-only for licence)

The full list of 36 is in the JSON twin.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • MRPC
  • MRPC Dev

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections