Datasets › RTE

RTE (Recognizing Textual Entailment)

archive 2025-07-28

The Recognizing Textual Entailment (RTE) datasets come from a series of textual entailment challenges. Data from RTE1, RTE2, RTE3 and RTE5 is combined. Examples are constructed based on news and Wikipedia text.

Benchmarks archive 2025-07-28

All 2 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Natural Language Inference RTE Vega v2 6B (KD-based prompt transfer) Accuracy 96% Toward Efficient Language Model Pretraining and... — 90 Compare
Classification RTE OPT-1.3B Test Accuracy 60.89% Achieving Dimension-Free Communication in Federated... ZidongLiu/DeComFL 2 Compare

Papers archive 2025-07-28

30 shown of 48 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 56. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Achieving Dimension-Free Communication in Federated Learning via Zeroth-Order Optimization 1 2 24 May 2024 ran 1 of 1 samples (0 unverified)
Not all layers are equally as important: Every Layer Counts BERT 0 4 3 Nov 2023 not harvested
The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning 2 1 23 May 2023 not harvested
PaLM 2 Technical Report 1 3 17 May 2023 not harvested
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions 1 5 27 Apr 2023 not harvested
BloombergGPT: A Large Language Model for Finance 2 4 30 Mar 2023 not harvested
Exploring the Benefits of Training Expert Language Models over Instruction Tuning 2 1 7 Feb 2023 ran 3 of 3 samples (0 unverified; 3 pointer-only for licence)
Hungry Hungry Hippos: Towards Language Modeling with State Space Models 3 5 28 Dec 2022 ran 7 of 15 samples (8 unverified)
OPT-IML: Scaling Language Model Instruction Meta Learning through the Lens of Generalization 1 6 22 Dec 2022 not harvested
Toward Efficient Language Model Pretraining and Downstream Adaptation via Self-Evolution: A Case Study on SuperGLUE 0 2 4 Dec 2022 not harvested
Knowledge-in-Context: Towards Knowledgeable Semi-Parametric Language Models 0 1 28 Oct 2022 not harvested
Guess the Instruction! Flipped Learning Makes Language Models Stronger Zero-Shot Learners 1 1 6 Oct 2022 not harvested
Ask Me Anything: A simple strategy for prompting language models 3 3 5 Oct 2022 ran 2 of 2 samples (0 unverified)
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale 4 1 15 Aug 2022 ran 2 of 5 samples (3 unverified)
AlexaTM 20B: Few-Shot Learning Using a Large-Scale Multilingual Seq2Seq Model 1 1 2 Aug 2022 ran 1 of 1 samples (0 unverified)
N-Grammer: Augmenting Transformers with latent n-grams 2 1 13 Jul 2022 ran 0 of 6 samples (6 unverified)
UL2: Unifying Language Learning Paradigms 2 2 10 May 2022 ran 0 of 16 samples (16 unverified)
PaLM: Scaling Language Modeling with Pathways 7 4 5 Apr 2022 ran 30 of 37 samples (7 unverified)
ST-MoE: Designing Stable and Transferable Sparse Expert Models 3 2 17 Feb 2022 ran 5 of 5 samples (0 unverified; 5 pointer-only for licence)
data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language 12 1 7 Feb 2022 ran 0 of 6 samples (6 unverified)
DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing 3 1 18 Nov 2021 ran 0 of 7 samples (7 unverified)
Finetuned Language Models Are Zero-Shot Learners 8 3 3 Sep 2021 ran 0 of 1 samples (1 unverified)
FNet: Mixing Tokens with Fourier Transforms 12 1 9 May 2021 ran 2 of 2 samples (0 unverified; 1 pointer-only for licence)
Entailment as Few-Shot Learner 3 2 29 Apr 2021 ran 1 of 3 samples (2 unverified)
How to Train BERT with an Academic Budget 4 1 15 Apr 2021 not harvested
Muppet: Massive Multi-task Representations with Pre-Finetuning 2 1 26 Jan 2021 not harvested
CLEAR: Contrastive Learning for Sentence Representation 0 1 31 Dec 2020 not harvested
RealFormer: Transformer Likes Residual Attention 5 1 21 Dec 2020 not harvested
A Statistical Framework for Low-bitwidth Training of Deep Neural Networks 2 1 27 Oct 2020 ran 1 of 4 samples (3 unverified; 1 pointer-only for licence)
Big Bird: Transformers for Longer Sequences 14 1 28 Jul 2020 ran 10 of 15 samples (5 unverified; 11 pointer-only for licence)

The full list of 48 is in the JSON twin.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • RTE

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections