Datasets › XSum

XSum

Introduced by Shashi Narayan et al. in Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization1 Jan 2018 archive 2025-07-28

The Extreme Summarization (XSum) dataset is a dataset for evaluation of abstractive single-document summarization systems. The goal is to create a short, one-sentence new summary answering the question “What is the article about?”. The dataset consists of 226,711 news articles accompanied with a one-sentence summary. The articles are collected from BBC articles (2010 to 2017) and cover a wide variety of domains (e.g., News, Politics, Sports, Weather, Business, Technology, Science, Health, Family, Education, Entertainment and Arts). The official random split contains 204,045 (90%), 11,332 (5%) and 11,334 (5) documents in training, validation and test sets, respectively.

Source: https://arxiv.org/pdf/1808.08745.pdf Image Source: https://arxiv.org/pdf/1808.08745.pdf

Benchmarks archive 2025-07-28

All 5 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

12 shown of 12 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 32. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Segmented Recurrent Transformer: An Efficient Sequence-to-Sequence Model 1 1 24 May 2023 not harvested
PaLM 2 Technical Report 1 3 17 May 2023 not harvested
Lift Yourself Up: Retrieval-augmented Text Generation with Self Memory 1 1 3 May 2023 ran 1 of 1 samples (0 unverified; 1 pointer-only for licence)
BRIO: Bringing Order to Abstractive Summarization 3 1 31 Mar 2022 ran 0 of 10 samples (10 unverified)
SummaReranker: A Multi-Task Mixture-of-Experts Re-ranking Framework for Abstractive Summarization 1 1 13 Mar 2022 ran 2 of 3 samples (1 unverified)
SimCLS: A Simple Framework for Contrastive Learning of Abstractive Summarization 2 1 3 Jun 2021 ran 4 of 6 samples (2 unverified; 1 pointer-only for licence)
Hierarchical Learning for Generation with Long Source Sequences 0 1 15 Apr 2021 not harvested
The GEM Benchmark: Natural Language Generation, its Evaluation and Metrics 0 1 2 Feb 2021 not harvested
PEGASUS: Pre-training with Extracted Gap-sentences for Abstractive Summarization 19 1 18 Dec 2019 ran 1 of 19 samples (18 unverified)
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension 47 1 29 Oct 2019 ran 22 of 53 samples (31 unverified; 7 pointer-only for licence)
Text Summarization with Pretrained Encoders 19 1 22 Aug 2019 ran 7 of 21 samples (14 unverified)
Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization 3 7 27 Aug 2018 ran 0 of 4 samples (4 unverified)

Dataset loaders archive 2025-07-28

9 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • XSum
  • X-Sum

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections