Datasets › XNLI
XNLI (Cross-lingual Natural Language Inference)
The Cross-lingual Natural Language Inference (XNLI) corpus is the extension of the Multi-Genre NLI (MultiNLI) corpus to 15 languages. The dataset was created by manually translating the validation and test sets of MultiNLI into each of those 15 languages. The English training set was machine translated for all languages. The dataset is composed of 122k train, 2490 validation and 5010 test examples.
Source: CamemBERT: a Tasty French Language Model Image Source: https://github.com/facebookresearch/XNLI
Benchmarks archive 2025-07-28
All 7 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
Papers archive 2025-07-28
12 shown of 12 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 349. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
Dataset loaders archive 2025-07-28
7 loaders as listed in the archive; links are outbound and not re-checked here.
Tasks archive 2025-07-28
License archive 2025-07-28
Attribution-NonCommercial 4.0 International
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- XNLI Zero-Shot English-to-French
- XNLI French
- XNLI Dev
- XNLI Chinese
- XNLI Chinese Dev
- XNLI Zero-Shot English-to-German
- XNLI Zero-Shot English-to-Spanish
- XNLI
8 variant names, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections