Datasets › DWIE
DWIE (Deutsche Welle corpus for Information Extraction)
The 'Deutsche Welle corpus for Information Extraction' (DWIE) is a multi-task dataset that combines four main Information Extraction (IE) annotation sub-tasks: (i) Named Entity Recognition (NER), (ii) Coreference Resolution, (iii) Relation Extraction (RE), and (iv) Entity Linking. DWIE is conceived as an entity-centric dataset that describes interactions and properties of conceptual entities on the level of the complete document.
Source: https://arxiv.org/abs/2009.12626
Benchmarks archive 2025-07-28
All 5 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Coreference Resolution | DWIE | REXEL Avg. F1 95.12 | REXEL: An End-to-end Model for Document-Level Relation... | amazon-science/e2e-docie | 3 | Compare |
| Named Entity Recognition (NER) | DWIE | REXEL F1-Hard 90.59 | REXEL: An End-to-end Model for Document-Level Relation... | amazon-science/e2e-docie | 3 | Compare |
| Relation Extraction | DWIE | REXEL F1-Hard 65.8 | REXEL: An End-to-end Model for Document-Level Relation... | amazon-science/e2e-docie | 3 | Compare |
| Document-level Closed Information Extraction | DWIE | REXEL F1-Hard 53.77 | REXEL: An End-to-end Model for Document-Level Relation... | amazon-science/e2e-docie | 1 | Compare |
| Document-level Relation Extraction | DWIE | VaeDiff-DocRE F1 0.7307 | VaeDiff-DocRE: End-to-end Data Augmentation Framework... | khaitran22/vaediff-docre | 1 | Compare |
Papers archive 2025-07-28
4 shown of 4 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 18. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| VaeDiff-DocRE: End-to-end Data Augmentation Framework for Document-level Relation Extraction | 1 | 1 | 18 Dec 2024 | not harvested |
| REXEL: An End-to-end Model for Document-Level Relation Extraction and Entity Linking | 1 | 4 | 19 Apr 2024 | ran 3 of 3 samples (0 unverified) |
| Injecting Knowledge Base Information into End-to-End Joint Entity and Relation Extraction and Coreference Resolution | 1 | 3 | 5 Jul 2021 | not harvested |
| DWIE: an entity-centric dataset for multi-task document-level information extraction | 2 | 3 | 26 Sep 2020 | ran 5 of 14 samples (9 unverified; 3 pointer-only for licence) |
Dataset loaders archive 2025-07-28
1 loader as listed in the archive; links are outbound and not re-checked here.
Tasks archive 2025-07-28
License archive 2025-07-28
GPL-3.0 License
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- DWIE
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections