Datasets › ACE 2004
ACE 2004 (ACE 2004 Multilingual Training Corpus)
ACE 2004 Multilingual Training Corpus contains the complete set of English, Arabic and Chinese training data for the 2004 Automatic Content Extraction (ACE) technology evaluation. The corpus consists of data of various types annotated for entities and relations and was created by Linguistic Data Consortium with support from the ACE Program, with additional assistance from the DARPA TIDES (Translingual Information Detection, Extraction and Summarization) Program. The objective of the ACE program is to develop automatic content extraction technology to support automatic processing of human language in text form. In September 2004, sites were evaluated on system performance in six areas: Entity Detection and Recognition (EDR), Entity Mention Detection (EMD), EDR Co-reference, Relation Detection and Recognition (RDR), Relation Mention Detection (RMD), and RDR given reference entities. All tasks were evaluated in three languages: English, Chinese and Arabic.
Benchmarks archive 2025-07-28
All 6 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Nested Named Entity Recognition | ACE 2004 | PromptNER [RoBERTa-large] F1 88.72 | PromptNER: Prompt Locating and Typing for Named Entity... | tricktreat/promptner | 24 | Compare |
| Relation Extraction | ACE 2004 | PL-Marker RE+ Micro F1 66.5 | Packed Levitated Marker for Entity and Relation Extraction | tomaarsen/spanmarkerner +1 | 11 | Compare |
| Named Entity Recognition (NER) | ACE 2004 | Ours: cross-sentence ALB F1 90.3 | A Frustratingly Easy Approach for Entity and Relation Extraction | princeton-nlp/PURE +1 | 9 | Compare |
| Nested Mention Recognition | ACE 2004 | BoningKnife F1 86.41 | BoningKnife: Joint Entity Mention Detection and Typing... | — | 7 | Compare |
| Entity Disambiguation | ACE2004 | KBED Micro-F1 93.4 | Improving Entity Disambiguation by Reasoning over a... | alexa/refined +2 | 6 | Compare |
| UIE | ACE 2004 | KnowCoder-7b-IE F1 score 86.2 | KnowCoder: Coding Structured Knowledge into LLMs for... | ICT-GoKnow/KnowCoder | 1 | Compare |
Papers archive 2025-07-28
30 shown of 40 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 51. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
The full list of 40 is in the JSON twin.
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- ACE 2004
- ACE2004
2 variant names, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections