Datasets › LDC2020T02
LDC2020T02 (Abstract Meaning Representation (AMR) Annotation Release 3.0)
Abstract Meaning Representation (AMR) Annotation Release 3.0 was developed by the Linguistic Data Consortium (LDC), SDL/Language Weaver, Inc., the University of Colorado's Computational Language and Educational Research group and the Information Sciences Institute at the University of Southern California. It contains a sembank (semantic treebank) of over 59,255 English natural language sentences from broadcast conversations, newswire, weblogs, web discussion forums, fiction and web text. This release adds new data to, and updates material contained in, Abstract Meaning Representation 2.0 (LDC2017T10), specifically: more annotations on new and prior data, new or improved PropBank-style frames, enhanced quality control, and multi-sentence annotations.
AMR captures "who is doing what to whom" in a sentence. Each sentence is paired with a graph that represents its whole-sentence meaning in a tree-structure. AMR utilizes PropBank frames, non-core semantic roles, within-sentence coreference, named entity annotation, modality, negation, questions, quantities, and so on to represent the semantic structure of a sentence largely independent of its syntax.
Benchmarks archive 2025-07-28
All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| AMR Parsing | LDC2020T02 | Graphene Smatch (MBSE paper) (IBM) Smatch 85.4 | Maximum Bayes Smatch Ensemble Distillation for AMR Parsing | IBM/transition-amr-parser +2 | 13 | Compare |
Papers archive 2025-07-28
8 shown of 8 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 9. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| Incorporating Graph Information in Transformer-based AMR Parsing | 1 | 2 | 23 Jun 2023 | not harvested |
| BiBL: AMR Parsing and Generation with Bidirectional Bayesian Learning | 1 | 2 | 1 Oct 2022 | not harvested |
| ATP: AMRize Then Parse! Enhancing AMR Parsing with PseudoAMRs | 2 | 1 | 19 Apr 2022 | not harvested |
| Graph Pre-training for AMR Parsing and Generation | 2 | 1 | 15 Mar 2022 | ran 3 of 5 samples (2 unverified) |
| Maximum Bayes Smatch Ensemble Distillation for AMR Parsing | 3 | 2 | 14 Dec 2021 | not harvested |
| Ensembling Graph Predictions for AMR Parsing | 1 | 2 | 18 Oct 2021 | ran 3 of 3 samples (0 unverified) |
| One SPRING to Rule Them Both: Symmetric AMR Semantic Parsing and Generation without a Complex Pipeline | 1 | 2 | 18 May 2021 | not harvested |
| AMR Parsing with Action-Pointer Transformer | 1 | 1 | 29 Apr 2021 | ran 1 of 2 samples (1 unverified) |
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
LDC User Agreement for Non-Members
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- LDC2020T02
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections