Datasets › LDC2020T02

LDC2020T02 (Abstract Meaning Representation (AMR) Annotation Release 3.0)

15 Jan 2020 archive 2025-07-28

Abstract Meaning Representation (AMR) Annotation Release 3.0 was developed by the Linguistic Data Consortium (LDC), SDL/Language Weaver, Inc., the University of Colorado's Computational Language and Educational Research group and the Information Sciences Institute at the University of Southern California. It contains a sembank (semantic treebank) of over 59,255 English natural language sentences from broadcast conversations, newswire, weblogs, web discussion forums, fiction and web text. This release adds new data to, and updates material contained in, Abstract Meaning Representation 2.0 (LDC2017T10), specifically: more annotations on new and prior data, new or improved PropBank-style frames, enhanced quality control, and multi-sentence annotations.

AMR captures "who is doing what to whom" in a sentence. Each sentence is paired with a graph that represents its whole-sentence meaning in a tree-structure. AMR utilizes PropBank frames, non-core semantic roles, within-sentence coreference, named entity annotation, modality, negation, questions, quantities, and so on to represent the semantic structure of a sentence largely independent of its syntax.

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
AMR Parsing LDC2020T02 Graphene Smatch (MBSE paper) (IBM) Smatch 85.4 Maximum Bayes Smatch Ensemble Distillation for AMR Parsing IBM/transition-amr-parser +2 13 Compare

Papers archive 2025-07-28

8 shown of 8 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 9. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Incorporating Graph Information in Transformer-based AMR Parsing 1 2 23 Jun 2023 not harvested
BiBL: AMR Parsing and Generation with Bidirectional Bayesian Learning 1 2 1 Oct 2022 not harvested
ATP: AMRize Then Parse! Enhancing AMR Parsing with PseudoAMRs 2 1 19 Apr 2022 not harvested
Graph Pre-training for AMR Parsing and Generation 2 1 15 Mar 2022 ran 3 of 5 samples (2 unverified)
Maximum Bayes Smatch Ensemble Distillation for AMR Parsing 3 2 14 Dec 2021 not harvested
Ensembling Graph Predictions for AMR Parsing 1 2 18 Oct 2021 ran 3 of 3 samples (0 unverified)
One SPRING to Rule Them Both: Symmetric AMR Semantic Parsing and Generation without a Complex Pipeline 1 2 18 May 2021 not harvested
AMR Parsing with Action-Pointer Transformer 1 1 29 Apr 2021 ran 1 of 2 samples (1 unverified)

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

LDC User Agreement for Non-Members

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • LDC2020T02

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections