Home › Datasets › task › Semantic Parsing
Semantic Parsing datasets
archive 2025-07-28
44 datasets carry the task tag "Semantic Parsing" (the task itself: Semantic Parsing), ordered by the archive's paper count. Page 1 of 1: 44 shown of 44. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Semantic Parsing datasets 1–44 of 44
FrameNet is a linguistic knowledge graph containing information about lexical and predicate argument semantics of the English language.
444 papers · 0 benchmarks
WikiSQL consists of a corpus of 87,726 hand-annotated SQL query and natural language question pairs.
267 papers · 4 benchmarks
SCAN (Simplified versions of the CommAI Navigation tasks)
SCAN is a dataset for grounded navigation which consists of a set of simple compositional navigation commands paired with the corresponding action sequences.
148 papers · 0 benchmarks
Spider dataset is used for evaluation in the paper "Structure-Grounded Pretraining for Text-to-SQL".
91 papers · 3 benchmarks
WikiTableQuestions is a question answering dataset over semi-structured tables.
79 papers · 2 benchmarks
CFQ (Compositional Freebase Questions)
A large and realistic natural language question answering dataset.
66 papers · 1 benchmark
Occluded REID is an occluded person dataset captured by mobile cameras, consisting of 2,000 images of 200 occluded persons (see Fig.
65 papers · 1 benchmark
ComplexWebQuestions is a dataset for answering complex questions that require reasoning over multiple web snippets.
62 papers · 2 benchmarks
The WebQuestionsSP dataset is released as part of our ACL-2016 paper “The Value of Semantic Parse Labeling for Knowledge Base Question Answering” [Yih, Richardson, Meek, Chang & Suh, 2016], in which we evaluated the value of gathering…
61 papers · 3 benchmarks
SParC (Semantic Parsing in Context)
SParC is a large-scale dataset for complex, cross-domain, and context-dependent (multi-turn) semantic parsing and text-to-SQL task (interactive natural language interfaces for relational databases).
59 papers · 2 benchmarks
NomBank is an annotation project at New York University that is related to the PropBank project at the University of Colorado.
50 papers · 0 benchmarks
A new large dataset with over 100,000 examples consisting of Java classes from online code repositories, and develop a new encoder-decoder architecture that models the interaction between the method documentation and the class environment.
46 papers · 1 benchmark
CoSQL (Conversational Text-to-SQL Challenge)
CoSQL is a corpus for building cross-domain, general-purpose database (DB) querying dialogue systems.
45 papers · 1 benchmark
A new large-scale geometry problem-solving dataset - 3,002 multi-choice geometry problems - dense annotations in formal language for the diagrams and text - 27,213 annotated diagram logic forms (literals) - 6,293 annotated text logic forms…
45 papers · 1 benchmark
Contains around 200K dialogs with a total of 1.6M turns.
40 papers · 0 benchmarks
The SQA dataset was created to explore the task of answering sequences of inter-related questions on HTML tables.
37 papers · 1 benchmark
Dataset is constructed from single intent dataset SNIPS.
26 papers · 2 benchmarks
TOPv2 (Task Oriented Parsing v2)
Task Oriented Parsing v2 (TOPv2) representations for intent-slot based dialog systems.
25 papers · 0 benchmarks
A large-scale dataset for Complex KBQA.
23 papers · 1 benchmark
AMR Bank (Abstract Meaning Representation)
The AMR Bank is a set of English sentences paired with simple, readable semantic representations.
22 papers · 1 benchmark
This dataset contains card descriptions of the card game Hearthstone and the code that implements them.
22 papers · 0 benchmarks
QuaRel is a crowdsourced dataset of 2771 multiple-choice story questions, including their logical forms.
20 papers · 0 benchmarks
One of the largest commonsense knowledge bases available, describing over 2 million disambiguated concepts and activities, connected by over 18 million assertions.
20 papers · 0 benchmarks
Groningen Meaning Bank is a semantic resource that anyone can edit and that integrates various semantic phenomena, including predicate-argument structure, scope, tense, thematic roles, animacy, pronouns, and rhetorical relations.
18 papers · 0 benchmarks
GraphQuestions is a characteristic-rich dataset designed for factoid question answering.
14 papers · 2 benchmarks
Fashion 144K is a novel heterogeneous dataset with 144,169 user posts containing diverse image, textual and meta information.
11 papers · 0 benchmarks
SEDE (Stack Exchange Data Explorer)
SEDE is a dataset comprised of 12,023 complex and diverse SQL queries and their natural language titles and descriptions, written by real users of the Stack Exchange Data Explorer out of a natural interaction.
11 papers · 1 benchmark
ComQA is a large dataset of real user questions that exhibit different challenging aspects such as compositionality, temporal reasoning, and comparisons.
9 papers · 0 benchmarks
A dataset of utterances, incorrect SQL interpretations and the corresponding natural language feedback.
6 papers · 0 benchmarks
SimpleQuestionsWikidata maps SimpleQuestions to Wikidata.
6 papers · 1 benchmark
1000 query triples on 120 tables.
5 papers · 0 benchmarks
PCFG SET (Probabilistic Context Free Grammar String Edit Task)
The Probabilistic Context Free Grammar String Edit Task (PCFG SET) dataset is a dataset with sequence to sequence problems specifically designed to test different aspects of compositional generalisation.
4 papers · 0 benchmarks
The Szeged Treebank is the largest fully manually annotated treebank of the Hungarian language.
4 papers · 0 benchmarks
aethel (Automatically Extracted Theorems from Lassy)
A dataset of approximately 75,000 phrases and sentences, syntactically analyzed as typelogical derivations (i.e.
4 papers · 0 benchmarks
Question Answering (QA) is a widely-used framework for developing and evaluating an intelligent machine.
3 papers · 0 benchmarks
Multilingual TOP is a dataset for multilingual semantic parsing with human-written sentences as opposed to machine translated ones.
3 papers · 0 benchmarks
The first dataset contains annotated natural language queries (i.e.
2 papers · 0 benchmarks
Spades (Semantic PArsing of DEclarative Sentences)
Datasets Spades contains 93,319 questions derived from clueweb09 sentences.
2 papers · 0 benchmarks
Schema2QA is the first large question answering dataset over real-world Schema.org data.
2 papers · 0 benchmarks
ViText2SQL is a dataset for the Vietnamese Text-to-SQL semantic parsing task, consisting of about 10K question and SQL query pairs.
2 papers · 0 benchmarks
Conic10K is an open-ended math problem dataset on conic sections in Chinese senior high school education.
1 paper · 0 benchmarks
Hinglish-TOP is a human annotated code-switched semantic parsing dataset containing 10k human annotations for Hindi-English (HINGLISH) code switched utterances, and over 170K CST5 generated code-switched utterances from the TOPv2 dataset.
1 paper · 0 benchmarks
Overnight is a dataset for semantic parsing in eight domains.
1 paper · 0 benchmarks
TurkQA consists of a selection of sentences from English Wikipedia articles, with questions and answers crowdsourced from workers on Amazon Mechanical Turk.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.