Datasets › WebQuestionsSP

WebQuestionsSP (WebQuestions Semantic Parses Dataset)

Introduced by Wen-tau Yih et al. in The Value of Semantic Parse Labeling for Knowledge Base Question Answering1 Aug 2016 archive 2025-07-28

The WebQuestionsSP dataset is released as part of our ACL-2016 paper “The Value of Semantic Parse Labeling for Knowledge Base Question Answering” [Yih, Richardson, Meek, Chang & Suh, 2016], in which we evaluated the value of gathering semantic parses, vs. answers, for a set of questions that originally comes from WebQuestions [Berant et al., 2013]. The WebQuestionsSP dataset contains full semantic parses in SPARQL queries for 4,737 questions, and “partial” annotations for the remaining 1,073 questions for which a valid parse could not be formulated or where the question itself is bad or needs a descriptive answer. This release also includes an evaluation script and the output of the STAGG semantic parsing system when trained using the full semantic parses. More detail can be found in the document and labeling instructions included in this release, as well as the paper.

Source: WebQuestions Semantic Parses Dataset

Benchmarks archive 2025-07-28

All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

12 shown of 12 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 61. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
SPARKLE: Enhancing SPARQL Generation with Direct KG Integration in Decoding 1 1 29 Jun 2024 not harvested
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 1 1 26 Jun 2024 ran 1 of 3 samples (2 unverified; 3 pointer-only for licence)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models 1 1 13 Oct 2023 ran 8 of 16 samples (8 unverified)
Bridging the KB-Text Gap: Leveraging Structured Knowledge-aware Pre-training for KBQA 1 1 28 Aug 2023 ran 10 of 10 samples (0 unverified; 10 pointer-only for licence)
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 2 1 14 Mar 2023 not harvested
ReaRev: Adaptive Reasoning for Question Answering over Knowledge Graphs 1 1 24 Oct 2022 not harvested
ReTraCk: A Flexible and Efficient Framework for Knowledge Base Question Answering 1 2 1 Aug 2021 not harvested
Case-based Reasoning for Natural Language Queries over Knowledge Bases 0 1 18 Apr 2021 not harvested
Improving Multi-hop Knowledge Base Question Answering by Learning Intermediate Supervision Signals 1 1 11 Jan 2021 not harvested
UniK-QA: Unified Representations of Structured and Unstructured Knowledge for Open-Domain Question Answering 1 2 29 Dec 2020 not harvested
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer 57 1 23 Oct 2019 ran 2 of 31 samples (29 unverified)
The Value of Semantic Parse Labeling for Knowledge Base Question Answering 0 1 1 Aug 2016 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • WebQSP-WD
  • WebQuestionsSP

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections