Home › Datasets › task › Text-To-SQL

Text-To-SQL datasets

archive 2025-07-28

21 datasets carry the task tag "Text-To-SQL" (the task itself: Text-To-SQL), ordered by the archive's paper count. Page 1 of 1: 21 shown of 21. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Text-To-SQL datasets 1–21 of 21

KITTI (Karlsruhe Institute of Technology and Toyota Technological Institute) is one of the most popular datasets for use in mobile robotics and autonomous driving.
3,661 papers · 137 benchmarks
Spider dataset is used for evaluation in the paper "Structure-Grounded Pretraining for Text-to-SQL".
91 papers · 3 benchmarks
A new large-scale question-answering dataset that requires reasoning on heterogeneous information.
70 papers · 1 benchmark
SParC (Semantic Parsing in Context)
SParC is a large-scale dataset for complex, cross-domain, and context-dependent (multi-turn) semantic parsing and text-to-SQL task (interactive natural language interfaces for relational databases).
59 papers · 2 benchmarks
CoSQL (Conversational Text-to-SQL Challenge)
CoSQL is a corpus for building cross-domain, general-purpose database (DB) querying dialogue systems.
45 papers · 1 benchmark
KaggleDBQA (KaggleDBQA: Realistic Text-to-SQL dataset)
KaggleDBQA is a challenging cross-domain and complex evaluation dataset of real Web databases, with domain-specific data types, original formatting, and unrestricted questions.
28 papers · 1 benchmark
BIRD (BIg Bench for LaRge-scale Database Grounded Text-to-SQL Evaluation) represents a pioneering, cross-domain dataset that examines the impact of extensive database contents on text-to-SQL parsing.
15 papers · 1 benchmark
SEDE (Stack Exchange Data Explorer)
SEDE is a dataset comprised of 12,023 complex and diverse SQL queries and their natural language titles and descriptions, written by real users of the Stack Exchange Data Explorer out of a natural interaction.
11 papers · 1 benchmark
Spider 2.0 is a comprehensive code generation agent task that includes 632 examples.
8 papers · 1 benchmark
A dataset of utterances, incorrect SQL interpretations and the corresponding natural language feedback.
6 papers · 0 benchmarks
ADVErsarial Table perturbAtion (ADVETA) is a robustness evaluation benchmark featuring natural and realistic ATPs.
5 papers · 0 benchmarks
MultiSpider is a large multilingual text-to-SQL dataset which covers seven languages (English, German, French, Spanish, Japanese, Chinese, and Vietnamese).
4 papers · 0 benchmarks
TriageSQL is a cross-domain text-to-SQL question intention classification benchmark that requires models to distinguish four types of unanswerable questions from answerable questions.
3 papers · 0 benchmarks
SQL-Eval is an open-source PostgreSQL evaluation dataset released by Defog, constructed based on Spider.
2 papers · 1 benchmark
ViText2SQL is a dataset for the Vietnamese Text-to-SQL semantic parsing task, consisting of about 10K question and SQL query pairs.
2 papers · 0 benchmarks
A synthetic dataset from an automobile manufacturer datasource.
1 paper · 0 benchmarks
MMSQL (Multi-Turn Multi-Type Text-to-SQL test suit)
A dataset for training and testing tin various problem types and multi-turn Q&A scenarios, including a training set, test set, and test scripts.
1 paper · 1 benchmark
The field of converting natural language into corresponding SQL queries using deep learning techniques has attracted significant attention in recent years.
1 paper · 0 benchmarks
TURSpider (TURSpider: A Turkish Text-to-SQL Dataset)
TURSpider is a novel Turkish Text-to-SQL dataset that includes complex queries, akin to those in the original Spider dataset.
0 papers · 0 benchmarks
TURSpider is a novel Turkish Text-to-SQL dataset that includes complex queries, akin to those in the original Spider dataset.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.