Home › Datasets › modality › Tabular
Tabular datasets
archive 2025-07-28
267 datasets carry the modality tag "Tabular", ordered by the archive's paper count. Page 1 of 6: 48 shown of 267. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Tabular datasets 1–48 of 267
The IMDb Movie Reviews dataset is a binary sentiment analysis dataset consisting of 50,000 reviews from the Internet Movie Database (IMDb) labeled as positive or negative.
1,787 papers · 9 benchmarks
The MovieLens datasets, first released in 1998, describe people’s expressed preferences for movies.
1,246 papers · 17 benchmarks
MIMIC-III (The Medical Information Mart for Intensive Care III)
The Medical Information Mart for Intensive Care III (MIMIC-III) dataset is a large, de-identified and publicly-available collection of medical records.
1,041 papers · 8 benchmarks
Netflix Prize consists of about 100,000,000 ratings for 17,770 movies given by 480,189 users.
370 papers · 1 benchmark
UNSW-NB15 is a network intrusion dataset.
156 papers · 3 benchmarks
NAS-Bench-101 is the first public architecture dataset for NAS research.
152 papers · 1 benchmark
WikiTableQuestions is a question answering dataset over semi-structured tables.
79 papers · 2 benchmarks
The data is related with direct marketing campaigns (phone calls) of a Portuguese banking institution.
69 papers · 0 benchmarks
Data Set Information: Extraction was done by Barry Becker from the 1994 Census database.
56 papers · 2 benchmarks
The Elephant MIL dataset is a benchmark used in multiple instance learning (MIL), which falls under the broader categories of image classification and content-based image retrieval.
47 papers · 1 benchmark
This dataset contains complex tables from the annual reports of S&P 500 companies with detailed table structure annotations to help table structure recognition and table data extraction.
35 papers · 0 benchmarks
The friedman1 data set is commonly used to test semi-supervised regression methods.
31 papers · 0 benchmarks
Pulsar candidates collected during the HTRU survey.
27 papers · 0 benchmarks
TCGA (The Cancer Genome Atlas)
23 papers · 2 benchmarks
This dataset contains card descriptions of the card game Hearthstone and the code that implements them.
22 papers · 0 benchmarks
Retrospectively collected medical data has the opportunity to improve patient care through knowledge discovery and algorithm development.
22 papers · 0 benchmarks
CAL500 (Computer Audition Lab 500)
CAL500 (Computer Audition Lab 500) is a dataset aimed for evaluation of music information retrieval systems.
21 papers · 0 benchmarks
Two datasets are provided.
21 papers · 0 benchmarks
The Amazon-Google dataset for entity resolution derives from the online retailers Amazon.com and the product search service of Google accessible through the Google Base Data API.
20 papers · 2 benchmarks
The Abt-Buy dataset for entity resolution derives from the online retailers Abt.com and Buy.com.
19 papers · 2 benchmarks
OpenXAI is the first general-purpose lightweight library that provides a comprehensive list of functions to systematically evaluate the quality of explanations generated by attribute-based explanation methods.
17 papers · 0 benchmarks
The T2Dv2 dataset consists of 779 tables originating from the English-language subset of the WebTables corpus.
14 papers · 4 benchmarks
ACS PUMS stands for American Community Survey (ACS) Public Use Microdata Sample (PUMS) and has been used to construct several tabular datasets for studying fairness in machine learning: - ACSIncome: to predict whether an individual’s…
11 papers · 0 benchmarks
The ToughTables (2T) dataset was created for the SemTab challenge and includes 180 tables in total.
11 papers · 4 benchmarks
Bank Account Fraud (BAF) is a large-scale, realistic suite of tabular datasets.
10 papers · 12 benchmarks
HANNA (HANNA, a large annotated dataset of Human-ANnotated NArratives for ASG evaluation.)
HANNA, a large annotated dataset of Human-ANnotated NArratives for Automatic Story Generation (ASG) evaluation, has been designed for the benchmarking of automatic metrics for ASG.
8 papers · 0 benchmarks
The dataset contains transactions made by credit cards in September 2013 by European cardholders.
8 papers · 2 benchmarks
Many e-shops have started to mark-up product data within their HTML pages using the schema.org vocabulary.
8 papers · 4 benchmarks
SOTAB V2 features two annotation tasks: Column Type Annotation (CTA) and Columns Property Annotation (CPA).
7 papers · 2 benchmarks
The WikiTables-TURL dataset was constructed by the authors of TURL and is based on the WikiTable corpus, which is a large collection of Wikipedia tables.
7 papers · 3 benchmarks
This data was extracted from the 1994 Census bureau database by Ronny Kohavi and Barry Becker (Data Mining and Visualization, Silicon Graphics).
6 papers · 1 benchmark
AnoShift (AnoShift: A Distribution Shift Benchmark for Unsupervised Anomaly Detection)
AnoShift is a large-scale anomaly detection benchmark, which focuses on splitting the test data based on its temporal distance to the training set, introducing three testing splits: IID, NEAR, and FAR.
6 papers · 1 benchmark
Median house prices for California districts derived from the 1990 census.
6 papers · 2 benchmarks
Concepticon (Concepticon. A Resource for the Linking of Concept Lists)
This resource, our Concepticon, links concept labels from different conceptlists to concept sets.
6 papers · 0 benchmarks
WDC Products is an entity matching benchmark which provides for the systematic evaluation of matching systems along combinations of three dimensions while relying on real-word data.
6 papers · 4 benchmarks
Diabetes (Diabetes 130-US Hospitals for Years 1999-2008)
What do the instances in this dataset represent?
5 papers · 3 benchmarks
ARAUS (Affective Responses to Augmented Urban Soundscapes)
Choosing optimal maskers for existing soundscapes to effect a desired perceptual change via soundscape augmentation is non-trivial due to extensive varieties of maskers and a dearth of benchmark datasets with which to compare and develop…
4 papers · 0 benchmarks
CI-MNIST (Correlated and Imbalanced MNIST)
CI-MNIST (Correlated and Imbalanced MNIST) is a variant of MNIST dataset with introduced different types of correlations between attributes, dataset features, and an artificial eligibility criterion.
4 papers · 0 benchmarks
DrivAerNet (A Parametric Car Dataset for Data-driven Aerodynamic Design and Graph-Based Drag Prediction)
DrivAerNet is a large-scale, high-fidelity CFD dataset of 3D industry-standard car shapes designed for data-driven aerodynamic design.
4 papers · 0 benchmarks
Human Activity Recognition (HAR) refers to the capacity of machines to perceive human actions.
4 papers · 0 benchmarks
MIMIC-IV-ED is a large, freely available database of emergency department (ED) admissions at the Beth Israel Deaconess Medical Center between 2011 and 2019.
4 papers · 0 benchmarks
The Musk dataset describes a set of molecules, and the objective is to detect musks from non-musks.
4 papers · 2 benchmarks
The Musk2 dataset is a set of 102 molecules of which 39 are judged by human experts to be musks and the remaining 63 molecules are judged to be non-musks.
4 papers · 1 benchmark
RTASC (ROBIN Technical Acquisition Speech Corpus)
The ROBIN Technical Acquisition Speech Corpus (ROBINTASC) was developed within the ROBIN project.
4 papers · 0 benchmarks
A dataset of real distributional shift across multiple large-scale tasks.
4 papers · 0 benchmarks
VizNet-Sato is a dataset from the authors of Sato and is based on the VizNet dataset.
4 papers · 2 benchmarks
The WikipediaGS dataset was created by extracting Wikipedia tables from Wikipedia pages.
4 papers · 2 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.