Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 40 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1873–1920 of 12,172
HPS (Human POSEitioning System Dataset)
HPS Dataset is a collection of 3D humans interacting with large 3D scenes (300-1000 m², up to 2500 m²).
23 papers · 0 benchmarks
HRSOD (High-Resolution Salient Object Detection)
There exist several datasets for saliency detection, but none of them is specifically designed for high-resolution salient object detection.
23 papers · 1 benchmark
HeadQA is a multi-choice question answering testbed to encourage research on complex reasoning.
23 papers · 1 benchmark
The ISNotes dataset is a corpus used for fine-grained Information Status (IS) classification.
23 papers · 0 benchmarks
ITOP (Invariant-Top View Dataset)
The ITOP dataset consists of 40K training and 10K testing depth images for each of the front-view and top-view tracks.
23 papers · 3 benchmarks
IXI (IXI Brain Development Dataset)
IXI Dataset is a collection of 600 MR brain images from normal, healthy subjects.
23 papers · 4 benchmarks
ImageNet-W(atermark) is a test set to evaluate models’ reliance on the newly found watermark shortcut in ImageNet, which is used to predict the carton class.
23 papers · 0 benchmarks
KAIST-Radar (K-Radar) is a novel large-scale object detection dataset and benchmark that contains 35K frames of 4D Radar tensor (4DRT) data with power measurements along the Doppler, range, azimuth, and elevation dimensions, together with…
23 papers · 0 benchmarks
KAIST Multispectral Pedestrian Dataset The KAIST Multispectral Pedestrian Dataset is imaging hardware consisting of a color camera, a thermal camera and a beam splitter to capture the aligned multispectral (RGB color + Thermal) images.
23 papers · 1 benchmark
A large-scale dataset for Complex KBQA.
23 papers · 1 benchmark
LitBank is an annotated dataset of 100 works of English-language fiction to support tasks in natural language processing and the computational humanities, described in more detail in the following publications: - David Bamman, Sejal Popat…
23 papers · 1 benchmark
MMAct is a large-scale dataset for multi/cross modal action understanding.
23 papers · 1 benchmark
Simplified Chinese dataset for NER in The Third International Chinese Language Processing Bakeoff (2006), provided by Microsoft Research Asia (MSRA).
23 papers · 3 benchmarks
We study visually grounded VideoQA in response to the emerging trends of utilizing pretraining techniques for video-language understanding.
23 papers · 1 benchmark
The NetHack Learning Environment (NLE) is a Reinforcement Learning environment based on NetHack 3.6.6.
23 papers · 1 benchmark
The NumGLUE dataset is a valuable resource developed by the Allen Institute for AI.
23 papers · 0 benchmarks
OpenImages V6 is a large-scale dataset , consists of 9 million training images, 41,620 validation samples, and 125,456 test samples.
23 papers · 2 benchmarks
The PASCAL FACE dataset is a dataset for face detection and face recognition.
23 papers · 1 benchmark
PIE-Bench (Prompt-based Image Editing Benchmark)
PIE-Bench comprises 700 images featuring 10 distinct editing types.
23 papers · 1 benchmark
The PhysioNet Challenge 2012 dataset is publicly available and contains the de-identified records of 8000 patients in Intensive Care Units (ICU).
23 papers · 5 benchmarks
PointCloud-C is the very first test-suite for point cloud robustness analysis under corruptions.
23 papers · 2 benchmarks
RECCON is a dataset for the task of recognizing emotion cause in conversations.
23 papers · 2 benchmarks
RST-DT (RST Discourse Treebank)
The Rhetorical Structure Theory (RST) Discourse Treebank consists of 385 Wall Street Journal articles from the Penn Treebank annotated with discourse structure in the RST framework along with human-generated extracts and abstracts…
23 papers · 2 benchmarks
Real 3D-AD is the first point cloud anomaly detection dataset for industrial products.
23 papers · 2 benchmarks
Question: I have five fingers but I am not alive.
23 papers · 1 benchmark
ShapeWorld is a new evaluation methodology and framework for multimodal deep learning models, with a focus on formal-semantic style generalization capabilities.
23 papers · 0 benchmarks
SituatedQA is an open-retrieval QA dataset where systems must produce the correct answer to a question given the temporal or geographical context.
23 papers · 0 benchmarks
Social-IQ is an unconstrained benchmark specifically designed to train and evaluate socially intelligent technologies.
23 papers · 0 benchmarks
TCGA (The Cancer Genome Atlas)
23 papers · 2 benchmarks
Question answering over knowledge graphs (KG-QA) is a vital topic in IR.
23 papers · 1 benchmark
Toyota Smarthome Trimmed has been designed for the activity classification task of 31 activities.
23 papers · 0 benchmarks
This is a 21 class land use image dataset meant for research purposes.
23 papers · 1 benchmark
VisEvent (Visible-Event benchmark) is a dataset constructed for the evaluation of tracking by combing visible and event cameras.
23 papers · 1 benchmark
WeatherBench 2 is an update to the global, medium-range (1–14 day) weather forecasting benchmark proposed by raspweatherbench2020, designed with the aim to accelerate progress in data-driven weather modeling.
23 papers · 0 benchmarks
mMARCO is a multilingual version of the MS MARCO passage ranking dataset comprising 8 languages that was created using machine translation.
23 papers · 0 benchmarks
Over 4 million frames of motion capture data for 100 different styles of locomotion.
22 papers · 0 benchmarks
Segmentation of robotic instruments is an important problem for robotic assisted minimially invasive surgery.
22 papers · 2 benchmarks
Car CAD models from "3d object detection and viewpoint estimation with a deformable 3d cuboid model" were used to generate the dataset.
22 papers · 0 benchmarks
A3D (AnAn Accident Detection)
A new dataset of diverse traffic accidents.
22 papers · 1 benchmark
To study the task of email subject line generation: automatically generating an email subject line from the email body.
22 papers · 1 benchmark
514 algebra word problems and associated equation systems gathered from Algebra.com.
22 papers · 1 benchmark
AMR Bank (Abstract Meaning Representation)
The AMR Bank is a set of English sentences paired with simple, readable semantic representations.
22 papers · 1 benchmark
Contains temporally labeled face tracks in video, where each face instance is labeled as speaking or not, and whether the speech is audible.
22 papers · 1 benchmark
The BACE dataset focuses on inhibitors of human beta-secretase 1 (BACE-1).
22 papers · 4 benchmarks
BAR (Biased Action Recognition)
Biased Action Recognition (BAR) dataset is a real-world image dataset categorized as six action classes which are biased to distinct places.
22 papers · 1 benchmark
BookTest is a new dataset similar to the popular Children’s Book Test (CBT), however more than 60 times larger.
22 papers · 0 benchmarks
BurstSR is a dataset consisting of smartphone bursts and high-resolution DSLR ground-truth
22 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.