Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 26 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1201–1248 of 12,172
CoSQL (Conversational Text-to-SQL Challenge)
CoSQL is a corpus for building cross-domain, general-purpose database (DB) querying dialogue systems.
45 papers · 1 benchmark
DART is a large dataset for open-domain structured data record to text generation.
45 papers · 3 benchmarks
DFEW (Dynamic Facial Expression in the Wild)
Recently, facial expression recognition (FER) in the wild has gained a lot of researchers’ attention because it is a valuable topic to enable the FER techniques to move from the laboratory to the real applications.
45 papers · 0 benchmarks
GEM (Generation, Evaluation, and Metrics)
Generation, Evaluation, and Metrics (GEM) is a benchmark environment for Natural Language Generation with a focus on its Evaluation, both through human annotations and automated Metrics.
45 papers · 1 benchmark
A new large-scale geometry problem-solving dataset - 3,002 multi-choice geometry problems - dense annotations in formal language for the diagrams and text - 27,213 annotated diagram logic forms (literals) - 6,293 annotated text logic forms…
45 papers · 1 benchmark
HICO (Humans Interacting with Common Objects)
HICO is a benchmark for recognizing human-object interactions (HOI).
45 papers · 1 benchmark
A new dataset of 1,001 human-human dialogs for movie recommendation with measures for successful recommendations.
45 papers · 0 benchmarks
LasHeR consists of 1224 visible and thermal infrared video pairs with more than 730K frame pairs in total.
45 papers · 1 benchmark
LiTS17 (Liver Tumor Segmentation Challenge 2017)
LiTS17 is a liver tumor segmentation benchmark.
45 papers · 3 benchmarks
MLSUM (MultiLingual SUMmarization)
A large-scale MultiLingual SUMmarization dataset.
45 papers · 4 benchmarks
DailyActivity3D dataset is a daily activity dataset captured by a Kinect device.
45 papers · 1 benchmark
MVTec 3D Anomaly Detection Dataset (MVTec 3D-AD) is a comprehensive 3D dataset for the task of unsupervised anomaly detection and localization.
45 papers · 4 benchmarks
RSTPReid (Real Scenario Text-based Person Re-identification)
RSTPReid contains 20505 images of 4,101 persons from 15 cameras.
45 papers · 2 benchmarks
STS Benchmark comprises a selection of the English datasets used in the STS tasks organized in the context of SemEval between 2012 and 2017.
45 papers · 6 benchmarks
The ScanNet200 benchmark studies 200-class 3D semantic segmentation - an order of magnitude more class categories than previous 3D scene understanding benchmarks.
45 papers · 3 benchmarks
Augments the video-description dataset TACoS with short and single sentence descriptions.
45 papers · 1 benchmark
The Talk2Car dataset finds itself at the intersection of various research domains, promoting the development of cross-disciplinary solutions for improving the state-of-the-art in grounding natural language into visual space.
45 papers · 0 benchmarks
This data set was prepared from 88 open-source YouTube cooking videos.
45 papers · 0 benchmarks
A novel dataset and benchmark, which features 1482 RGB-D scans of 478 environments across multiple time steps.
44 papers · 4 benchmarks
Contains aesthetic scores and meaningful attributes assigned to each image by multiple human raters.
44 papers · 0 benchmarks
The dataset contains over 15K images of 20 people (6 females and 14 males - 4 people were recorded twice).
44 papers · 1 benchmark
COCO-O(ut-of-distribution) contains 6 domains (sketch, cartoon, painting, weather, handmake, tattoo) of COCO objects which are hard to be detected by most existing detectors.
44 papers · 1 benchmark
EmotionLines contains a total of 29245 labeled utterances from 2000 dialogues.
44 papers · 1 benchmark
GPTFuzzer is a fascinating project that explores red teaming of large language models (LLMs) using auto-generated jailbreak prompts.
44 papers · 0 benchmarks
How2Sign (A Large-scale Multimodal Dataset for Continuous American Sign Language)
The How2Sign is a multimodal and multiview continuous American Sign Language (ASL) dataset consisting of a parallel corpus of more than 80 hours of sign language videos and a set of corresponding modalities including speech, English…
44 papers · 3 benchmarks
MC-TACO is a dataset of 13k question-answer pairs that require temporal commonsense comprehension.
44 papers · 0 benchmarks
MPDD (Metal Parts Defect Detection Dataset)
MPDD is a dataset aimed at benchmarking visual defect detection methods in industrial metal parts manufacturing.
44 papers · 1 benchmark
OCNLI (Original Chinese Natural Language Inference)
OCNLI stands for Original Chinese Natural Language Inference.
44 papers · 0 benchmarks
Occ3D is a dataset for 3D occupancy prediction, which aims to estimate the detailed occupancy and semantics of objects from multi-view images.
44 papers · 1 benchmark
Oxford105k is the combination of the Oxford5k dataset and 99782 negative images crawled from Flickr using 145 most popular tags.
44 papers · 0 benchmarks
The PanoContext dataset contains 500 annotated cuboid layouts of indoor environments such as bedrooms and living rooms.
44 papers · 1 benchmark
This dataset contains the traffic data in San Bernardino from July to August in 2016, with 170 detectors on 8 roads with a time interval of 5 minutes.
44 papers · 1 benchmark
The SCUT-CTW1500 dataset contains 1,500 images: 1,000 for training and 500 for testing.
44 papers · 3 benchmarks
TurkCorpus, a dataset with 2,359 original sentences from English Wikipedia, each with 8 manual reference simplifications.
44 papers · 1 benchmark
For understanding multimodal language used in expressing humor.
44 papers · 0 benchmarks
CARS196 is composed of 16,185 car images of 196 classes.
43 papers · 4 benchmarks
Dataset contains 33,010 molecule-description pairs split into 80\%/10\%/10\% train/val/test splits.
43 papers · 4 benchmarks
DensePASS - a novel densely annotated dataset for panoramic segmentation under cross-domain conditions, specifically built to study the Pinhole-to-Panoramic transfer and accompanied with pinhole camera training examples obtained from…
43 papers · 1 benchmark
A large-scale hierarchical dataset of diverse student activities collected by Santa, a multi-platform self-study solution equipped with artificial intelligence tutoring system.
43 papers · 1 benchmark
Powered by the ImageNet dataset, unsupervised learning on large-scale data has made significant advances for classification tasks.
43 papers · 6 benchmarks
Although large language models (LLMs) demonstrate impressive performance for many language tasks, most of them can only handle texts a few thousand tokens long, limiting their applications on longer sequence inputs, such as books, reports,…
43 papers · 1 benchmark
A large dataset of musculoskeletal radiographs containing 40,561 images from 14,863 studies, where each study is manually labeled by radiologists as either normal or abnormal.
43 papers · 0 benchmarks
MusicNet is a collection of 330 freely-licensed classical music recordings, together with over 1 million annotated labels indicating the precise time of each note in every recording, the instrument that plays each note, and the note's…
43 papers · 1 benchmark
A unified benchmark on searching for both topology and size, for (almost) any up-to-date NAS algorithm.
43 papers · 5 benchmarks
Occluded-DukeMTMC contains 15,618 training images, 17,661 gallery images, and 2,210 occluded query images.
43 papers · 1 benchmark
POP909 is a dataset which contains multiple versions of the piano arrangements of 909 popular songs created by professional musicians.
43 papers · 0 benchmarks
Accurate modeling of priors over 3D human pose is fundamental to many problems in computer vision.
43 papers · 0 benchmarks
QA-SRL was proposed as an open schema for semantic roles, in which the relation between an argument and a predicate is expressed as a natural-language question containing the predicate (“Where was someone educated?”) whose answer is the…
43 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.