Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 23 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1057–1104 of 12,172
CVC-ClinicDB is an open-access dataset of 612 images with a resolution of 384×288 from 31 colonoscopy sequences.It is used for medical image segmentation, in particular polyp detection in colonoscopy videos.
54 papers · 1 benchmark
CrossTask dataset contains instructional videos, collected for 83 different tasks.
54 papers · 1 benchmark
DAQUAR (DAtaset for QUestion Answering on Real-world images) is a dataset of human question answer pairs about images.
54 papers · 0 benchmarks
The DDIExtraction 2013 task relies on the DDI corpus which contains MedLine abstracts on drug-drug interactions as well as documents describing drug-drug interactions from the DrugBank database.
54 papers · 3 benchmarks
The Epinions dataset is built form a who-trust-whom online social network of a general consumer review site Epinions.com.
54 papers · 2 benchmarks
FGNet is a dataset for age estimation and face recognition across ages.
54 papers · 2 benchmarks
Jericho is a learning environment for man-made Interactive Fiction (IF) games.
54 papers · 0 benchmarks
Common corruptions dataset for MNIST.
54 papers · 0 benchmarks
MovieNet is a holistic dataset for movie understanding.
54 papers · 1 benchmark
The Multi-Domain Sentiment Dataset contains product reviews taken from Amazon.com from many product types (domains).
54 papers · 1 benchmark
PartiPrompts (P2) is a rich set of over 1600 prompts in English that we release as part of this work.
54 papers · 0 benchmarks
PandaSet is a dataset produced by a complete, high-precision autonomous vehicle sensor kit with a no-cost commercial license.
54 papers · 0 benchmarks
Room-Across-Room (RxR) is a multilingual dataset for Vision-and-Language Navigation (VLN) for Matterport3D environments.
54 papers · 1 benchmark
The SEMAINE videos dataset contains spontaneous data capturing the audiovisual interaction between a human and an operator undertaking the role of an avatar with four personalities: Poppy (happy), Obadiah (gloomy), Spike (angry) and…
54 papers · 1 benchmark
UAVid is a high-resolution UAV semantic segmentation dataset as a complement, which brings new challenges, including large scale variation, moving object recognition and temporal consistency preservation.
54 papers · 2 benchmarks
WikiSum is a dataset based on English Wikipedia and suitable for a task of multi-document abstractive summarization.
54 papers · 0 benchmarks
Wikidata5m is a million-scale knowledge graph dataset with aligned corpus.
54 papers · 1 benchmark
BEHAVE is a full body human-object interaction dataset with multi-view RGBD frames and corresponding 3D SMPL and object fits along with the annotated contacts between them.
53 papers · 3 benchmarks
This corpus includes annotations of cancer-related PubMed articles, covering 3 full papers (PMID:24651010, PMID:11777939, PMID:15630473) as well as the result sections of 46 additional PubMed papers.
53 papers · 1 benchmark
CUHK02 (CUHK Person Re-identification Dataset)
CUHK02 is a dataset for person re-identification.
53 papers · 0 benchmarks
DRCD (Delta Reading Comprehension Dataset)
Delta Reading Comprehension Dataset (DRCD) is an open domain traditional Chinese machine reading comprehension (MRC) dataset.
53 papers · 0 benchmarks
FED (Fine-grained Evaluation of Dialog)
The FED dataset is constructed by annotating a set of human-system and human-human conversations with eighteen fine-grained dialog qualities.
53 papers · 0 benchmarks
HRF (High-Resolution Fundus)
The HRF dataset is a dataset for retinal vessel segmentation which comprises 45 images and is organized as 15 subsets.
53 papers · 3 benchmarks
Large language models (LLMs), after being aligned with vision models and integrated into vision-language models (VLMs), can bring impressive improvement in image reasoning tasks.
53 papers · 1 benchmark
The ICDAR2003 dataset is a dataset for scene text recognition.
53 papers · 1 benchmark
MLDoc (Multilingual Document Classification Corpus)
Multilingual Document Classification Corpus (MLDoc) is a cross-lingual document classification dataset covering English, German, French, Spanish, Italian, Russian, Japanese and Chinese.
53 papers · 8 benchmarks
The MMVP (Multimodal Visual Patterns) Benchmark focuses on identifying "CLIP-blind pairs" – images that appear similar to the CLIP model despite having clear visual differences.
53 papers · 1 benchmark
MathInstruct is a meticulously curated instruction tuning dataset that combines data from 13 mathematical rationale datasets.
53 papers · 0 benchmarks
PAQ (Probably Asked Questions)
Probably Asked Questions (PAQ) is a very large resource of 65M automatically-generated QA-pairs.
53 papers · 0 benchmarks
PRM800K is a process supervision dataset containing 800,000 step-level correctness labels for model-generated solutions to problems from the MATH dataset.
53 papers · 0 benchmarks
RICH (Real scenes, Interaction, Contact and Humans)
Inferring human-scene contact (HSC) is the first step toward understanding how humans interact with their surroundings.
53 papers · 1 benchmark
A benchmark for action spotting in soccer videos.
53 papers · 1 benchmark
Consists of 100 challenging video sequences captured from real-world traffic scenes (over 140,000 frames with rich annotations, including occlusion, weather, vehicle category, truncation, and vehicle bounding boxes) for object detection,…
53 papers · 2 benchmarks
Virtual KITTI 2 is an updated version of the well-known Virtual KITTI dataset which consists of 5 sequence clones from the KITTI tracking benchmark.
53 papers · 2 benchmarks
VoiceBank + DEMAND (Noisy speech database for training speech enhancement algorithms and TTS models)
VoiceBank+DEMAND is a noisy speech database for training speech enhancement algorithms and TTS models.
53 papers · 1 benchmark
ACDC (Automated Cardiac Diagnosis Challenge)
The goal of the Automated Cardiac Diagnosis Challenge (ACDC) challenge is to: - compare the performance of automatic methods on the segmentation of the left ventricular endocardium and epicardium as the right ventricular endocardium for…
52 papers · 5 benchmarks
EntailmentBank is a dataset that contains multistep entailment trees.
52 papers · 0 benchmarks
ICDAR 2015 was a scene text detection used for the ICDAR 2015 conference.
52 papers · 2 benchmarks
InfographicVQA is a dataset that comprises a diverse collection of infographics along with natural language questions and answers annotations.
52 papers · 1 benchmark
The Machine Translation of Noisy Text (MTNT) dataset is a Machine Translation dataset that consists of noisy comments on Reddit and professionally sourced translation.
52 papers · 0 benchmarks
The MultiMNIST dataset is generated from MNIST.
52 papers · 1 benchmark
NJU2K is a large RGB-D dataset containing 1,985 image pairs.
52 papers · 1 benchmark
The Object Discovery dataset was collected by downloading images from Internet for airplane, car and horse.
52 papers · 1 benchmark
Pick-a-Pic dataset was created by logging user interactions with the Pick-a-Pic web application for text-to image generation.
52 papers · 0 benchmarks
There exist previous works [6, 10] that constructed referring segmentation datasets for videos.
52 papers · 3 benchmarks
The Scene UNderstanding (SUN) database contains 899 categories and 130,519 images.
52 papers · 8 benchmarks
Sprites (2D Video Game Character Sprites)
The Sprites dataset contains 60 pixel color images of animated characters (sprites).
52 papers · 3 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.