Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 15 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 673–720 of 12,172
Mapillary Vistas Dataset is a diverse street-level imagery dataset with pixel‑accurate and instance‑specific human annotations for understanding street scenes around the world.
101 papers · 3 benchmarks
The McMaster dataset is a dataset for color demosaicing, which contains 18 cropped images of size 500×500.
101 papers · 5 benchmarks
Brax is a differentiable physics engine that simulates environments made up of rigid bodies, joints, and actuators.
100 papers · 0 benchmarks
CORD (Consolidated Receipt Dataset for Post-OCR Parsing)
OCR is inevitably linked to NLP since its final output is in text.
100 papers · 1 benchmark
The CUKL-SYSY dataset is a large scale benchmark for person search, containing 18,184 images and 8,432 identities.
100 papers · 2 benchmarks
ConvAI2 (Conversational Intelligence Challenge 2)
The ConvAI2 NeurIPS competition aimed at finding approaches to creating high-quality dialogue agents capable of meaningful open domain conversation.
100 papers · 1 benchmark
Extended GTEA Gaze+ EGTEA Gaze+ is a large-scale dataset for FPV actions and gaze.
100 papers · 3 benchmarks
Recent breakthroughs in diffusion models, multimodal pretraining, and efficient finetuning have led to an explosion of text-to-image generative models.
100 papers · 1 benchmark
The dataset consists of 4,738 pairs of images of 232 different scenes including reference pairs.
100 papers · 6 benchmarks
CLUE (Chinese Language Understanding Evaluation Benchmark)
CLUE is a Chinese Language Understanding Evaluation benchmark.
99 papers · 8 benchmarks
The CrowdPose dataset contains about 20,000 images and a total of 80,000 human poses with 14 labeled keypoints.
99 papers · 2 benchmarks
The Image Shadow Triplets dataset (ISTD) is a dataset for shadow understanding that contains 1870 image triplets of shadow image, shadow mask, and shadow-free image.
99 papers · 2 benchmarks
The LUNA16 (LUng Nodule Analysis) dataset is a dataset for lung segmentation.
99 papers · 0 benchmarks
The SYSU-MM01 is a dataset collected for the Visible-Infrared Re-identification problem.
99 papers · 2 benchmarks
IDD (Indian Driving Dataset)
IDD is a dataset for road scene understanding in unstructured environments used for semantic segmentation and object detection for autonomous driving.
98 papers · 1 benchmark
QuALITY (Question Answering with Long Input Texts, Yes!)
QuALITY (Question Answering with Long Input Texts, Yes!) is a multiple-choice question answering dataset for long document comprehension.
98 papers · 1 benchmark
Contains 145k captions for 28k images.
98 papers · 1 benchmark
This work presents two new benchmark datasets (CIFAR-10N, CIFAR-100N), equipping the training dataset of CIFAR-10 and CIFAR-100 with human-annotated real-world noisy labels that we collect from Amazon Mechanical Turk.
97 papers · 6 benchmarks
Kuzushiji-MNIST is a drop-in replacement for the MNIST dataset (28x28 grayscale, 70,000 images).
97 papers · 2 benchmarks
The ListOps examples are comprised of summary operations on lists of single digit integers, written in prefix notation.
97 papers · 0 benchmarks
The Medical Segmentation Decathlon is a collection of medical image segmentation datasets.
97 papers · 1 benchmark
An open access benchmark dataset comprising of 13,975 CXR images across 13,870 patient cases, with the largest number of publicly available COVID-19 positive cases to the best of the authors' knowledge.
96 papers · 1 benchmark
FIGER (Fine-Grained Entity Recognition)
The FIGER dataset is an entity recognition dataset where entities are labelled using fine-grained system 112 tags, such as person/doctor, art/writtenwork and building/hotel.
96 papers · 2 benchmarks
HM3D (Habitat-Matterport 3D)
Habitat-Matterport 3D (HM3D) is a large-scale dataset of 1,000 building-scale 3D reconstructions from a diverse set of real-world locations.
96 papers · 0 benchmarks
The ImageCLEF-DA dataset is a benchmark dataset for ImageCLEF 2014 domain adaptation challenge, which contains three domains: Caltech-256 (C), ImageNet ILSVRC 2012 (I) and Pascal VOC 2012 (P).
96 papers · 1 benchmark
Indian Pines is a Hyperspectral image segmentation dataset.
96 papers · 1 benchmark
Tudataset: A collection of benchmark datasets for learning with graphs
96 papers · 1 benchmark
RAVEN consists of 1,120,000 images and 70,000 RPM (Raven's Progressive Matrices) problems, equally distributed in 7 distinct figure configurations.
96 papers · 0 benchmarks
TORCS (The Open Racing Car Simulator)
TORCS (The Open Racing Car Simulator) is a driving simulator.
96 papers · 0 benchmarks
UAVDT (Unmanned Aerial Vehicle Benchmark Object Detection and Tracking)
UAVDT is a large scale challenging UAV Detection and Tracking benchmark (i.e., about 80, 000 representative frames from 10 hours raw videos) for 3 important fundamental tasks, i.e., object DETection (DET), Single Object Tracking (SOT) and…
96 papers · 2 benchmarks
WikiArt contains painting from 195 different artists.
96 papers · 2 benchmarks
Consists of 8,422 blurry and sharp image pairs with 65,784 densely annotated FG human bounding boxes.
95 papers · 4 benchmarks
Kinetics-700 is a video dataset of 650,000 clips that covers 700 human action classes.
95 papers · 3 benchmarks
Math23K (Math23K for Math Word Problem Solving)
Math23K is a dataset created for math word problem solving, contains 23, 162 Chinese problems crawled from the Internet.
95 papers · 1 benchmark
DexYCB is a dataset for capturing hand grasping of objects.
94 papers · 2 benchmarks
As far as we know, there only exists one large camouflaged object testing dataset, the COD10K, while the sizes of other testing datasets are less than 300.
94 papers · 1 benchmark
The sleep-edf database contains 197 whole-night PolySomnoGraphic sleep recordings, containing EEG, EOG, chin EMG, and event markers.
94 papers · 5 benchmarks
T-LESS is a dataset for estimating the 6D pose, i.e.
94 papers · 2 benchmarks
The VGG Face dataset is face identity recognition dataset that consists of 2,622 identities.
94 papers · 0 benchmarks
Aachen Day-Night is a dataset designed for benchmarking 6DOF outdoor visual localization in changing conditions.
93 papers · 1 benchmark
The CUHK-PEDES dataset is a caption-annotated pedestrian dataset.
93 papers · 3 benchmarks
JAFFE (Japanese Female Facial Expression)
The JAFFE dataset consists of 213 images of different facial expressions from 10 different Japanese female subjects.
93 papers · 4 benchmarks
SUN360 (Scene UNderstanding 360° panorama)
The goal of the SUN360 panorama database is to provide academic researchers in computer vision, computer graphics and computational photography, cognition and neuroscience, human perception, machine learning and data mining, with a…
93 papers · 1 benchmark
VOC 2012 (The PASCAL Visual Object Classes Challenge 2012)
see detailed use case on code implementation of the paper 'Tell Me Where To Look: Guided Attention Inference Networks'
93 papers · 0 benchmarks
xView is one of the largest publicly available datasets of overhead imagery.
93 papers · 1 benchmark
CBT (Children’s Book Test)
Children’s Book Test (CBT) is designed to measure directly how well language models can exploit wider linguistic context.
92 papers · 1 benchmark
The Hopkins 155 dataset consists of 156 video sequences of two or three motions.
92 papers · 1 benchmark
MEAD (A Large-scale Audio-visual Dataset for Emotional Talking-face Generation)
Multi-view Emotional Audio-visual Dataset
92 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.