12,172 datasets listed, ordered by the archive's paper count. Page 6 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
GPQA stands for Graduate-Level Google-Proof Q&A Benchmark.
270 papers · 1 benchmark
ImageNet-Sketch data set consists of 50,889 images, approximately 50 images for each of the 1000 ImageNet classes.
268 papers · 3 benchmarks
WikiSQL consists of a corpus of 87,726 hand-annotated SQL query and natural language question pairs.
267 papers · 4 benchmarks
Description: 10,000 People - Human Pose Recognition Data.
265 papers · 1 benchmark
DensePose-COCO is a large-scale ground-truth dataset with image-to-surface correspondences manually annotated on 50K COCO images and train DensePose-RCNN, to densely regress part-specific UV coordinates within every human region at…
265 papers · 1 benchmark
AwA (Animals with Attributes)
Animals with Attributes (AwA) was a dataset for benchmarking transfer-learning algorithms, in particular attribute base classification.
264 papers · 3 benchmarks
EMNIST (extended MNIST) has 4 times more data than MNIST.
264 papers · 10 benchmarks
The New York Times Annotated Corpus contains over 1.8 million articles written and published by the New York Times between January 1, 1987 and June 19, 2007 with article metadata provided by the New York Times Newsroom, the New York Times…
262 papers · 9 benchmarks
BSDS500 (Berkeley Segmentation Dataset 500)
Berkeley Segmentation Data Set 500 (BSDS500) is a standard benchmark for contour detection.
261 papers · 8 benchmarks
NAS-Bench-201 is a benchmark (and search space) for neural architecture search.
260 papers · 4 benchmarks
The NCI1 dataset comes from the cheminformatics domain, where each input graph is used as representation of a chemical compound: each vertex stands for an atom of the molecule, and edges between vertices represent bonds between atoms.
260 papers · 2 benchmarks
The VizWiz-VQA dataset originates from a natural visual question answering setting where blind people each took an image and recorded a spoken question about it, together with 10 crowdsourced answers per visual question.
260 papers · 7 benchmarks
COLLAB is a scientific collaboration dataset.
259 papers · 2 benchmarks
The LOL dataset is composed of 500 low-light and normal-light image pairs and divided into 485 training pairs and 15 testing pairs.
257 papers · 2 benchmarks
LibriTTS is a multi-speaker English corpus of approximately 585 hours of read English speech at 24kHz sampling rate, prepared by Heiga Zen with the assistance of Google Speech and Google Brain team members.
257 papers · 1 benchmark
The MS-Celeb-1M dataset is a large-scale face recognition dataset consists of 100K identities, and each identity has about 100 facial images.
257 papers · 0 benchmarks
WebVid contains 10 million video clips with captions, sourced from the web.
257 papers · 1 benchmark
SNIPS (SNIPS Natural Language Understanding benchmark)
The SNIPS Natural Language Understanding benchmark is a dataset of over 16,000 crowdsourced queries distributed among 7 user intents of various complexity: SearchCreativeWork (e.g.
256 papers · 6 benchmarks
The ActivityNet Captions dataset is built on ActivityNet v1.3 which includes 20k YouTube untrimmed videos with 100k caption annotations.
255 papers · 6 benchmarks
OntoNotes 5.0 is a large corpus comprising various genres of text (news, conversational telephone speech, weblogs, usenet newsgroups, broadcast, talk shows) in three languages (English, Chinese, and Arabic) with structural information…
254 papers · 12 benchmarks
ZINC is a free database of commercially-available compounds for virtual screening.
251 papers · 5 benchmarks
100DOH (100 Days Of Hands Dataset)
The 100 Days Of Hands Dataset (100DOH) is a large-scale video dataset containing hands and hand-object interactions.
249 papers · 0 benchmarks
Foggy Cityscapes is a synthetic foggy dataset which simulates fog on real scenes.
249 papers · 7 benchmarks
JHMDB (Joint-annotated Human Motion Data Base)
JHMDB is an action recognition dataset that consists of 960 video sequences belonging to 21 actions.
249 papers · 9 benchmarks
The HPatches is a recent dataset for local patch descriptor evaluation that consists of 116 sequences of 6 images with known homography.
248 papers · 4 benchmarks
The ICDAR 2013 dataset consists of 229 training images and 233 testing images, with word-level annotations provided.
246 papers · 3 benchmarks
IJB-C (IARPA Janus Benchmark-C)
The IJB-C dataset is a video-based face recognition dataset.
246 papers · 3 benchmarks
KITTI-360 is a large-scale dataset that contains rich sensory information and full annotations.
246 papers · 7 benchmarks
SIDD (Smartphone Image Denoising Dataset)
SIDD is an image denoising dataset containing 30,000 noisy images from 10 scenes under different lighting conditions using five representative smartphone cameras.
245 papers · 2 benchmarks
AI2-Thor is an interactive environment for embodied AI.
243 papers · 1 benchmark
IMDB-MULTI is a relational dataset that consists of a network of 1000 actors or actresses who played roles in movies in IMDB.
243 papers · 3 benchmarks
The UTKFace dataset is a large-scale face dataset with long age span (range from 0 to 116 years old).
243 papers · 4 benchmarks
YFCC100M is a that dataset contains a total of 100 million media objects, of which approximately 99.2 million are photos and 0.8 million are videos, all of which carry a Creative Commons license.
243 papers · 0 benchmarks
MathVista (Mathematical Reasoning of in Visual Contexts)
MathVista is a consolidated Mathematical reasoning benchmark within Visual contexts.
242 papers · 0 benchmarks
The WebQuestions dataset is a question answering dataset using Freebase as the knowledge base and contains 6,642 question-answer pairs.
241 papers · 4 benchmarks
The LIDC-IDRI dataset contains lesion annotations from four experienced thoracic radiologists.
240 papers · 6 benchmarks
MIMIC-CXR from Massachusetts Institute of Technology presents 371,920 chest X-rays associated with 227,943 imaging studies from 65,079 patients.
240 papers · 3 benchmarks
MoleculeNet is a large scale benchmark for molecular machine learning.
240 papers · 1 benchmark
Omniverse Isaac Gym is a GPU-based physics simulation platform developed by NVIDIA.
240 papers · 2 benchmarks
GOT-10k (Generic Object Tracking Benchmark)
The GOT-10k dataset contains more than 10,000 video segments of real-world moving objects and over 1.5 million manually labelled bounding boxes.
239 papers · 2 benchmarks
CK+ (Extended Cohn-Kanade dataset)
The Extended Cohn-Kanade (CK+) dataset contains 593 video sequences from a total of 123 different subjects, ranging from 18 to 50 years of age with a variety of genders and heritage.
238 papers · 2 benchmarks
ChestX-ray14 is a medical imaging dataset which comprises 112,120 frontal-view X-ray images of 30,805 (collected from the year of 1992 to 2015) unique patients with the text-mined fourteen common disease labels, mined from the text…
237 papers · 6 benchmarks
The Pascal3D+ multi-view dataset consists of images in the wild, i.e., images of object categories exhibiting high variability, captured under uncontrolled settings, in cluttered scenes and under many different poses.
237 papers · 1 benchmark
The Sketch dataset contains over 20,000 sketches evenly distributed over 250 object categories.
237 papers · 1 benchmark
Charades-STA is a new dataset built on top of Charades by adding sentence temporal annotations.
236 papers · 4 benchmarks
TUM RGB-D is an RGB-D dataset.
235 papers · 1 benchmark
AwA2 (Animals with Attributes 2)
Animals with Attributes 2 (AwA2) is a dataset for benchmarking transfer-learning algorithms, such as attribute base classification and zero-shot learning.
231 papers · 5 benchmarks
DAVIS16 is a dataset for video object segmentation which consists of 50 videos in total (30 videos for training and 20 for testing).
231 papers · 4 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.