Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 29 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1345–1392 of 12,172
LTCC contains 17,119 person images of 152 identities, and each identity is captured by at least two cameras.
38 papers · 2 benchmarks
ManiSkill is a large-scale learning-from-demonstrations benchmark for articulated object manipulation with visual input (point cloud and image).
38 papers · 0 benchmarks
We advocate the use of curated, comprehensive benchmark suites of machine learning datasets, backed by standardized OpenML-based interfaces and complementary software toolkits written in Python, Java and R.
38 papers · 0 benchmarks
The RGBT234 dataset is a comprehensive video dataset specifically designed for RGB-T (Red-Green-Blue and Thermal) tracking purposes.
38 papers · 1 benchmark
The SUN Attribute dataset consists of 14,340 images from 717 scene categories, and each category is annotated with a taxonomy of 102 discriminate attributes.
38 papers · 2 benchmarks
The Synthesized Lakh (Slakh) Dataset is a dataset for audio source separation that is synthesized from the Lakh MIDI Dataset v0.1 using professional-grade sample-based virtual instruments.
38 papers · 3 benchmarks
Tatoeba is a free collection of example sentences with translations geared towards foreign language learners.
38 papers · 2 benchmarks
TyDiQA is the gold passage version of the Typologically Diverse Question Answering (TyDiWA) dataset, a benchmark for information-seeking question answering, which covers nine languages.
38 papers · 1 benchmark
URMP (University of Rochester Multi-Modal Musical Performance)
URMP (University of Rochester Multi-Modal Musical Performance) is a dataset for facilitating audio-visual analysis of musical performances.
38 papers · 2 benchmarks
WORD (Whole abdominal Organs Dataset)
WORD is a dataset for organ semantic segmentation that contains 150 abdominal CT volumes (30,495 slices) and each volume has 16 organs with fine pixel-level annotations and scribble-based sparse annotation, which may be the largest dataset…
38 papers · 0 benchmarks
WaveFake is a dataset for audio deepfake detection.
38 papers · 0 benchmarks
Tolokers is a crowdsourcing platform workers network based on data provided by Toloka.
38 papers · 1 benchmark
The A3D dataset is a step forward to make autonomous driving safer for pedestrians and the public in the real world.
37 papers · 0 benchmarks
CV-Bench (Cambrian Vision-Centric Benchmark)
The Cambrian Vision-Centric Benchmark (CV-Bench) is designed to address the limitations of existing vision-centric benchmarks by providing a comprehensive evaluation framework for multimodal large language models (MLLMs).
37 papers · 0 benchmarks
The EgoGesture dataset contains 2,081 RGB-D videos, 24,161 gesture samples and 2,953,224 frames from 50 distinct subjects.
37 papers · 2 benchmarks
HOC (Hallmarks of Cancer)
The Hallmarks of Cancer (HOC) corpus consists of 1852 PubMed publication abstracts manually annotated by experts according to the Hallmarks of Cancer taxonomy.
37 papers · 1 benchmark
HumanAct12 is a new 3D human motion dataset adopted from the polar image and 3D pose dataset PHSPD, with proper temporal cropping and action annotating.
37 papers · 2 benchmarks
Learn2Reg is a dataset for medical image registration.
37 papers · 2 benchmarks
NN-HAZE is an image dehazing dataset.
37 papers · 2 benchmarks
Open Images V4 offers large scale across several dimensions: 30.1M image-level labels for 19.8k concepts, 15.4M bounding boxes for 600 object classes, and 375k visual relationship annotations involving 57 classes.
37 papers · 1 benchmark
PIPAL (Perceptual Image Processing ALgorithms IQA Dataset)
PIPAL training set contains 200 reference images, 40 distortion types, 23k distortion images, and more than one million human ratings.
37 papers · 0 benchmarks
PeMS04 is a traffic forecasting benchmark.
37 papers · 2 benchmarks
The SQA dataset was created to explore the task of answering sequences of inter-related questions on HTML tables.
37 papers · 1 benchmark
The SemEval-2018 hypernym discovery evaluation benchmark (Camacho-Collados et al.
37 papers · 3 benchmarks
TextOCR is a dataset to benchmark text recognition on arbitrary shaped scene-text.
37 papers · 0 benchmarks
WMT 2018 is a collection of datasets used in shared tasks of the Third Conference on Machine Translation.
37 papers · 4 benchmarks
Watercolor2k is a dataset used for cross-domain object detection which contains 2k watercolor images with image and instance-level annotations.
37 papers · 3 benchmarks
The Yelp Reviews Polarity dataset is obtained from the Yelp Dataset Challenge in 2015 (1,569,264 samples that have review text).
37 papers · 0 benchmarks
node classification on genius
37 papers · 2 benchmarks
iHarmony4 is a synthesized dataset for Image Harmonization.
37 papers · 1 benchmark
3DSSG provides 3D semantic scene graphs for 3RScan.
36 papers · 1 benchmark
Adversarial GLUE (AdvGLUE) is a new multi-task benchmark to quantitatively and thoroughly explore and evaluate the vulnerabilities of modern large-scale language models under various types of adversarial attacks.
36 papers · 1 benchmark
BRATS 2013 is a brain tumor segmentation dataset consists of synthetic and real images, where each of them is further divided into high-grade gliomas (HG) and low-grade gliomas (LG).
36 papers · 2 benchmarks
CelebV-HQ is a large-scale video facial attributes dataset with annotations.
36 papers · 3 benchmarks
The color FERET database is a dataset for face recognition.
36 papers · 3 benchmarks
DRealSR (Diverse Real-world image Super-Resolution)
DRealSR establishes a Super Resolution (SR) benchmark with diverse real-world degradation processes, mitigating the limitations of conventional simulated image degradation.
36 papers · 1 benchmark
DialogRE is the first human-annotated dialogue-based relation extraction dataset, containing 1,788 dialogues originating from the complete transcripts of a famous American television situation comedy Friends.
36 papers · 1 benchmark
Doc2Dial (Doc2Dial: Document-grounded Dialogue)
For goal-oriented document-grounded dialogs, it often involves complex contexts for identifying the most relevant information, which requires better understanding of the inter-relations between conversations and documents.
36 papers · 0 benchmarks
EBM-NLP annotates PICO (Participants, Interventions, Comparisons and Outcomes) spans in clinical trial abstracts.
36 papers · 1 benchmark
Fashion-Gen consists of 293,008 high definition (1360 x 1360 pixels) fashion images paired with item descriptions provided by professional stylists.
36 papers · 0 benchmarks
GVGAI (General Video Game AI)
The General Video Game AI (GVGAI) framework is widely used in research which features a corpus of over 100 single-player games and 60 two-player games.
36 papers · 0 benchmarks
A hand-object interaction dataset with 3D pose annotations of hand and object.
36 papers · 2 benchmarks
In this project, we introduce InfoSeek, a visual question answering dataset tailored for information-seeking questions that cannot be answered with only common sense knowledge.
36 papers · 2 benchmarks
JTA is a dataset for people tracking in urban scenarios by exploiting a photorealistic videogame.
36 papers · 1 benchmark
LSUI (Large Scale Underwater Image Dataset)
We released a large-scale underwater image (LSUI) dataset including 5004 image pairs, which involve richer underwater scenes (lighting conditions, water types and target categories) and better visual quality reference images than the…
36 papers · 1 benchmark
The Lakh MIDI dataset is a collection of 176,581 unique MIDI files, 45,129 of which have been matched and aligned to entries in the Million Song Dataset.
36 papers · 0 benchmarks
MAD (Movie Audio Descriptions) is an automatically curated large-scale dataset for the task of natural language grounding in videos or natural language moment retrieval.
36 papers · 2 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.