Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 55 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 2593–2640 of 12,172
Node classification on Texas with the fixed 48%/32%/20% splits provided by Geom-GCN.
14 papers · 2 benchmarks
TimeTravel contains 29,849 counterfactual rewritings, each with the original story, a counterfactual event, and human-generated revision of the original story compatible with the counterfactual event.
14 papers · 1 benchmark
TransNAS-Bench-101 is a Neural Architecture Search (NAS) benchmark dataset containing network performance across seven tasks, covering classification, regression, pixel-level prediction, and self-supervised tasks.
14 papers · 0 benchmarks
The UCR Anomaly Archive is a collection of 250 uni-variate time series collected in human medicine, biology, meteorology and industry.
14 papers · 2 benchmarks
UrbanLoco is a mapping/localization dataset collected in highly-urbanized environments with a full sensor-suite.
14 papers · 0 benchmarks
V3C (Vimeo Creative Commons Collection)
The Vimeo Creative Commons Collection, in short V3C, is a collection of 28’450 videos (with overall length of about 3’800 h) published under creative commons license on Vimeo.
14 papers · 0 benchmarks
VegFru is a domain-specific dataset for fine-grained visual categorization.
14 papers · 0 benchmarks
ViTT (Video Timeline Tags)
The ViTT dataset consists of human produced segment-level annotations for 8,169 videos.
14 papers · 2 benchmarks
VisIT-Bench is a new vision-language instruction following benchmark inspired by real-world use cases.
14 papers · 0 benchmarks
Created for MVS tasks and is a large-scale multi-view aerial dataset generated from a highly accurate 3D digital surface model produced from thousands of real aerial images with precise camera parameters.
14 papers · 0 benchmarks
The WNLI dataset is a part of the GLUE benchmark used for Natural Language Inference (NLI).
14 papers · 1 benchmark
mvor (Multi-View Operating Room)
Multi-View Operating Room (MVOR) dataset consists of 732 synchronized multi-view frames recorded by three RGB-D cameras in a hybrid OR during real clinical interventions.
14 papers · 0 benchmarks
speechocean762 is an open-source speech corpus designed for pronunciation assessment use, consisting of 5000 English utterances from 250 non-native speakers, where half of the speakers are children.
14 papers · 3 benchmarks
the 160x160 subset of the GasHisSDB dataset.
13 papers · 0 benchmarks
This dataset accompanies our paper on synthesizing the 3D Ken Burns effect from a single image.
13 papers · 0 benchmarks
Provides a wide range of raw sensor data that is accessible on almost any modern-day smartphone together with a high-quality ground-truth track.
13 papers · 0 benchmarks
ANTIQUE is a collection of 2,626 open-domain non-factoid questions from a diverse set of categories.
13 papers · 0 benchmarks
ASAP (Aligned Scores and Performances)
ASAP is a dataset of 222 digital musical scores aligned with 1068 performances (more than 92 hours) of Western classical piano music.
13 papers · 2 benchmarks
This dataset includes reviews (ratings, text, helpfulness votes), product metadata (descriptions, category information, price, brand, and image features), and links (also viewed/also bought graphs).
13 papers · 2 benchmarks
A large-scale cloze-style biomedical MRC dataset.
13 papers · 1 benchmark
This dataset was introduced by the TrackNetV2 work.
13 papers · 1 benchmark
BRATS 2016 is a brain tumor segmentation dataset.
13 papers · 0 benchmarks
BrixIA Covid-19 is a large dataset of CXR images corresponding to the entire amount of images taken for both triage and patient monitoring in sub-intensive and intensive care units during one month (between March 4th and April 4th 2020) of…
13 papers · 0 benchmarks
The dataset contains 21 full-HD videos, each around 1 hr long, captured at six different locations.
13 papers · 1 benchmark
CBIS-DDSM (Curated Breast Imaging Subset of Digital Database for Screening Mammography)
This CBIS-DDSM (Curated Breast Imaging Subset of DDSM) is an updated and standardized version of the Digital Database for Screening Mammography (DDSM) .
13 papers · 2 benchmarks
Source: ICDAR 2019 CROHME + TFD: Competition on Recognition of Handwritten Mathematical Expressions and Typeset Formula Detection
13 papers · 1 benchmark
A large set of images of cats and dogs.
13 papers · 5 benchmarks
CirCor DigiScope is currently the largest pediatric heart sound dataset.
13 papers · 2 benchmarks
Detecting vehicles and representing their position and orientation in the three dimensional space is a key technology for autonomous driving.
13 papers · 3 benchmarks
DNA-Rendering is a large-scale, high-fidelity repository of human performance data for neural actor rendering.
13 papers · 0 benchmarks
the dataset contains data about hydrogen storage in metal hydrides
13 papers · 2 benchmarks
The Dayton dataset is a dataset for ground-to-aerial (or aerial-to-ground) image translation, or cross-view image synthesis.
13 papers · 4 benchmarks
DialFact is a testing benchmark dataset of 22,245 annotated conversational claims, paired with pieces of evidence from Wikipedia.
13 papers · 0 benchmarks
EGFxSet (Electric Guitar Effects Dataset, ISMIR 2022 LBD)
EGFxSet (Electric Guitar Effects dataset) features recordings for all clean tones in a 22-fret Stratocaster, recorded with 5 different pickup configurations, also processed through 12 popular guitar effects.
13 papers · 0 benchmarks
EPRSTMT (E-commerce Product Review Dataset for Sentiment Analysis)
The EPRSTMT dataset, also known as EPR-sentiment, is a binary sentiment analysis dataset based on product reviews on an e-commerce platform.
13 papers · 0 benchmarks
This dataset contains benchmark scores for EQ-Bench, a novel benchmark designed to evaluate aspects of emotional intelligence in Large Language Models (LLMs).
13 papers · 1 benchmark
EVE (End-to-end Video-based Eye-tracking)
EVE (End-to-end Video-based Eye-tracking) is a dataset for eye-tracking.
13 papers · 0 benchmarks
The EgoDexter dataset provides both 2D and 3D pose annotations for 4 testing video sequences with 3190 frames.
13 papers · 0 benchmarks
EuRoC MAV is a visual-inertial datasets collected on-board a Micro Aerial Vehicle (MAV).
13 papers · 1 benchmark
Large-scale single-object tracking dataset, containing 108 sequences with a total length of 1.5 hours.
13 papers · 1 benchmark
FaceVerse (FaceVerse-High Quality 3D Face Dataset)
FaceVerse-High Quality 3D Face Dataset contains 2,688 high-quality head scans (21 expressions from 128 identities) captured by a dense DLSR rig.
13 papers · 0 benchmarks
FewGLUE consists of a random selection of 32 training examples from the SuperGLUE training sets and up to 20,000 unlabeled examples for each SuperGLUE task.
13 papers · 0 benchmarks
Flare7K, the first nighttime flare removal dataset, which is generated based on the observation and statistic of real-world nighttime lens flares.
13 papers · 1 benchmark
GCDC (Grammarly Corpus of Discourse Coherence)
A corpus of real-world texts.
13 papers · 2 benchmarks
Prophesee’s GEN1 Automotive Detection Dataset is the largest Event-Based Dataset to date.
13 papers · 1 benchmark
GUE (Genome Understanding Evaluation)
A collection of $28$ datasets across $7$ tasks constructed for genome language model evaluation.
13 papers · 7 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.