Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 24 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1105–1152 of 12,172
The Weibo NER dataset is a Chinese Named Entity Recognition dataset drawn from the social media website Sina Weibo.
52 papers · 2 benchmarks
WikiLingua includes ~770k article and summary pairs in 18 languages from WikiHow.
52 papers · 1 benchmark
3,859 high-resolution YouTube videos, 2,985 training videos, 421 validation videos and 453 test videos.
52 papers · 1 benchmark
ACE 2004 (ACE 2004 Multilingual Training Corpus)
ACE 2004 Multilingual Training Corpus contains the complete set of English, Arabic and Chinese training data for the 2004 Automatic Content Extraction (ACE) technology evaluation.
51 papers · 6 benchmarks
Rendered synthetically using a library of standard 3D objects, and tests the ability to recognize compositions of object movements that require long-term reasoning.
51 papers · 3 benchmarks
The Event-Camera Dataset is a collection of datasets with an event-based camera for high-speed robotics.
51 papers · 2 benchmarks
Fishyscapes is a public benchmark for uncertainty estimation in a real-world task of semantic segmentation for urban driving.
51 papers · 2 benchmarks
KuaiRec is a real-world dataset collected from the recommendation logs of the video-sharing mobile app Kuaishou.
51 papers · 0 benchmarks
The large-scale MUSIC-AVQA dataset of musical performance contains 45,867 question-answer pairs, distributed in 9,288 videos for over 150 hours.
51 papers · 1 benchmark
PubTabNet is a large dataset for image-based table recognition, containing 568k+ images of tabular data annotated with the corresponding HTML representation of the tables.
51 papers · 1 benchmark
QReCC contains 14K conversations with 81K question-answer pairs.
51 papers · 0 benchmarks
SiW (Spoofing in the Wild)
SiW provides live and spoof videos from 165 subjects.
51 papers · 1 benchmark
The Snow100K dataset consists of 1) 100k synthesized snowy images 2) corresponding snow-free ground truth images 3) snow masks 4) 1,329 realistic snowy images The images of 2) and 4) were downloaded via the Flickr api, and were manually…
51 papers · 0 benchmarks
Representation and learning of commonsense knowledge is one of the foundational problems in the quest to enable deep language understanding.
51 papers · 1 benchmark
The Tumblr GIF (TGIF) dataset contains 100K animated GIFs and 120K sentences describing visual content of the animated GIFs.
51 papers · 1 benchmark
TOFU (Task of Fictitious Unlearning)
The TOFU dataset serves as a benchmark for evaluating unlearning performance of large language models on realistic tasks.
51 papers · 0 benchmarks
The xBD dataset contains over 45,000KM2 of polygon labeled pre and post disaster imagery.
51 papers · 2 benchmarks
Consists of 1.3 million records of U.S.
50 papers · 2 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
50 papers · 1 benchmark
DrivingStereo contains over 180k images covering a diverse set of driving scenarios, which is hundreds of times larger than the KITTI Stereo dataset.
50 papers · 0 benchmarks
ECQA (Explanations for CommonsenseQA)
This repository contains the publicly released dataset, code, and models for the Explanations for CommonsenseQA paper presented at ACL-IJCNLP 2021.
50 papers · 0 benchmarks
Lesion segmentation data includes the original image, paired with the expert manual tracing of the lesion boundaries in the form of a binary mask.
50 papers · 1 benchmark
The InterHand2.6M dataset is a large-scale real-captured dataset with accurate GT 3D interacting hand poses, used for 3D hand pose estimation The dataset contains 2.6M labeled single and interacting hand frames.
50 papers · 2 benchmarks
NomBank is an annotation project at New York University that is related to the PropBank project at the University of Colorado.
50 papers · 0 benchmarks
PlotQA is a VQA dataset with 28.9 million question-answer pairs grounded over 224,377 plots on data from real-world sources and questions based on crowd-sourced question templates.
50 papers · 5 benchmarks
The Quick Draw Dataset is a collection of 50 million drawings across 345 categories, contributed by players of the game Quick, Draw!.
50 papers · 0 benchmarks
Quoref is a QA dataset which tests the coreferential reasoning capability of reading comprehension systems.
50 papers · 0 benchmarks
emrQA has 1 million question-logical form and 400,000+ questionanswer evidence pairs.
50 papers · 0 benchmarks
2018 Data Science Bowl (2018 Data Science Bowl Find the nuclei in divergent images to advance medical discovery)
This dataset contains a large number of segmented nuclei images.
49 papers · 1 benchmark
CMRC 2018 (Chinese Machine Reading Comprehension 2018)
CMRC 2018 is a dataset for Chinese Machine Reading Comprehension.
49 papers · 0 benchmarks
DENSE (Depth Estimation oN Synthetic Events)
DENSE (Depth Estimation oN Synthetic Events) is a new dataset with synthetic events and perfect ground truth.
49 papers · 1 benchmark
DVQA (Data Visualizations via Question Answering)
DVQA is a synthetic question-answering dataset on images of bar-charts.
49 papers · 1 benchmark
The DukeMTMC-VideoReID (Duke Multi-Tracking Multi-Camera Video-based ReIDentification) dataset is a subset of the DukeMTMC for video-based person re-ID.
49 papers · 2 benchmarks
FSS-1000 is a 1000 class dataset for few-shot segmentation.
49 papers · 1 benchmark
The Foursquare dataset consists of check-in data for different cities.
49 papers · 0 benchmarks
Letter (Letter Recognition Data Set)
Letter Recognition Data Set is a handwritten digit dataset.
49 papers · 2 benchmarks
MKQA (Multilingual Knowledge Questions and Answers)
Multilingual Knowledge Questions and Answers (MKQA) is an open-domain question answering evaluation set comprising 10k question-answer pairs aligned across 26 typologically diverse languages (260k question-answer pairs in total).
49 papers · 0 benchmarks
MMKG is a collection of three knowledge graphs for link prediction and entity matching research.
49 papers · 3 benchmarks
MOSE (Complex Video Object Segmentation)
CoMplex video Object SEgmentation (MOSE) is a dataset to study the tracking and segmenting objects in complex environments.
49 papers · 2 benchmarks
The OPUS-MT benchmark is a systematic collection of results from these models, focusing on verifiable translation performance and large coverage in terms of languages and domains.
49 papers · 0 benchmarks
The Pavia University dataset is a hyperspectral image dataset which gathered by a sensor known as the reflective optics system imaging spectrometer (ROSIS-3) over the city of Pavia, Italy.
49 papers · 1 benchmark
A new large-scale benchmark consisting of both synthetic and real-world hazy images, called REalistic Single Image DEhazing (RESIDE).
49 papers · 5 benchmarks
TAO (Tracking Any Object Dataset)
TAO is a federated dataset for Tracking Any Object, containing 2,907 high resolution videos, captured in diverse environments, which are half a minute long on average.
49 papers · 1 benchmark
WildDeepfake is a dataset for real-world deepfakes detection which consists of 7,314 face sequences extracted from 707 deepfake videos that are collected completely from the internet.
49 papers · 0 benchmarks
The inD dataset is a new dataset of naturalistic vehicle trajectories recorded at German intersections.
49 papers · 0 benchmarks
Roman-empire is a word dependency graph based on the Roman Empire article from the English Wikipedia.
49 papers · 1 benchmark
3D-FUTURE (3D FUrniture shape with TextURE) is a 3D dataset that contains 20,240 photo-realistic synthetic images captured in 5,000 diverse scenes, and 9,992 involved unique industrial 3D CAD shapes of furniture with high-resolution…
48 papers · 0 benchmarks
It's a synthetic dataset, which contains 1000 graphs divided into two classes according to the motif they contain: either a “house” or a five-node cycle.
48 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.