Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 32 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 1489–1536 of 12,172
We build a large-scale, comprehensive, and high-quality synthetic dataset for city-scale neural rendering researches.
33 papers · 0 benchmarks
MeQSum is a dataset for medical question summarization.
33 papers · 1 benchmark
Music21 is an untrimmed video dataset crawled by keyword query from Youtube.
33 papers · 0 benchmarks
The Pinterest dataset contains more than 1 million images associated to Pinterest users’ who have “pinned” them.
33 papers · 1 benchmark
SEVIR (Storm EVent ImagRy)
SEVIR is an annotated, curated and spatio-temporally aligned dataset containing over 10,000 weather events that each consist of 384 km x 384 km image sequences spanning 4 hours of time.
33 papers · 2 benchmarks
SciBench is a large-scale scientific problem-solving benchmark suite that aims to systematically examine the reasoning capabilities required for complex scientific problem solving.
33 papers · 0 benchmarks
SCIREX is a document level IE dataset that encompasses multiple IE tasks, including salient entity identification and document level N-ary relation identification from scientific articles.
33 papers · 2 benchmarks
We describe the SemEval task of extracting keyphrases and relations between them from scientific documents, which is crucial for understanding which publications describe which processes, tasks and materials.
33 papers · 1 benchmark
TallyQA is a large-scale dataset for open-ended counting.
33 papers · 2 benchmarks
A large-scale dataset for transparent object segmentation, named Trans10K, consisting of 10,428 images of real scenarios with carefully manual annotations, which are 10 times larger than the existing datasets.
33 papers · 1 benchmark
VLCS is a dataset to test for domain generalization.
33 papers · 1 benchmark
WMCA (Wide Multi Channel Presentation Attack)
The Wide Multi Channel Presentation Attack (WMCA) database consists of 1941 short video recordings of both bonafide and presentation attacks from 72 different identities.
33 papers · 1 benchmark
WMT 2015 is a collection of datasets used in shared tasks of the Tenth Workshop on Statistical Machine Translation.
33 papers · 2 benchmarks
WMT 2020 is a collection of datasets used in shared tasks of the Fifth Conference on Machine Translation.
33 papers · 0 benchmarks
WikiEvents is a document-level event extraction benchmark dataset which includes complete event and coreference annotation.
33 papers · 1 benchmark
Wrench is a benchmark platform for thorough and standardized evaluation of Weak Supervision (WS).
33 papers · 0 benchmarks
amazon-ratings is a product co-purchasing network based on data from SNAP datasets
33 papers · 1 benchmark
🤖 Robo3D - The nuScenes-C Benchmark nuScenes-C is an evaluation benchmark heading toward robust and reliable 3D perception in autonomous driving.
33 papers · 2 benchmarks
AUTSL (Ankara University Turkish Sign Language Dataset)
The Ankara University Turkish Sign Language Dataset (AUTSL) is a large-scale, multimode dataset that contains isolated Turkish sign videos.
32 papers · 1 benchmark
To analyze how these capabilities would mesh together in a natural conversation, and compare the performance of different architectures and training schemes.
32 papers · 1 benchmark
CholecT50 is a dataset of endoscopic videos of laparoscopic cholecystectomy surgery introduced to enable research on fine-grained action recognition in laparoscopic surgery.
32 papers · 5 benchmarks
Consists of (i) five relevant responses for each context and (ii) five adversarially crafted irrelevant responses for each context.
32 papers · 0 benchmarks
Dress Code is a new dataset for image-based virtual try-on composed of image pairs coming from different catalogs of YOOX NET-A-PORTER.
32 papers · 1 benchmark
EMDB contains in-the-wild videos of human activity recorded with a hand-held iPhone.
32 papers · 2 benchmarks
Ego4D is a massive-scale egocentric video dataset and benchmark suite.
32 papers · 6 benchmarks
Electricity (Individual household electric power consumption Data Set)
Abstract: Measurements of electric power consumption in one household with a one-minute sampling rate over a period of almost 4 years.
32 papers · 6 benchmarks
EmoBank is a corpus of 10k English sentences balancing multiple genres, annotated with dimensional emotion metadata in the Valence-Arousal-Dominance (VAD) representation format.
32 papers · 0 benchmarks
Current benchmarks for facial expression recognition (FER) mainly focus on static images, while there are limited datasets for FER in videos.
32 papers · 0 benchmarks
IPM NEL (Derczynski IPM Named Entity Linking)
This data is for the task of named entity recognition and linking/disambiguation over tweets.
32 papers · 1 benchmark
ImageNet-P consists of noise, blur, weather, and digital distortions.
32 papers · 1 benchmark
M3Exam is a multilingual, multimodal, and multilevel benchmark designed for evaluating Large Language Models (LLMs).
32 papers · 0 benchmarks
The MQ2007 dataset consists of queries, corresponding retrieved documents and labels provided by human experts.
32 papers · 0 benchmarks
MS-CXR (Making the Most of Text Semantics to Improve Biomedical Vision-Language Processing)
The MS-CXR dataset provides 1162 image–sentence pairs of bounding boxes and corresponding phrases, collected across eight different cardiopulmonary radiological findings, with an approximately equal number of pairs for each finding.
32 papers · 0 benchmarks
MVTec Logical Constraints Anomaly Detection (MVTec LOCO AD) dataset is intended for the evaluation of unsupervised anomaly localization algorithms.
32 papers · 1 benchmark
MedQuAD (Medical Question Answering Dataset)
MedQuAD includes 47,457 medical question-answer pairs created from 12 NIH websites (e.g.
32 papers · 0 benchmarks
ModelNet40-C is a comprehensive dataset to benchmark the corruption robustness of 3D point cloud recognition.
32 papers · 2 benchmarks
The goal of NICO Challenge is to facilitate the OOD (Out-of-Distribution) generalization in visual recognition through promoting the research on the intrinsic learning mechanisms with native invariance and generalization ability.
32 papers · 1 benchmark
NWPU-Crowd consists of 5,109 images, in a total of 2,133,375 annotated heads with points and boxes.
32 papers · 2 benchmarks
P-Stance: A Large Dataset for Stance Detection in Political Domain 2021
32 papers · 1 benchmark
PACO (Parts and Attributes of Common Objects)
Parts and Attributes of Common Objects (PACO) is a detection dataset that goes beyond traditional object boxes and masks and provides richer annotations such as part masks and attributes.
32 papers · 0 benchmarks
PIRM (Perceptual Image Restoration and Manipulation)
The PIRM dataset consists of 200 images, which are divided into two equal sets for validation and testing.
32 papers · 1 benchmark
PUBHEALTH is a comprehensive dataset for explainable automated fact-checking of public health claims.
32 papers · 0 benchmarks
SHREC (SHape REtrieval Contest)
The SHREC dataset contains 14 dynamic gestures performed by 28 participants (all participants are right handed) and captured by the Intel RealSense short range depth camera.
32 papers · 8 benchmarks
We introduce our new dataset, Spaces, to provide a more challenging shared dataset for future view synthesis research.
32 papers · 0 benchmarks
SportsMOT (SportsMOT: A Large Multi-Object Tracking Dataset in Multiple Sports Scenes)
Motivation Multi-object tracking (MOT) is a fundamental task in computer vision, aiming to estimate objects (e.g., pedestrians and vehicles) bounding boxes and identities in video sequences.
32 papers · 3 benchmarks
TinyFace is a large scale face recognition benchmark to facilitate the investigation of natively LRFR (Low Resolution Face Recognition) at large scales (large gallery population sizes) in deep learning.
32 papers · 0 benchmarks
UCF-CC-50 is a dataset for crowd counting and consists of images of extremely dense crowds.
32 papers · 0 benchmarks
UT Zappos50K is a large shoe dataset consisting of 50,025 catalog images collected from Zappos.com.
32 papers · 2 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.