Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 77 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 3649–3696 of 12,172
HBW (Human Bodies in the Wild)
Human Bodies in the Wild (HBW) is a validation and test set for body shape estimation.
7 papers · 0 benchmarks
The Human Related version of UBnormal ("UBnormal: New Benchmark for Supervised Open-Set Video Anomaly Detection," Acsintoae et al.) was introduced by Flaborea et al.
7 papers · 1 benchmark
The HandNet dataset contains depth images of 10 participants' hands non-rigidly deforming in front of a RealSense RGB-D camera.
7 papers · 0 benchmarks
HiREST (HIerarchical REtrieval and STep-captioning)
HiREST (HIerarchical REtrieval and STep-captioning) dataset is a benchmark that covers hierarchical information retrieval and visual/textual stepwise summarization from an instructional video corpus.
7 papers · 0 benchmarks
Hindi Visual Genome is a multimodal dataset consisting of text and images suitable for English-Hindi multimodal machine translation task and multimodal research.
7 papers · 0 benchmarks
Horse-10 is an animal pose estimation dataset.
7 papers · 1 benchmark
The Hotels-50K dataset consists of over 1 million images from 50,000 different hotels around the world.
7 papers · 0 benchmarks
Houston is a hyperspectral image classification dataset.
7 papers · 1 benchmark
Human-Art is a versatile human-centric dataset to bridge the gap between natural and artificial scenes.
7 papers · 1 benchmark
HyperRED (Hyper-Relational Extraction Dataset)
HyperRED is a dataset for the new task of hyper-relational extraction, which extracts relation triplets together with qualifier information such as time, quantity or location.
7 papers · 1 benchmark
ICB (Image Compression Benchmark)
A carefully chosen set of high-resolution high-precision natural images suited for compression algorithm evaluation.
7 papers · 6 benchmarks
IQUAD (Interactive Question Answering Dataset)
IQUAD is a dataset for Visual Question Answering in interactive environments.
7 papers · 0 benchmarks
Includes 500 categories from the list in the Wikipedia and 399,726 images, a more comprehensive food dataset that surpasses existing popular benchmark datasets by category coverage and data volume.
7 papers · 0 benchmarks
ISRUC-Sleep is a polysomnographic (PSG) dataset.
7 papers · 2 benchmarks
The IWSLT 2017 translation dataset.
7 papers · 1 benchmark
This dataset contains 5955 painting images (from WikiCommons) : a train set of 2978 images and a test set of 2977 images (for classification task).
7 papers · 1 benchmark
ImageNet-9 consists of images with different amounts of background and foreground signal, which you can use to measure the extent to which your models rely on image backgrounds.
7 papers · 1 benchmark
Paper: Improved automatic keyword extraction given more linguistic knowledge Doi: 10.3115/1119355.1119383
7 papers · 2 benchmarks
Interiorverse is a high-quality indoor scene dataset with rich details, including complex furniture and decorations and it is rendered with GGX BRDF model, which has stronger material modeling capability than any BRDF models.
7 papers · 0 benchmarks
JGLUE, Japanese General Language Understanding Evaluation, is built to measure the general NLU ability in Japanese.
7 papers · 0 benchmarks
Dataset for document shadow removal
7 papers · 0 benchmarks
The odometry benchmark consists of 22 stereo sequences, saved in loss less png format: We provide 11 sequences (00-10) with ground truth trajectories for training and 11 sequences (11-21) without ground truth for evaluation.
7 papers · 1 benchmark
KITTI360-EX is a dataset for outer- and inner FoV expansion.
7 papers · 1 benchmark
KUAKE Query-Query Relevance, a dataset used to evaluate the relevance of the content expressed in two queries, is used for the KUAKE-QQR task.
7 papers · 1 benchmark
KUAKE Query Title Relevance, a dataset used to estimate the relevance of the title of a query document, is used for the KUAKE-QTR task.
7 papers · 1 benchmark
The KUMC dataset for polyp detection and classification was collected from the University of Kansas Medical Center.
7 papers · 0 benchmarks
The Kannada-MNIST dataset is a drop-in substitute for the standard MNIST dataset for the Kannada language.
7 papers · 0 benchmarks
Consists of faces extracted from pre-modern Japanese artwork.
7 papers · 0 benchmarks
KiTS19 (The 2019 Kidney and Kidney Tumor Segmentation Challenge)
The 2021 Kidney and Kidney Tumor Segmentation challenge (abbreviated KiTS21) is a competition in which teams compete to develop the best system for automatic semantic segmentation of renal tumors and surrounding anatomy.
7 papers · 1 benchmark
Language-molecule models have emerged as an exciting direction for molecular discovery and understanding.
7 papers · 1 benchmark
LDV (Large-scale Diverse Video)
LDV is a dataset for video enhancement.
7 papers · 0 benchmarks
LEVEN (Legal Event Detection Dataset)
Overview LEVEN is the largest Legal Event Detection dataset as well as the largest Chinese Event Detection dataset.
7 papers · 0 benchmarks
LIVE-ETRI (ETRI-LIVE Space-Time Subsampled Video Quality (STSVQ) Database)
The video deployed parameter space is continuously increasing to provide more realistic and immersive experiences to global streaming and social media viewers.
7 papers · 1 benchmark
LSVTD is a large scale video text dataset for promoting the video text spotting community, which contains 100 text videos from 22 different real-life scenarios.
7 papers · 0 benchmarks
LaFAN1 (Ubisoft La Forge Animation Dataset)
Ubisoft La Forge Animation Dataset ("LAFAN1") Ubisoft La Forge Animation dataset and accompanying code for the SIGGRAPH 2020 paper Robust Motion In-betweening.
7 papers · 1 benchmark
Large Scale Composed Image Retrieval (LaSCo) is a new dataset for Composed Image Retrieval (CoIR), x10 times larger than current ones.
7 papers · 1 benchmark
A new question answering dataset constructed from play-by-play live broadcast.
7 papers · 0 benchmarks
This work proposes Long-RVOS, a large-scale benchmark for long-term video object segmentation.
7 papers · 1 benchmark
A self-driving dataset for motion prediction, containing over 1,000 hours of data.
7 papers · 0 benchmarks
MAPS (Midi Aligned Piano Dataset)
MAPS – standing for MIDI Aligned Piano Sounds – is a database of MIDI-annotated piano recordings.
7 papers · 1 benchmark
MIPE (Improving Paratope and Epitope Prediction by Multi-Modal Contrastive Learning and Interaction Informativeness Estimation)
Datasets.
7 papers · 1 benchmark
MLPF (Simulated particle-level dataset of ttbar with PU200 using Pythia8+Delphes3 for machine learned particle flow (MLPF))
Dataset of 50,000 top quark-antiquark (ttbar) events produced in proton-proton collisions at 14 TeV, overlaid with minimum bias events corresponding to a pileup of 200 on average.
7 papers · 0 benchmarks
The MMBody dataset provides human body data with motion capture, GT mesh, Kinect RGBD, and millimeter wave sensor data.
7 papers · 0 benchmarks
The MMVP-VLM (Multimodal Visual Patterns - Visual Language Models) Benchmark is specifically designed to systematically evaluate the performance of recent CLIP-based models in understanding and processing visual patterns.
7 papers · 0 benchmarks
This is a 3D action recognition dataset, also known as 3D Action Pairs dataset.
7 papers · 1 benchmark
MSSD (Music Streaming Sessions Dataset)
The Spotify Music Streaming Sessions Dataset (MSSD) consists of 160 million streaming sessions with associated user interactions, audio features and metadata describing the tracks streamed during the sessions, and snapshots of the…
7 papers · 1 benchmark
MSU BASED (MSU BASED Video Deblurring Dataset and Benchmark)
Qualitative dataset with real blurred videos, created by using beam-splitter setup in lab environment
7 papers · 1 benchmark
Collects the data by scraping Wikipedia and then utilize crowdsourcing to collect question-answer pairs.
7 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.