Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 25 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 1153–1200 of 3,239
The 7-Scenes dataset is a collection of tracked RGB-D camera frames.
7 papers · 0 benchmarks
Aachen Day-Night v1.1 dataset is an extended version of the original Aachen Day-Night dataset.
7 papers · 1 benchmark
We introduce ArtBench-10, the first class-balanced, high-quality, cleanly annotated, and standardized dataset for benchmarking artwork generation.
7 papers · 1 benchmark
ArtiFact (Artificial and Factual Image Dataset for Synthetic Image Detection)
The ArtiFact dataset is a large-scale image dataset that aims to include a diverse collection of real and synthetic images from multiple categories, including Human/Human Faces, Animal/Animal Faces, Places, Vehicles, Art, and many other…
7 papers · 0 benchmarks
The Atari Grand Challenge dataset is a large dataset of human Atari 2600 replays.
7 papers · 0 benchmarks
BANDON is a dataset for building change detection with off-nadir aerial images dataset, which is composed of off-Nadir image pairs of urban and rural areas.
7 papers · 0 benchmarks
Bamboo Dataset is a mega-scale and information-dense dataset for both classification and detection pre-training.
7 papers · 0 benchmarks
BigDetection is a new large-scale benchmark to build more general and powerful object detection systems.
7 papers · 1 benchmark
BraTS 2020 (RSNA-ASNR-MICCAI Brain Tumor Segmentation BraTS Challenge)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
7 papers · 0 benchmarks
CASME II (Chinese Academy of Sciences Micro-Expression II)
The Chinese Academy of Sciences Micro-Expression dataset (CASME II) consists of 255 videos, elicited from 26 participants.
7 papers · 1 benchmark
ChangeSim is a dataset aimed at online scene change detection (SCD) and more.
7 papers · 2 benchmarks
The CheXmask Database presents a comprehensive, uniformly annotated collection of chest radiographs, constructed from five public databases: ChestX-ray8, Chexpert, MIMIC-CXR-JPG, Padchest and VinDr-CXR.
7 papers · 0 benchmarks
The Cross-dataset Testbed is a Decaf7 based cross-dataset image classification dataset, which contains 40 categories of images from 3 domains: 3,847 images in Caltech256, 4,000 images in ImageNet, and 2,626 images for SUN.
7 papers · 0 benchmarks
Danish Fungi 2020 (DF20) is a fine-grained dataset and benchmark.
7 papers · 1 benchmark
The Diabetic Foot Ulcers dataset (DFUC2021) is a dataset for analysis of pathology, focusing on infection and ischaemia.
7 papers · 0 benchmarks
DFW (Disguised Faces in the Wild)
Contains over 11000 images of 1000 identities with different types of disguise accessories.
7 papers · 3 benchmarks
DOLPHINS (Dataset for Collaborative Perception enabled Harmonious and Interconnected Self-driving)
Vehicle-to-Everything (V2X) network has enabled collaborative perception in autonomous driving, which is a promising solution to the fundamental defect of stand-alone intelligence including blind zones and long-range perception.
7 papers · 0 benchmarks
FPL (First-Person Locomotion)
Supports new task that predicts future locations of people observed in first-person videos.
7 papers · 0 benchmarks
FRLL-Morphs is a dataset of morphed faces based on images selected from the publicly available Face Research London Lab dataset [1].
7 papers · 0 benchmarks
Freiburg Groceries is a groceries classification dataset consisting of 5000 images of size 256x256, divided into 25 categories.
7 papers · 0 benchmarks
GeoDE is a geographically diverse dataset with 61,940 images from 40 classes and 6 world regions, and no personally identifiable information, collected through crowd-sourcing.
7 papers · 0 benchmarks
Grocery Store is a dataset of natural images of grocery items.
7 papers · 0 benchmarks
HBW (Human Bodies in the Wild)
Human Bodies in the Wild (HBW) is a validation and test set for body shape estimation.
7 papers · 0 benchmarks
The HandNet dataset contains depth images of 10 participants' hands non-rigidly deforming in front of a RealSense RGB-D camera.
7 papers · 0 benchmarks
Hindi Visual Genome is a multimodal dataset consisting of text and images suitable for English-Hindi multimodal machine translation task and multimodal research.
7 papers · 0 benchmarks
Horse-10 is an animal pose estimation dataset.
7 papers · 1 benchmark
The Hotels-50K dataset consists of over 1 million images from 50,000 different hotels around the world.
7 papers · 0 benchmarks
Houston is a hyperspectral image classification dataset.
7 papers · 1 benchmark
ICB (Image Compression Benchmark)
A carefully chosen set of high-resolution high-precision natural images suited for compression algorithm evaluation.
7 papers · 6 benchmarks
IQUAD (Interactive Question Answering Dataset)
IQUAD is a dataset for Visual Question Answering in interactive environments.
7 papers · 0 benchmarks
Includes 500 categories from the list in the Wikipedia and 399,726 images, a more comprehensive food dataset that surpasses existing popular benchmark datasets by category coverage and data volume.
7 papers · 0 benchmarks
This dataset contains 5955 painting images (from WikiCommons) : a train set of 2978 images and a test set of 2977 images (for classification task).
7 papers · 1 benchmark
ImageNet-9 consists of images with different amounts of background and foreground signal, which you can use to measure the extent to which your models rely on image backgrounds.
7 papers · 1 benchmark
Interiorverse is a high-quality indoor scene dataset with rich details, including complex furniture and decorations and it is rendered with GGX BRDF model, which has stronger material modeling capability than any BRDF models.
7 papers · 0 benchmarks
The odometry benchmark consists of 22 stereo sequences, saved in loss less png format: We provide 11 sequences (00-10) with ground truth trajectories for training and 11 sequences (11-21) without ground truth for evaluation.
7 papers · 1 benchmark
KITTI360-EX is a dataset for outer- and inner FoV expansion.
7 papers · 1 benchmark
The KUMC dataset for polyp detection and classification was collected from the University of Kansas Medical Center.
7 papers · 0 benchmarks
The Kannada-MNIST dataset is a drop-in substitute for the standard MNIST dataset for the Kannada language.
7 papers · 0 benchmarks
Large Scale Composed Image Retrieval (LaSCo) is a new dataset for Composed Image Retrieval (CoIR), x10 times larger than current ones.
7 papers · 1 benchmark
The MMBody dataset provides human body data with motion capture, GT mesh, Kinect RGBD, and millimeter wave sensor data.
7 papers · 0 benchmarks
MegaAge is a large dataset that consists of 41,941 faces annotated with age posterior distributions.
7 papers · 0 benchmarks
MobilityAids is a dataset for perception of people and their mobility aids.
7 papers · 0 benchmarks
NDD20 (Northumberland Dolphin Dataset 2020)
Northumberland Dolphin Dataset 2020 (NDD20) is a challenging image dataset annotated for both coarse and fine-grained instance segmentation and categorisation.
7 papers · 0 benchmarks
NIGHTS (Novel Image Generations with Human-Tested Similarity)
A dataset of human similarity judgments over image pairs that are alike in diverse ways.
7 papers · 0 benchmarks
NumtaDB (Assembled Bengali Handwritten Digits)
To benchmark Bengali digit recognition algorithms, a large publicly available dataset is required which is free from biases originating from geographical location, gender, and age.
7 papers · 0 benchmarks
OMMO is a new benchmark for several outdoor NeRF-based tasks, such as novel view synthesis, surface reconstruction, and multi-modal NeRF.
7 papers · 0 benchmarks
ORBIT is a real-world few-shot dataset and benchmark grounded in a real-world application of teachable object recognizers for people who are blind/low vision.
7 papers · 2 benchmarks
OST300 is an outdoor scene dataset with 300 test images of outdoor scenes, and a training set of 7 categories of images with rich textures.
7 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.