Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 20 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 913–960 of 3,239

ADAM (Adam: automatic detection challenge on age-related macular degeneration)
ADAM is organized as a half day Challenge, a Satellite Event of the ISBI 2020 conference in Iowa City, Iowa, USA.
12 papers · 1 benchmark
Aesthetic Visual Analysis is a dataset for aesthetic image assessment that contains over 250,000 images along with a rich variety of meta-data including a large number of aesthetic scores for each image, semantic labels for over 60…
12 papers · 1 benchmark
BenchLMM (BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models)
Large Multimodal Models (LMMs) such as GPT-4V and LLaVA have shown remarkable capabilities in visual reasoning with common image styles.
12 papers · 1 benchmark
Bongard-HOI testifies to which extent your few-shot visual learner can quickly induce the true HOI concept from a handful of images and perform reasoning with it.
12 papers · 1 benchmark
The COCO-MLT is created from MS COCO-2017, containing 1,909 images from 80 classes.
12 papers · 2 benchmarks
Update on 3DIdent, where we introduce six additional object classes (Hare, Dragon, Cow, Armadillo, Horse, and Head), and impose a causal graph over the latent variables.
12 papers · 1 benchmark
FLAME (Fire Luminosity Airborne-based Machine learning Evaluation)
FLAME is a fire image dataset collected by drones during a prescribed burning piled detritus in an Arizona pine forest.
12 papers · 1 benchmark
HIC (Hands in Action)
The Hands in action dataset (HIC) dataset has RGB-D sequences of hands interacting with objects.
12 papers · 0 benchmarks
HRSC2016 (High resolution ship collections 2016)
High-resolution ship collections 2016 (HRSC2016) is a data set used for scientific research.
12 papers · 1 benchmark
An in-the-wild stereo image dataset, comprising 49,368 image pairs contributed by users of the Holopix mobile social platform.
12 papers · 0 benchmarks
HyperKvasir dataset contains 110,079 images and 374 videos where it captures anatomical landmarks and pathological and normal findings.
12 papers · 2 benchmarks
A new large dataset for illumination estimation.
12 papers · 0 benchmarks
IRS (Indoor Robotics Stereo)
IRS is an open dataset for indoor robotics vision tasks, especially disparity and surface normal estimation.
12 papers · 0 benchmarks
100 tasks from LIBERO-100 suite.
12 papers · 1 benchmark
Math-Vision (Math-V) dataset is a meticulously curated collection of 3,040 high-quality mathematical problems with visual contexts sourced from real math competitions.
12 papers · 1 benchmark
MMNeedle (Multimodal Needle in a Haystack)
We introduce the MultiModal Needle-in-a-haystack (MMNeedle) benchmark, specifically designed to assess the long-context capabilities of MLLMs.
12 papers · 1 benchmark
MoVi (Large Multipurpose Motion and Video Dataset)
Contains 60 female and 30 male actors performing a collection of 20 predefined everyday actions and sports movements, and one self-chosen movement.
12 papers · 1 benchmark
MuST-Cinema is a Multilingual Speech-to-Subtitles corpus ideal for building subtitle-oriented machine and speech translation systems.
12 papers · 0 benchmarks
ONCE-3DLanes (Monocular 3D Lane Detection Dataset)
ONCE-3DLanes is a real-world autonomous driving dataset with lane layout annotation in 3D space.
12 papers · 0 benchmarks
OPA (Object Placement Assessment)
Object-Placement-Assessment (OPA) is a task consisting on verifying whether a composite image is plausible in terms of the object placement.
12 papers · 0 benchmarks
Partial iLIDS is a dataset for occluded person person re-identification.
12 papers · 0 benchmarks
People-Art is an object detection dataset which consists of people in 43 different styles.
12 papers · 2 benchmarks
RECON (RECON Outdoor Navigation Dataset)
https://sites.google.com/view/recon-robot/dataset
12 papers · 0 benchmarks
A dataset of color images corrupted by natural noise due to low-light conditions, together with spatially and intensity-aligned low noise images of the same scenes.
12 papers · 1 benchmark
RICE (Remote sensing Image Cloud rEmoving)
RICE is a remote sensing image dataset for cloud removal.
12 papers · 1 benchmark
SILK (Synth It Like KITTI)
An important factor in advancing autonomous driving systems is simulation.
12 papers · 0 benchmarks
SODA10M is a large-scale object detection benchmark for standardizing the evaluation of different self-supervised and semi-supervised approaches by learning from raw data.
12 papers · 0 benchmarks
Semi-iNat (Semi-Supervised iNaturalist)
Semi-iNat is a challenging dataset for semi-supervised classification with a long-tailed distribution of classes, fine-grained categories, and domain shifts between labeled and unlabeled data.
12 papers · 0 benchmarks
SpaceNet 7 (Multi-Temporal Urban Development SpaceNet Dataset)
Satellite imagery analytics have numerous human development and disaster response applications, particularly when time series methods are involved.
12 papers · 0 benchmarks
TJU-DHD is a high-resolution dataset for object detection and pedestrian detection.
12 papers · 2 benchmarks
The TrajNet Challenge represents a large multi-scenario forecasting benchmark.
12 papers · 2 benchmarks
The TrashCan dataset is an instance-segmentation dataset of underwater trash.
12 papers · 0 benchmarks
We construct the long-tailed version of VOC from its 2012 train-val set.
12 papers · 2 benchmarks
VOT2014 (Visual Object Tracking Challenge 2014)
The dataset comprises 25 short sequences showing various objects in challenging backgrounds.
12 papers · 1 benchmark
Consists of over 39,000 images originating from people who are blind that are each paired with five captions.
12 papers · 0 benchmarks
WildScenes is a bi-modal benchmark dataset consisting of multiple large-scale, sequential traversals in natural environments, including semantic annotations in high-resolution 2D images and dense 3D LiDAR point clouds, and accurate 6-DoF…
12 papers · 2 benchmarks
4D-OR includes a total of 6734 scenes, recorded by six calibrated RGB-D Kinect sensors 1 mounted to the ceiling of the OR, with one frame-per-second, providing synchronized RGB and depth images.
11 papers · 3 benchmarks
AeBAD (Aero-engine Blade Anomaly Detection Dataset)
Unlike previous datasets that focus on detecting the diversity of defect categories (like MVTec AD and VisA), AeBAD is centered on the diversity of domains within the same data category.
11 papers · 3 benchmarks
BCNB (Early Breast Cancer Core-Needle Biopsy WSI)
Breast cancer (BC) has become the greatest threat to women’s health worldwide.
11 papers · 0 benchmarks
BCN20000 is a dataset composed of 19,424 dermoscopic images of skin lesions captured from 2010 to 2016 in the facilities of the Hospital Clínic in Barcelona.
11 papers · 0 benchmarks
BreakHis (Breast Cancer Histopathological Database)
The Breast Cancer Histopathological Image Classification (BreakHis) is composed of 9,109 microscopic images of breast tumor tissue collected from 82 patients using different magnifying factors (40X, 100X, 200X, and 400X).
11 papers · 5 benchmarks
COST (COCO Segmentation Text)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
11 papers · 0 benchmarks
Casia V1 is a dataset for forgery classification.
11 papers · 2 benchmarks
This is the home of a collaborative data collection effort by U.
11 papers · 1 benchmark
Five classic grayscale images commonly used for image quality assessment tasks.
11 papers · 4 benchmarks
Crello (Crello dataset)
Crello dataset consists of design templates obtained from online design service, crello.com.
11 papers · 0 benchmarks
Depth in the Wild is a dataset for single-image depth perception in the wild, i.e., recovering depth from a single image taken in unconstrained settings.
11 papers · 0 benchmarks
DocILE is a large dataset of business documents for the tasks of Key Information Localization and Extraction and Line Item Recognition.
11 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.