Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 52 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2449–2496 of 3,239
The IAW dataset contains 420 Ikea furniture pieces from 14 common categories e.g.
1 paper · 0 benchmarks
- Revision: v1.0.0-full-20210527a - DOI: 10.5281/zenodo.4817662 - Authors: J.
1 paper · 0 benchmarks
The dataset is composed of Hematoxylin and eosin (H&E) stained breast histology microscopy and whole-slide images.
1 paper · 1 benchmark
We introduce the IDRCell100K image dataset, a collection of biological images, purposefully curated from the extensive and varied Image Data Resource platform.
1 paper · 0 benchmarks
ILIAS (ILIAS: Instance-Level Image retrieval At Scale)
ILIAS is a large-scale test dataset for evaluation on Instance-Level Image retrieval At Scale.
1 paper · 0 benchmarks
A collection of test sets for evaluating base and chat LLMs (incl.
1 paper · 0 benchmarks
IMFW (Indian Masked Faces In The Wild)
Indian Masked faces in the wild Database is collected into three sets:(i) Indian Celebrity, (ii) Instagram and (iii) Indian Crowd.
1 paper · 0 benchmarks
IMPACT Patent (A Large-scale Integrated Multimodal Patent Analysis and Creation Dataset for Design Patents)
It is a large-scale multimodal patent dataset with detailed captions for design patent figures.
1 paper · 1 benchmark
The INRIA Dense Light Field Dataset (DLFD) is a dataset for testing depth estimation methods in a light field.
1 paper · 0 benchmarks
The INRIA Sprse Light Field Dataset (SLFD) is a dataset for testing depth estimation methods in a light field.
1 paper · 0 benchmarks
This data set contains over 600GB of multimodal data from a Mars analog mission, including accurate 6DoF outdoor ground truth, indoor-outdoor transitions with continuous cross-domain ground truth, and indoor data with Optitrack…
1 paper · 0 benchmarks
This dataset contains 65 DFIs acquired from patients with POAG at the University of Iowa Hospitals and Clinics.
1 paper · 2 benchmarks
We establish the first large benchmark called IRBFD to facilitate the research in the area of nonuniformity correction and infrared UAV target detection, which consists of 50,000 manually labeled infrared images with various nonuniformity…
1 paper · 0 benchmarks
IRMA (15,363 IRMA images of 193 categories for ImageCLEFmed 2009)
This collection compiles anonymous radiographs, which have been arbitrarly selected from routine at the Department of Diagnostic Radiology, Aachen University of Technology (RWTH), Aachen, Germany.
1 paper · 0 benchmarks
IRV2V (IRregular V2V Dataset)
To facilitate research on asynchrony for collaborative perception, we simulate the first collaborative perception dataset with different temporal asynchronies based on CARLA, named IRregular V2V(IRV2V).
1 paper · 1 benchmark
We introduce a new synthetic test set named IS3 for interactive sound source localization.
1 paper · 0 benchmarks
ISBNet is a dataset of images of recyclables.
1 paper · 1 benchmark
ISEKAI dataset’s images are generated by Midjourney’s text-to-image model using well-crafted instructions.
1 paper · 0 benchmarks
ISOD (Indoor Small Object Dataset)
ISOD contains 2,000 manually labelled RGB-D images from 20 diverse sites, each featuring over 30 types of small objects randomly placed amidst the items already present in the scenes.
1 paper · 0 benchmarks
ISP-AD (The Industrial Screen Printing Anomaly Detection Dataset)
The ISP-AD Dataset is a large-scale anomaly detection dataset, representing a real-world industrial use case.
1 paper · 0 benchmarks
The ITCPR dataset is a comprehensive collection specifically designed for the Zero-Shot Composed Person Retrieval (ZS-CPR) task.
1 paper · 1 benchmark
ITDD (Industrial Textile Defect Detection)
The Industrial Textile Defect Detection (ITDD) dataset includes 1885 industrial textile images categorized into 4 categories: cotton fabric, dyed fabric, hemp fabric, and plaid fabric.
1 paper · 2 benchmarks
The IUSTPersonReID dataset was developed to address limitations in existing person re-identification datasets by including cultural and environmental contexts unique to Islamic countries, especially Iran and Iraq.
1 paper · 1 benchmark
IVM-Mix-1M provide over 1M image-instruction pairs with corresponding instruction-relevant mask labels.
1 paper · 0 benchmarks
Icon645 is a large-scale dataset of icon images that cover a wide range of objects: 645,687 colored icons 377 different icon classes These collected icon classes are frequently mentioned in the IconQA questions.
1 paper · 0 benchmarks
IllusionAnimalstest Dataset Characteristics IllusionAnimalstest is a generated dataset based on a synthetic collection of animal images, including 10 animal classes: cat, dog, pigeon, butterfly, elephant, horse, deer, snake, fish, and…
1 paper · 0 benchmarks
IllusionChartest Dataset Characteristics IllusionChartest is a generated dataset containing 3,300 samples of images that feature sequences of 3 to 5 random characters.
1 paper · 0 benchmarks
IllusionFashionMNISTtest Dataset Characteristics IllusionFashionMNISTtest is a generated dataset derived from the FashionMNIST dataset.
1 paper · 0 benchmarks
IllusionMNISTtest Dataset Characteristics IllusionMNISTtest is a generated dataset derived from the MNIST dataset.
1 paper · 0 benchmarks
Image Caption Quality Dataset is a dataset of crowdsourced ratings for machine-generated image captions.
1 paper · 0 benchmarks
This publicly available dataset contains 1613 RGB-D images of field-grown broccoli plants.
1 paper · 0 benchmarks
This ImageNet version contains only 50 training images per class while the original testing set remains unchanged.
1 paper · 1 benchmark
The training and validation data are subsets of the training split of the Imagenet 2012.
1 paper · 0 benchmarks
We build a new evaluation set by adding spotting words to the images of ImageNet 2012 evaluation sets.
1 paper · 0 benchmarks
This dataset consists of ~350k JPEG images of streetlight columns installed on a public road infrastructure located in the city of Bristol, UK.
1 paper · 0 benchmarks
ImagiFilter focusses on photographic and/or natural images, a very common use-case in computer vision research.
1 paper · 0 benchmarks
As a first step towards building models that can recognise immune cells in WSIs, we introduce Immunocto, a high-resolution (40 x magnification) massive database of 2,310,257 immune cells distributed across 4 immune cell subtypes (CD4…
1 paper · 0 benchmarks
There was no predefined dataset of party symbols to be usedas a benchmark.
1 paper · 0 benchmarks
We present two multi-modal datasets, one for Main Board IPOs, and the other for Small and Medium Enterprises (SME) IPOs.
1 paper · 0 benchmarks
Indiscapes2, a new large-scale diverse dataset of Indic manuscripts with semantic layout annotations.
1 paper · 0 benchmarks
The Deepfake face detection task involves a facial image of unknown authenticity for testing.
1 paper · 0 benchmarks
The dfdindoor dataset contains 110 images for training and 29 images for testing.
1 paper · 0 benchmarks
IndraEye (IndraEye: Infrared Electro-Optical Drone-based Aerial Object Detection Dataset)
Deep neural networks (DNNs) have demonstrated superior performance when trained on well-illuminated environments, given that the images are captured through an Electro-Optical (EO) camera, which offers rich texture content.
1 paper · 0 benchmarks
InfraParis is a novel and versatile dataset supporting multiple tasks across three modalities: RGB, depth, and infrared.
1 paper · 0 benchmarks
InstaCities1M is a dataset of social media images with associated text.
1 paper · 0 benchmarks
A dataset for image editing containing >450k samples of: 1.
1 paper · 0 benchmarks
This newly curated synthetic dataset specifies an additional reference region to guide image harmonization.
1 paper · 0 benchmarks
Iran's Built Heritage Binary Image Classification Dataset contains approximately 10,500 CHB images gathered from four different sources: i) The archives of Iran’s cultural heritage ministry ii) The author’s (M.B) personal archives iii)…
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.