Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 63 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2977–3024 of 3,239
The YFCC100M Fine-Grained Geolocation dataset is a subset of 100 a set of 36,146 YFCC100M images that had Flickr tags that could be identified as corresponding to one of the labels in the iNaturalist 2017 dataset.
1 paper · 0 benchmarks
An instance segmentation dataset of yeast cells in microstructures.
1 paper · 0 benchmarks
Data for the paper entitled Quantifying yeast colony morphologies with feature engineering from time-lapse photography by A.
1 paper · 0 benchmarks
YesBut Dataset (https://yesbut-dataset.github.io) Understanding satire and humor is a challenging task for even current Vision-Language models.
1 paper · 0 benchmarks
YorkTag provides pairs of sharp/blurred images containing fiducial markers and is proposed to train and qualitatively and quantitatively evaluate our model.
1 paper · 0 benchmarks
YouTube-GDD (YouTube-GDD: A challenging gun detection dataset with rich contextual information)
YouTubeGun Detection Dataset is collected from 343 high-definition YouTube videos and contains 5000 well-chosen images, in which 16064 instances of gun and 9046 instances of person are annotated.
1 paper · 0 benchmarks
The first and the one open dataset for Russian finger- spelling, contained 1,593 annotated phrases and over 37 thousand HD+ videos.
1 paper · 1 benchmark
A large-scale traffic sign and traffic light dataset with accurate 3D positioning and temporally consistent 3D bounding boxes of traffic management objects from up to 200 meters away.
1 paper · 0 benchmarks
A large-scale training dataset suffering from the defocus spread effect (DSE) is synthesized by applying an α-matte boundary defocus model to the VOC 2012 dataset.
1 paper · 0 benchmarks
A fully synthetic dataset of drones generated using structured domain randomization.
1 paper · 0 benchmarks
cryoPPP (CryoPPP: A Large Expert-Curated Cryo-EM Image Dataset for Machine Learning Protein Particle Picking)
The CryoPPP dataset consists of 34 ground truth data and metadata for 335 EMPIAR IDs.
1 paper · 0 benchmarks
deepMTJ (Muscle-Tendon Junction Tracking in Ultrasound Images)
deepMTJ: Muscle-Tendon Junction Tracking in Ultrasound Images ------------------------------------------------------------- deepMTJ is a machine learning approach for automatically tracking of muscle-tendon junctions (MTJ) in ultrasound…
1 paper · 1 benchmark
The database is comprised of two datasets, the 4-Class eTRIMS Dataset with 4 annotated object classes and the 8-Class eTRIMS Dataset with 8 annotated object classes.
1 paper · 0 benchmarks
fruit-SALAD is a synthetic image dataset with 10,000 generated images of fruit depictions.
1 paper · 0 benchmarks
The iNaturalist Fine-Grained Geolocation dataset is an extension of the iNaturalist dataset with complementary geolocation information.
1 paper · 0 benchmarks
Easily generate simple continual learning benchmarks.
1 paper · 0 benchmarks
"The Chicago Face Database was developed at the University of Chicago by Debbie S.
1 paper · 0 benchmarks
the dataset is a monkey doo doo dataset
1 paper · 0 benchmarks
If you want to known more about this dataset and new method, please read our paper link.
1 paper · 0 benchmarks
To encourage reproducible research, a labeled MultiRAW dataset containing>7k RAW images acquired using multiple camera sensors is made publicly accessible for RAW-domain processing.
1 paper · 0 benchmarks
mwBTFreddy dataset is a resource developed to support flash flood damage assessment in urban Malawi, specifically focusing on the impacts of Cyclone Freddy in 2023.
1 paper · 0 benchmarks
The pic2kal benchmark for calorie prediction contains 308,000 images from over 70,000 recipes including photographs, ingredients and instructions, matched with nutritional information.
1 paper · 0 benchmarks
rc_49 (rc_49 Grasping Dataset)
Includes several sets of synthetic stereo images labelled with grasp rectangles representing parallel-jaw grasps (Cornell-like format).
1 paper · 0 benchmarks
The results-A dataset is a dataset consisting of 22 infrared images commonly used for testing performance of Infrared Image Super-Resolution models.
1 paper · 1 benchmark
The sRGB2XYZ dataset contains ~1,200 pairs of camera-rendered sRGB and the corresponding scene-referred CIE XYZ images (971 training, 50 validation, and 244 testing images).
1 paper · 0 benchmarks
These datasets, ComCo and SimCo, designed for evaluating multi-object representation in Vision-Language Models (VLMs).
1 paper · 0 benchmarks
The simply-CLEVR dataset aims to provide a benchmark dataset that can be used for transparent quantitative evaluation of explanation methods (aka heatmaps/XAI methods).
1 paper · 0 benchmarks
The synRailObs contains following categories: - Person - Rocks - Vehicles - Moto-cars - Animals Each directory contain images and correspoing yolo-format annotations and masks, which can be leveraged in both object detection and…
1 paper · 0 benchmarks
The synethetic dataset (10000 pairs of images and region, 2.95GB) is shared with the code (hdf5 dataset format).
1 paper · 0 benchmarks
topex-printer is a dataset containing 102 machine parts of a label printing machine.
1 paper · 0 benchmarks
Microscopy is a cornerstone of biomedical research, enabling detailed study of biological structures at multiple scales.
1 paper · 0 benchmarks
This dataset contains the ground truth for urban changes occurred in Mariupol, Ukraine for the time frame 2017-2020.
1 paper · 0 benchmarks
VQA NLE synthetic dataset, made with LLaVA-1.5 using features from GQA dataset.
1 paper · 0 benchmarks
The dataset consists of 53,189 wikiHow articles across various categories of everyday tasks, 155,265 methods, and 772,294 steps with corresponding images.
1 paper · 1 benchmark
AASCE (Accurate Automated Spinal Curvature Estimation)
The purpose of this challenge is to investigate (semi-)automatic spinal curvature estimation algorithms.
0 papers · 0 benchmarks
ABODA (Abandoned Object Dataset)
ABandoned Objects DAtaset (ABODA) is a new public dataset for abandoned object detection.
0 papers · 0 benchmarks
ADFI (Anomaly Detection Datasets for Visual Inspection)
ADFI Dataset is an image dataset for anomaly detection methods with a focus on industrial inspection.
0 papers · 0 benchmarks
ALFI (Annotations for Label-Free Images)
ALFI (Annotations for Label-Free Images) is a dataset of images and annotations for label-free microscopy imaging.
0 papers · 0 benchmarks
AMUSE (Automotive Multi-Sensor Dataset)
The automotive multi-sensor (AMUSE) dataset consists of inertial and other complementary sensor data combined with monocular, omnidirectional, high frame rate visual data taken in real traffic scenes during multiple test drives.
0 papers · 0 benchmarks
The first public dataset dedicated for Latin (French) and Arabic Scene Text Detection in Highway panels.
0 papers · 0 benchmarks
AiTLAS: Benchmark Arena is an open-source benchmark framework for evaluating state-of-the-art deep learning approaches for image classification in Earth Observation (EO).
0 papers · 0 benchmarks
AntM2C (Ant-Group Multi-Scenario Multi-Modal CTR dataset)
We release a large-scale Multi-Scenario Multi-Modal CTR dataset named AntM2C, built from real industrial data from Alipay.
0 papers · 0 benchmarks
The photo fixation of apple fruitlets was done in the LatHort orchard in Dobele, at the development of fruit (BBCH stage 76-78).
0 papers · 0 benchmarks
The photo fixation of apple fruits was done in the LatHort orchard in Dobele, at the maturity of fruit and seed (BBCH stage 81-85).
0 papers · 0 benchmarks
Dataset contains images with apples infected by scab.
0 papers · 0 benchmarks
Dataset contains images with apple leaves infected by scab.
0 papers · 0 benchmarks
Contain Arabic handwritten digits images (60000 training and 10000 testing images).
0 papers · 0 benchmarks
This dataset contains video shots for two different classes: tigers and cars.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.