Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 49 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2305–2352 of 3,239
The Daimler Monocular Pedestrian Detection dataset is a dataset for pedestrian detection in urban environments.
1 paper · 0 benchmarks
A dataset of images obtained from DALL-E 3 for 67 countries and 10 concept classes, similar to DollarStreet images.
1 paper · 0 benchmarks
DanbooRegion is a dataset consists of 5377 in-the-wild illustration downloaded from the Danbooru2018 and region segment map annotation pairs samples are provided as at 1024px 8-bit RGB images, and region segment maps as int-32 index images.
1 paper · 0 benchmarks
Danish Airs and Grounds (DAG) is a large collection of street-level and aerial images targeting such cases.
1 paper · 0 benchmarks
This dataset contains around 218K sentences, with 1.5 million words, from 30 different books designed for Post-OCR text correction.
1 paper · 0 benchmarks
The dataset is generated from the study of computational reproducibility of Jupyter notebooks from biomedical publications.
1 paper · 0 benchmarks
This repository contains the dataset for the study of the computational reproducibility of Jupyter notebooks from biomedical publications.
1 paper · 0 benchmarks
~1M Flickr images from the XX century-aged from the 1910s to 1990s.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
This dataset inclue multi-spectral acquisition of vegetation for the conception of new DeepIndices.
1 paper · 1 benchmark
The Deep Thermal Imaging dataset consists of two main datasets: - DeepTherm I (Indoor materials) - 15 indoor materials were used to create the dataset DeepTherm I which consists of 14,860 processed thermal images (average count of data for…
1 paper · 0 benchmarks
There are 537 RGB jpg images of cracks and corresponding png binary segmentation masks of crack: a training set with 300 images and a testing set with 237 images.
1 paper · 0 benchmarks
DeepGraviLens is a data set of simulated gravitational lenses consisting of images associated with brightness variation time series.
1 paper · 0 benchmarks
DeepLocCross is a localization dataset that contains RGB-D stereo images captured at 1280 x 720 pixels at a rate of 20 Hz.
1 paper · 0 benchmarks
Two versions of the dataset are offered: one is the full dataset used to train the models in DeformPAM, and the other is a mini dataset for easier examination.
1 paper · 0 benchmarks
DensePose-Track is a dataset of videos where selected frames are annotated in the traditional DensePose manner.
1 paper · 0 benchmarks
Provide: 10 pickle files 8 pickle files are used to generate depth maps 2 pickle files are data of fronto parallel texture with their ground truth depth Each pickle file contains the parameter of the camera system (aperture size, optical…
1 paper · 0 benchmarks
A dataset of 100K synthetic images of skin lesions, ground-truth (GT) segmentations of lesions and healthy skin, GT segmentations of seven body parts (head, torso, hips, legs, feet, arms and hands), and GT binary masks of non-skin regions…
1 paper · 0 benchmarks
A curated dataset of 221 question-answer-rationale triples capturing visualization design decisions and the reasoning behind them, derived from real-world student-authored narratives.
1 paper · 0 benchmarks
DialogCC is a large-scale multi-modal dialogue dataset, which covers diverse real-world topics and various images per dialogue.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Disaster is a dataset that contains images collected from various sources for three different disasters: fire, water and land.
1 paper · 0 benchmarks
DivShift North American West Coast DivShift Paper | Extended Version | Code https://doi.org/10.1609/aaai.v39i27.35060 Highlighting biases through partitions of volunteer-collected biodiversity data and ecologically relevant context for…
1 paper · 0 benchmarks
Doc3DShade extends Doc3D with realistic lighting and shading.
1 paper · 0 benchmarks
An adaption of the MVTec Anomaly Detection dataset, presented in the paper "Domain-independent detection of known anomalies".
1 paper · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
The Drag100 dataset is introduced in the paper "GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models"¹.
1 paper · 0 benchmarks
About the Dataset: 4 classes of drinking waste: Aluminium Cans, Glass bottles, PET (plastic) bottles and HDPE (plastic) Milk bottles.
1 paper · 1 benchmark
A synthetic dataset including driving under adverse weather conditions | Autonomous Driving
1 paper · 0 benchmarks
This dataset contains videos where a flying drone (hexacopter) is captured with multiple consumer-grade cameras (smartphones, compact cameras, gopro,...) with highly accurate 3D drone trajectory ground truth recorderd by a precise…
1 paper · 0 benchmarks
For the Drone-vs-Bird Detection Challenge 2021, 77 different video sequences have been made available as training data.
1 paper · 1 benchmark
From DroneDeploy: We’ve collected a dataset of aerial orthomosaics and elevation images.
1 paper · 1 benchmark
人群计数旨在识别物体的数量,在智能交通、城市管理和安全监控中发挥着重要作用。由于比例变化、照明变化、遮挡和较差的成像条件,尤其是在夜间和雾霾条件下,人群计数的任务非常具有挑战性。 在本文中,我们提出了一个基于无人机的 RGB-Thermal 人群计数数据集 (DroneRGBT),该数据集由 3600…
1 paper · 1 benchmark
Seven different types of dry beans were used in this research, taking into account the features such as form, shape, type, and structure by the market situation.
1 paper · 0 benchmarks
We study dynamic appearance models of both relightable (BRDF) and non-relightable (RGB).
1 paper · 0 benchmarks
EAPD (Expert-labeled Aesthetics Perception Database)
An expert benchmark aiming to comprehensively evaluate the aesthetic perception capacities of MLLMs.
1 paper · 0 benchmarks
EBHI-Seg is a dataset containing 5,170 images of six types of tumor differentiation stages and the corresponding ground truth images.
1 paper · 0 benchmarks
EGC-FPHFS (Early Gastric Cancer Data from First People's Hospital of Foshan)
High-resolution early gastric cancer (EGC) detection and analysis: Patient Data:Datasets often include images from patients diagnosed with gastric cancer, specifically distinguishing between early gastric cancer (EGC) and Non -pathogenic…
1 paper · 1 benchmark
EGO-CH-Gaze (Learning to Detect Attended Objects in Cultural Sites with Gaze Signals and Weak Object Supervision)
To study the problem of weakly supervised attended object detection in cultural sites, we collected and labeled a dataset of egocentric images acquired from subjects visiting a cultural site.
1 paper · 0 benchmarks
For Emotion Interpretation task
1 paper · 2 benchmarks
first everyday task dataset featuring COT outputs, diverse task designs, detailed re-plan processes, along with SFT and DPO sub-datasets.
1 paper · 0 benchmarks
ENSeg Dataset Overview This dataset represents an enhanced subset of the ENS dataset.
1 paper · 1 benchmark
EPIC-ROI builds on top of the EPIC-KITCHENS dataset, and consists of 103 diverse images with pixel-level annotations for regions where human hands frequently touch in everyday interaction.
1 paper · 0 benchmarks
EPIC-STATES builds upon the raw data in the EPIC-KITCHENS dataset and consists of 10 object state categories: open, close, in-hand, out-of-hand, whole, cut, raw, cooked, peeled, unpeeled.
1 paper · 0 benchmarks
ESP (Evaluation for Styled Prompt)
ESP dataset (Evaluation for Styled Prompt dataset) is a benchmark for zero-shot domain-conditional caption generation.
1 paper · 0 benchmarks
The ETHZ Shape dataset contains images of five diverse shape-based classes, collected from Flickr and Google Images.
1 paper · 0 benchmarks
The EXPO-HD Dataset is a dataset of Expo whiteboard markers for the purpose of instance segmentation.
1 paper · 0 benchmarks
EarlyNSD (Early Nutrient Stress Detection of Plants)
Early detection of plant nutritional deficiencies, followed by corrective actions, is essential for sustaining crop yield.
1 paper · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.