Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 48 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2257–2304 of 3,239
CelebAGaze consists of 25283 high-resolution celebrity images that are collected from CelebA and the Internet.
1 paper · 0 benchmarks
The dataset has 93 image stacks and their corresponding Extended Depth of Field (EDF) image acquired from cases with grades Nagative, LSIL or HSIL (The Bethesda System): - Negative: 16 - LSIL: 46 - HSIL: 31 The ground truth includes the…
1 paper · 0 benchmarks
Scene change detection (SCD) dataset tailored for generalizable SCD algorithm.
1 paper · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Charlotte-ThermalFace is a thermal face dataset.
1 paper · 0 benchmarks
A novel remote sensing dataset for evaluating a geospatial machine learning model's ability to learn long range dependencies and spatial context understanding.
1 paper · 1 benchmark
The Chess Recognition Dataset (ChessReD) comprises a diverse collection of images of chess formations captured using smartphone cameras; a sensor choice made to ensure real-world applicability.
1 paper · 0 benchmarks
The Chess Recognition Dataset 2K (ChessReD2K) comprises a diverse collection of images of chess formations captured using smartphone cameras; a sensor choice made to ensure real-world applicability.
1 paper · 0 benchmarks
ChiQA is a dataset designed for visual question answering tasks that not only measures the relatedness but also measures the answerability, which demands more fine-grained vision and language reasoning.
1 paper · 0 benchmarks
The Chinese Traditional Painting dataset for style transfer contains 1000 content images and 100 style images.
1 paper · 0 benchmarks
CiNAT Birds 2021 (Cross-View iNaturalist-2021 Birds) dataset contains ground-level images of bird species along with satellite images associated with the geolocation of the ground-level images.
1 paper · 0 benchmarks
The Cifar10Mnist dataset is created using CIFAR-10 and MNIST data sources.
1 paper · 0 benchmarks
Ciona17 is a semantic segmentation dataset with pixel-level annotations pertaining to invasive species in a marine environment.
1 paper · 0 benchmarks
Clarkson Fingerprint Generator consists of a dataset of 50K synthetically generated fingerprints.
1 paper · 0 benchmarks
Clickable heat-map visualizations of the experiments run to quantify the Classic ECN AQM problem and to evaluate the success of the Classic AQM Detection and Fall-back algorithm.
1 paper · 0 benchmarks
Clickbait PDFs (From Attachments to SEO: Click Here to Learn More about Clickbait PDFs!)
The paper presents a study of Clickbait PDFs, which are PDF documents leading to various attacks on the Web.
1 paper · 0 benchmarks
CoCaHis (Colon Cancer Histology Dataset)
Highlights • Publicly available dataset with 82 H&E stained images of frozen sections.
1 paper · 0 benchmarks
This dataset provides simulated flood inundation maps of Abu Dhabi's coast under 174 different shoreline protection scenarios.
1 paper · 1 benchmark
The combinatorial 3D shape dataset is composed of 406 instances of 14 classes.
1 paper · 0 benchmarks
Computer Vision Arxiv Figures dataset consists of 88,645 images that more closely resemble the structure of our visual prompts.
1 paper · 0 benchmarks
ConQA (Conceptual Query Answering)
ConQA is a dataset created using the intersection between VisualGenome and MS-COCO.
1 paper · 2 benchmarks
This is a video and image segmentation dataset for human head and shoulders, relevant for creating elegant media for videoconferencing and virtual reality applications.
1 paper · 0 benchmarks
ConsInv is a stereo RGB + IMU dataset designed for Dynamic SLAM testing and contains two subsets: - ConsInv-Indoors contains sequences in an office setting where small objects are moved.
1 paper · 0 benchmarks
This dataset contains 12,500 meter images acquired in the field by the employees of the Energy Company of Paraná (Copel), which directly serves more than 4 million consuming units, across 395 cities and 1,113 locations (i.e., districts,…
1 paper · 1 benchmark
Probing cross-modal capabilities of Vision & Language models with a counting task.
1 paper · 0 benchmarks
The Creative Visual Storytelling Anthology is a collection of 100 author responses to an improved creative visual storytelling exercise over a sequence of three images.
1 paper · 0 benchmarks
CropCOCO is a validation-only dataset of COCO val 2017 images cropped such that some keypoints annotations are outside of the image.
1 paper · 0 benchmarks
Cross Modal Automatic Commenting (CMAC) is a task which aims to automatically generate comments for graphic news.
1 paper · 0 benchmarks
A large synthetic multi-camera crowd counting dataset with a large number of scenes and camera views to capture many possible variations, which avoids the difficulty of collecting and annotating such a large real dataset.
1 paper · 0 benchmarks
The standard evaluation protocol of Cross-View Time dataset allows for certain cameras to be shared between training and testing sets.
1 paper · 1 benchmark
To study the data-scarcity mitigation for learning-based visual localization methods via sim-to-real transfer, we curate and now present the CrossLoc benchmark datasets—a multimodal aerial sim-to-real data available for flights above…
1 paper · 0 benchmarks
This dataset concentrates on the activities of the crowd for a fine-grained image classification task, named as Crowd Activity dataset, as automatically understanding crowd activity is meaningful for social security.
1 paper · 0 benchmarks
Abstract: Through digitization, maintaining and promoting cultural heritage is being strengthened.
1 paper · 0 benchmarks
The Curated AFD dataset is a curated version of the Asian Face Dataset (AFD) for face recognition research.
1 paper · 0 benchmarks
D3DFACS (Dynamic 3D Facial Action Coding System Database)
The D3DFACS dataset is a dynamic 3D facial expression data set based on the Facial Action Coding System.
1 paper · 0 benchmarks
DADE (Driving Agents in Dynamic Environments)
The DADE dataset, short for Driving Agents in Dynamic Environments, is a synthetic dataset designed for the training and evaluation of methods for the task of semantic segmentation in the context of autonomous driving agents navigating…
1 paper · 0 benchmarks
DARai (Daily Activity Recordings for AI and ML applications)
Daily Activity Recordings for Artificial Intelligence (DARai, pronounced "Dahr-ree") is a multimodal, hierarchically annotated dataset constructed to understand human activities in real-world settings.
1 paper · 0 benchmarks
A comprehensive object-instance ReID dataset with multiple indoor object instances under varying lighting conditions.
1 paper · 0 benchmarks
A real world dataset for benchmarking global localization in complex indoor environments.
1 paper · 0 benchmarks
A ProcTHOR created synthetic dataset for benchmarking global localization in complex indoor environments.
1 paper · 0 benchmarks
DAVIS-Edit is a curated testing benchmark for video editing.
1 paper · 0 benchmarks
Object Detection data set created from the engine DeepGTAV, which is based on the video game GTAV.
1 paper · 0 benchmarks
DIGITal (Digitally Generated Numerals)
Digitally Generated Numerals (DIGITal) Description The Digitally Generated Numerals (DIGITal) dataset consists of 100,000 image pairs representing digits from 0 to 9.
1 paper · 0 benchmarks
Contains ~60000 HD images of Deformable Linear Objects (DLOs) generated using blender.
1 paper · 0 benchmarks
Medical VQA dataset built from the IDRiD and eOphta datasets.
1 paper · 0 benchmarks
DRIFT (Domain-Adaptive Regression for Forest Monitoring)
The DRIFT dataset includes 25k image patches collected in five European countries sourced from aerial and nanosatellite image archives.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.