Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 48 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2257–2304 of 3,239

CelebAGaze consists of 25283 high-resolution celebrity images that are collected from CelebA and the Internet.
1 paper · 0 benchmarks
The dataset has 93 image stacks and their corresponding Extended Depth of Field (EDF) image acquired from cases with grades Nagative, LSIL or HSIL (The Bethesda System): - Negative: 16 - LSIL: 46 - HSIL: 31 The ground truth includes the…
1 paper · 0 benchmarks
Scene change detection (SCD) dataset tailored for generalizable SCD algorithm.
1 paper · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Charlotte-ThermalFace is a thermal face dataset.
1 paper · 0 benchmarks
ChesapeakeRSC (Chesapeake Roads Spatial Context)
A novel remote sensing dataset for evaluating a geospatial machine learning model's ability to learn long range dependencies and spatial context understanding.
1 paper · 1 benchmark
ChessReD (Chess Recognition Dataset)
The Chess Recognition Dataset (ChessReD) comprises a diverse collection of images of chess formations captured using smartphone cameras; a sensor choice made to ensure real-world applicability.
1 paper · 0 benchmarks
ChessReD2K (Chess Recognition Dataset 2K)
The Chess Recognition Dataset 2K (ChessReD2K) comprises a diverse collection of images of chess formations captured using smartphone cameras; a sensor choice made to ensure real-world applicability.
1 paper · 0 benchmarks
ChiQA (Chinese VQA)
ChiQA is a dataset designed for visual question answering tasks that not only measures the relatedness but also measures the answerability, which demands more fine-grained vision and language reasoning.
1 paper · 0 benchmarks
The Chinese Traditional Painting dataset for style transfer contains 1000 content images and 100 style images.
1 paper · 0 benchmarks
CiNAT-Birds-2021 (Cross-View iNaturalist Birds 2021)
CiNAT Birds 2021 (Cross-View iNaturalist-2021 Birds) dataset contains ground-level images of bird species along with satellite images associated with the geolocation of the ground-level images.
1 paper · 0 benchmarks
The Cifar10Mnist dataset is created using CIFAR-10 and MNIST data sources.
1 paper · 0 benchmarks
Ciona17 is a semantic segmentation dataset with pixel-level annotations pertaining to invasive species in a marine environment.
1 paper · 0 benchmarks
Clarkson Fingerprint Generator consists of a dataset of 50K synthetically generated fingerprints.
1 paper · 0 benchmarks
Clickable heat-map visualizations of the experiments run to quantify the Classic ECN AQM problem and to evaluate the success of the Classic AQM Detection and Fall-back algorithm.
1 paper · 0 benchmarks
Clickbait PDFs (From Attachments to SEO: Click Here to Learn More about Clickbait PDFs!)
The paper presents a study of Clickbait PDFs, which are PDF documents leading to various attacks on the Web.
1 paper · 0 benchmarks
CoCaHis (Colon Cancer Histology Dataset)
Highlights • Publicly available dataset with 82 H&E stained images of frozen sections.
1 paper · 0 benchmarks
Coastal Inundation Maps with Floodwater Depth Values (Simulated Flood Inundation Maps of Abu Dhabi's Coast Under Different Shoreline Protection Scenarios)
This dataset provides simulated flood inundation maps of Abu Dhabi's coast under 174 different shoreline protection scenarios.
1 paper · 1 benchmark
The combinatorial 3D shape dataset is composed of 406 instances of 14 classes.
1 paper · 0 benchmarks
Computer Vision Arxiv Figures dataset consists of 88,645 images that more closely resemble the structure of our visual prompts.
1 paper · 0 benchmarks
ConQA (Conceptual Query Answering)
ConQA is a dataset created using the intersection between VisualGenome and MS-COCO.
1 paper · 2 benchmarks
This is a video and image segmentation dataset for human head and shoulders, relevant for creating elegant media for videoconferencing and virtual reality applications.
1 paper · 0 benchmarks
ConsInv is a stereo RGB + IMU dataset designed for Dynamic SLAM testing and contains two subsets: - ConsInv-Indoors contains sequences in an office setting where small objects are moved.
1 paper · 0 benchmarks
This dataset contains 12,500 meter images acquired in the field by the employees of the Energy Company of Paraná (Copel), which directly serves more than 4 million consuming units, across 395 cities and 1,113 locations (i.e., districts,…
1 paper · 1 benchmark
Counting Probe (Counting Probe based on Visual7W)
Probing cross-modal capabilities of Vision & Language models with a counting task.
1 paper · 0 benchmarks
Creative Visual Storytelling Anthology (ARL Creative Visual Storytelling Anthology)
The Creative Visual Storytelling Anthology is a collection of 100 author responses to an improved creative visual storytelling exercise over a sequence of three images.
1 paper · 0 benchmarks
CropCOCO is a validation-only dataset of COCO val 2017 images cropped such that some keypoints annotations are outside of the image.
1 paper · 0 benchmarks
Cross Modal Automatic Commenting (CMAC) is a task which aims to automatically generate comments for graphic news.
1 paper · 0 benchmarks
A large synthetic multi-camera crowd counting dataset with a large number of scenes and camera views to capture many possible variations, which avoids the difficulty of collecting and annotating such a large real dataset.
1 paper · 0 benchmarks
The standard evaluation protocol of Cross-View Time dataset allows for certain cameras to be shared between training and testing sets.
1 paper · 1 benchmark
To study the data-scarcity mitigation for learning-based visual localization methods via sim-to-real transfer, we curate and now present the CrossLoc benchmark datasets—a multimodal aerial sim-to-real data available for flights above…
1 paper · 0 benchmarks
This dataset concentrates on the activities of the crowd for a fine-grained image classification task, named as Crowd Activity dataset, as automatically understanding crowd activity is meaningful for social security.
1 paper · 0 benchmarks
Abstract: Through digitization, maintaining and promoting cultural heritage is being strengthened.
1 paper · 0 benchmarks
The Curated AFD dataset is a curated version of the Asian Face Dataset (AFD) for face recognition research.
1 paper · 0 benchmarks
D3DFACS (Dynamic 3D Facial Action Coding System Database)
The D3DFACS dataset is a dynamic 3D facial expression data set based on the Facial Action Coding System.
1 paper · 0 benchmarks
DADE (Driving Agents in Dynamic Environments)
The DADE dataset, short for Driving Agents in Dynamic Environments, is a synthetic dataset designed for the training and evaluation of methods for the task of semantic segmentation in the context of autonomous driving agents navigating…
1 paper · 0 benchmarks
DARai (Daily Activity Recordings for AI and ML applications)
Daily Activity Recordings for Artificial Intelligence (DARai, pronounced "Dahr-ree") is a multimodal, hierarchically annotated dataset constructed to understand human activities in real-world settings.
1 paper · 0 benchmarks
A comprehensive object-instance ReID dataset with multiple indoor object instances under varying lighting conditions.
1 paper · 0 benchmarks
A real world dataset for benchmarking global localization in complex indoor environments.
1 paper · 0 benchmarks
A ProcTHOR created synthetic dataset for benchmarking global localization in complex indoor environments.
1 paper · 0 benchmarks
DAVIS-Edit is a curated testing benchmark for video editing.
1 paper · 0 benchmarks
DGTA-Cattle (DeepGTAV-Cattle)
Object Detection data set created from the engine DeepGTAV, which is based on the video game GTAV.
1 paper · 0 benchmarks
DIGITal (Digitally Generated Numerals)
Digitally Generated Numerals (DIGITal) Description The Digitally Generated Numerals (DIGITal) dataset consists of 100,000 image pairs representing digits from 0 to 9.
1 paper · 0 benchmarks
DLO Instance Segmentation dataset (DLO Instance Segmentation dataset generated by Blender)
Contains ~60000 HD images of Deformable Linear Objects (DLOs) generated using blender.
1 paper · 0 benchmarks
DME VQA dataset (Diabetic Macular Edema VQA dataset)
Medical VQA dataset built from the IDRiD and eOphta datasets.
1 paper · 0 benchmarks
DRIFT (Domain-Adaptive Regression for Forest Monitoring)
The DRIFT dataset includes 25k image patches collected in five European countries sourced from aerial and nanosatellite image archives.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.