Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 47 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2209–2256 of 3,239

BrazilDAM is a multi sensor and multitemporal dataset that consists of multispectral images of ore tailings dams throughout Brazil.
1 paper · 0 benchmarks
BreastRates4 ([MIMBCD-UI] UTA4: Rates Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
1 paper · 0 benchmarks
Burned Area Delineation from Satellite Imagery (A Dataset for Burned Area Delineation and Severity Estimation from Satellite Imagery)
The dataset contains 73 satellite images of different forests damaged by wildfires across Europe with a resolution of up to 10m per pixel.
1 paper · 1 benchmark
CAD-EdgeTune dataset is acquired using a Husarion ROSbot 2.0 and ROSbot 2.0 Pro with the collection speed set to 5 frames per second from a suburban university environment.
1 paper · 0 benchmarks
CAESAR-Radi (CAESAR-Radi: SAR-Ship-Dataset)
This dataset labeled by SAR experts was created using 102 Chinese Gaofen-3 images and 108 Sentinel-1 images.
1 paper · 0 benchmarks
CAGUI (Chinese Android GUI Benchmark)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
CANDOR Corpus (CANDOR = Conversation: A Naturalistic Dataset of Online Recordings)
The CANDOR corpus is a large, novel, multimodal corpus of 1,656 recorded conversations in spoken English.
1 paper · 0 benchmarks
This dataset contains synthetic images extracted from the CARLA simulator along with rich information extracted from the deferred rendering pipeline of Unreal Engine 4.
1 paper · 0 benchmarks
CAT (Context Adjustment Training)
CAT is a specialized dataset for co-saliency detection - one of the core tasks in the field of computer vision.
1 paper · 0 benchmarks
CAT is a specialized dataset for co-saliency detection.
1 paper · 0 benchmarks
CBTex (Synthetic CardBoard Textures)
Dataset of >200 synthetic cardboard texture images that were rendered with DoubeGum's cardboard shader in Blender.
1 paper · 0 benchmarks
CCIHP (Characterized Crowd Instance-level Human Parsing)
CCIHP dataset is devoted to fine-grained description of people in the wild with localized & characterized semantic attributes.
1 paper · 0 benchmarks
CCSE (Chinese Character Stroke Extraction)
Chinese Character Stroke Extraction (CCSE) is a benchmark containing two large-scale datasets: Kaiti CCSE (CCSE-Kai) and Handwritten CCSE (CCSE-HW).
1 paper · 0 benchmarks
CD-HARD comprises 102 images featuring vehicles with oblique license plates sourced from the Cars dataset.
1 paper · 0 benchmarks
Given the difficulty to handle planetary data we provide downloadable files in PNG format from the missions Chang'E-3 and Chang'E-4.
1 paper · 0 benchmarks
CEMS-W (CEMS Wildires)
The dataset includes annotations for burned area delineation and land cover segmentation, with a focus on European soil.
1 paper · 1 benchmark
CFC-DAOD (Caltech Fish Counting – Domain Adaptive Object Detection)
CFC-DAOD is a domain adaptation extension to the Caltech Fish Counting domain generalization benchmark.
1 paper · 1 benchmark
An image sequence dataset of growing snowflakes in HDF5 format.
1 paper · 0 benchmarks
CISOL (Construction Industry Steel Ordering Lists Dataset)
The Construction Industry Steel Ordering Lists (CISOL) dataset comprises table-centric, real-world documents from the construction industry, annotated to facilitate the testing and training of deep learning models for table detection (TD)…
1 paper · 2 benchmarks
CLCXray (Cutters and Liquid Containers X-ray Dataset)
The CLCXray dataset contains 9,565 X-ray images, in which 4,543 X-ray images (real data) are obtained from the real subway scene and 5,022 X-ray images (simulated data) are scanned from manually designed baggage.
1 paper · 1 benchmark
CLEVR-MRT (CLEVR: Mental Rotation Tests)
CLEVR Mental Rotation Tests (CLEVR-MRT) is a new version of the CLEVR dataset.
1 paper · 0 benchmarks
CLOUD (CLOUD Dataset)
The CLOUD dataset is a set of Optical Coherence Tomography of the Anterior Segment images (AS-OCT) used to the automatic identification and representation of the cornea-contact lens relationship.
1 paper · 0 benchmarks
CLPD (China License Plate Dataset)
The CLPD dataset comprises 1200 images that encompass various regions within mainland China.
1 paper · 0 benchmarks
Contains a dataset of 241 Chinese dishes with 191,811 images.
1 paper · 0 benchmarks
CNFOOD-241 Contains a dataset of 241 Chinese dishes with 191,811 images.
1 paper · 1 benchmark
COCO Earthquake is a dataset similar to Common Objects in Context (COCO) used for cracking segmentation.
1 paper · 0 benchmarks
COCO-OOC goes beyond standard object detection to ask the question: Which objects are out-of-context (OOC)?
1 paper · 1 benchmark
COCO-WAN (Medium noise) (Benchmarking Label Noise in Instance Segmentation: Spatial Noise Matters)
The COCO-WAN benchmark is designed to assess the impact of weakly annotations (combined with auto-annotation tools) noise on instance segmentation models.
1 paper · 1 benchmark
COMFORT (Consistent Multilingual Frame of Reference Test)
COMFORT is an evaluation protocol to systematically assess the spatial reasoning capabilities of VLMs.
1 paper · 0 benchmarks
COQE (Containers Of liQuid contEnt)
Contains more than 5,000 images of 10,000 liquid containers in context labelled with volume, amount of content, bounding box annotation, and corresponding similar 3D CAD models.
1 paper · 0 benchmarks
COVIDx CXR-3 is an open access benchmark dataset that we generated, comprising 30,882 CXR images across 17,026 patient cases.
1 paper · 1 benchmark
we introduce COph100, a novel and challenging dataset known as the Comprehensive Ophthalmology Retinal Image Registration dataset for infants with a wide range of image quality issues constituting the public "RIDIRP" database.
1 paper · 0 benchmarks
CPCXR (COVID-19 Posteroanterior Chest X-Ray fused)
The COVID-19 Posteroanterior Chest X-Ray fused (CPCXR) dataset is generated by the fusion of three publicly available datasets: COVID-19 cxr image, Radiological Society of North America (RSNA), and U.S.
1 paper · 0 benchmarks
CPM-Real is a dataset consisting of 3895 images representing real - makeup styles.
1 paper · 0 benchmarks
In this repository you can find all the elaborate results that were used for the simulated evaluation of an innovative, optimized for real-life use, STC-based, multi-robot Coverage Path Planning (mCPP) algorithm.
1 paper · 0 benchmarks
CRCDX (TCGA-CRC-DX)
Histological images of colorectal cancer, derived from the TCGA database
1 paper · 0 benchmarks
CROSS (Cross-Reference Omnidirectional Stitching IQA)
Cross-Reference Omnidirectional Stitching IQA is a novel omnidirectional image dataset containing stitched images as well as dual-fisheye images captured from standard quarters of 0◦, 90◦ , 180◦ and 270◦.
1 paper · 0 benchmarks
Over 20,000 annotated synthetic images and web-scraped images of bicyclists with bounding box annotations in Pascal VOC format.
1 paper · 1 benchmark
The CUHK Face Alignment Database is dataset with 13,466 face images, among which 5, 590 images are from LFW and the remaining 7, 876 images are downloaded from the web.
1 paper · 0 benchmarks
CUHK-QA is a dataset for natural language-based person search using iterative questioning.
1 paper · 0 benchmarks
CUHK01 (CUHK Person Re-identification)
This dataset contains 971 identities from two disjoint camera views.
1 paper · 0 benchmarks
This is the dataset released along with the publication: CUTS: A Deep Learning and Topological Framework for Multigranular Unsupervised Medical Image Segmentation [[ArXiv]](https://arxiv.org/abs/2209.11359)…
1 paper · 0 benchmarks
In this dataset an uppertorso humanoid robot with 7-DOF arm explored 100 different objects belonging to 20 different categories using 10 behaviors: Look, Crush, Grasp, Hold, Lift, Drop, Poke, Push, Shake and Tap.
1 paper · 0 benchmarks
A collaborative effort between researchers at the Vascular Imaging Lab located at the University of Calgary and the Medical Image Computing Lab located at the University of Campinas (UNICAMP) originated the Calgary Campinas public brain…
1 paper · 0 benchmarks
Calliar is a dataset for Arabic calligraphy.
1 paper · 0 benchmarks
Canon RAW Low Light (Canon Camera Low Light RAW Image Dataset)
The goal of this project is to present two new datasets that seek to expand the capability of the Learning to See in the Dark Low-light enhancement CNN for the Canon 6D DSLR, and explore how the network performs when modified in various…
1 paper · 2 benchmarks
The CapMIT1003 database contains captions and clicks collected for images from the MIT1003 database, for which reference eye scanpath are available.
1 paper · 1 benchmark
Cattle data set, which was introduced in a paper.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.