Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 31 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 1441–1488 of 3,239
AeroRIT is a hyperspectral dataset to facilitate aerial hyperspectral scene understanding.
4 papers · 0 benchmarks
AutoChart is a dataset for chart-to-text generation, a task that consists on generating analytical descriptions of visual plots.
4 papers · 0 benchmarks
AwA Pose is a large scale animal keypoint dataset with ground truth annotations for keypoint detection of quadruped animals from images.
4 papers · 0 benchmarks
BCCD is a small-scale dataset for blood cells detection.
4 papers · 0 benchmarks
As part of an ongoing worldwide effort to comprehend and monitor insect biodiversity, we present the BIOSCAN-5M Insect dataset to the machine learning community.
4 papers · 0 benchmarks
Bentham manuscripts refers to a large set of documents that were written by the renowned English philosopher and reformer Jeremy Bentham (1748-1832).
4 papers · 1 benchmark
This self-driving dataset collected in Brno, Czech Republic contains data from four WUXGA cameras, two 3D LiDARs, inertial measurement unit, infrared camera and especially differential RTK GNSS receiver with centimetre accuracy.
4 papers · 0 benchmarks
Unsupervised Domain Adaptation demonstrates great potential to mitigate domain shifts by transferring models from labeled source domains to unlabeled target domains.
4 papers · 3 benchmarks
This dataset is an OSN-transmitted (OSN = Online Social Network) version of the CASIA dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (OSN = Online Social Network) version of the CASIA dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (OSN = Online Social Network) version of the CASIA dataset.
4 papers · 1 benchmark
The COVID-19 pandemic raises the problem of adapting face recognition systems to the new reality, where people may wear surgical masks to cover their noses and mouths.
4 papers · 1 benchmark
CDS2K is a benchmark for Concealed scene understanding (CSU), which is a hot computer vision topic aiming to perceive objects with camouflaged properties.
4 papers · 0 benchmarks
CEDAR Signature is a database of off-line signatures for signature verification.
4 papers · 1 benchmark
CHOCOLATE (Captions Have Often ChOsen Lies About The Evidence)
CHOCOLATE is a benchmark for detecting and correcting factual inconsistency in generated chart captions.
4 papers · 4 benchmarks
CI-MNIST (Correlated and Imbalanced MNIST)
CI-MNIST (Correlated and Imbalanced MNIST) is a variant of MNIST dataset with introduced different types of correlations between attributes, dataset features, and an artificial eligibility criterion.
4 papers · 0 benchmarks
The quality of AI-generated images has rapidly increased, leading to concerns of authenticity and trustworthiness.
4 papers · 2 benchmarks
CSAW-S is a dataset of mammography images which includes expert annotations of tumors and non-expert annotations of breast anatomy and artifacts in the image.
4 papers · 0 benchmarks
CUB-GHA (CUB Gaze-based Human Attention)
CUB-GHA is a dataset for fine-grained classification with human attention annotations.
4 papers · 0 benchmarks
The COVID-19 pandemic raises the problem of adapting face recognition systems to the new reality, where people may wear surgical masks to cover their noses and mouths.
4 papers · 1 benchmark
The ChineseLP dataset contains 411 vehicle images (mostly of passenger cars) with Chinese license plates (LPs).
4 papers · 1 benchmark
CocoDoom is a collection of pre-recorded data extracted from Doom gaming sessions along with annotations in the MS Coco format.
4 papers · 0 benchmarks
This dataset is an OSN-transmitted (Online Social Network) version of the Columbia dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (Online Social Network) version of the Columbia dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (Online Social Network) version of the Columbia dataset.
4 papers · 1 benchmark
Concadia is a publicly available Wikipedia-based corpus, which consists of 96,918 images with corresponding English-language descriptions, captions, and surrounding context.
4 papers · 0 benchmarks
We present the CrackVision12k dataset, a collection of 12,000 crack images derived from 13 publicly available crack datasets.
4 papers · 1 benchmark
DABS (Domain-Agnostic Benchmark for Self-supervised learning)
DABS is a domain-agnostic benchmark for self-supervised learning to encourage research and progress towards domain-agnostic methods.
4 papers · 1 benchmark
DAMON (Dense Annotation of 3D Human Object contact in Natural Images)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
4 papers · 0 benchmarks
The DCM dataset is composed of 772 annotated images from 27 golden age comic books.
4 papers · 3 benchmarks
DDI-100 (Distorted Document Images)
The DDI-100 dataset is a synthetic dataset for text detection and recognition based on 7000 real unique document pages and consists of more than 100000 augmented images.
4 papers · 0 benchmarks
This dataset is an OSN-transmitted (Online Social Network) version of the DSO dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (Online Social Network) version of the DSO dataset.
4 papers · 1 benchmark
This dataset is an OSN-transmitted (Online Social Network) version of the DSO dataset.
4 papers · 1 benchmark
A large-scale anime image database with 4.2m+ images annotated with 130m+ text tags describing image contents in detail; it can be useful for machine learning purposes such as image recognition and generation.
4 papers · 0 benchmarks
The dataset consists of over 350,000 public domain patent drawings collected from the United States Patent and Trademark Office (USPTO).
4 papers · 1 benchmark
DiaMOS Plant (A Dataset for Diagnosis and Monitoring Plant Disease)
Abstract The classification and recognition of foliar diseases is an increasingly developing field of research, where the concepts of machine and deep learning are used to support agricultural stakeholders.
4 papers · 0 benchmarks
DiagSet is a histopathological dataset for prostate cancer detection.
4 papers · 0 benchmarks
The images in DukeMTMC-attribute dataset comes from Duke University.
4 papers · 1 benchmark
This dataset provides a large number of training and testing example which is sufficient for a deep learning approach to address Dunhuang Grotto Painting restoration.
4 papers · 0 benchmarks
EDEN (Enclosed garDEN) is a multimodal synthetic dataset, a dataset for nature-oriented applications.
4 papers · 0 benchmarks
EDUB-Seg (Egocentric Dataset of the University of Barcelona – Segmentation)
Egocentric Dataset of the University of Barcelona – Segmentation (EDUB-Seg) is a dataset for egocentric event segmentation acquired by the Narrative Clip, which takes a picture every 30 seconds.
4 papers · 0 benchmarks
EMU (Edited Media Understanding)
48k question-answer pairs written in rich natural language.
4 papers · 0 benchmarks
ESAD (SARAS Endoscopic Surgeon Action Detection)
ESAD is a large-scale dataset designed to tackle the problem of surgeon action detection in endoscopic minimally invasive surgery.
4 papers · 0 benchmarks
ETHEC (ETH Entomological Collection (ETHEC) Dataset)
It includes 47,978 butterfly images with a 4-level label-hierarchy.
4 papers · 0 benchmarks
EVJVQA (English-Japanese-Vietnamese Visual Question Answering)
EVJVQA, the first multilingual Visual Question Answering dataset with three languages: English, Vietnamese, and Japanese, is released in this task.
4 papers · 0 benchmarks
The endoscopic SLAM dataset (EndoSLAM) is a dataset for depth estimation approach for endoscopic videos.
4 papers · 0 benchmarks
A SAR version of the EuroSAT dataset.
4 papers · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.