Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 60 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2833–2880 of 3,239
The SuSy Dataset combines authentic photographs and AI-generated images designed for training and evaluating synthetic image detection models.
1 paper · 0 benchmarks
Super-CLEVR-3D is a visual question answering (VQA) dataset where the questions are about the explicit 3D configuration of the objects from images (i.e.
1 paper · 0 benchmarks
A dataset of images containing leaves from 15 tree classes.
1 paper · 0 benchmarks
SyDog (A Synthetic Dog Dataset)
SyDog is a synthetic dataset of dogs containing ground truth pose and bounding box coordinates which was generated using the game engine, Unity.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
This is a pose estimation dataset, consisting of symmetric 3D shapes where multiple orientations are visually indistinguishable.
1 paper · 0 benchmarks
50K synthetic renders of the human foot, with surface normals, masks and keypoints.
1 paper · 0 benchmarks
3D Computer Graphics is leveraged to generate a large and diverse dataset for training bike rotation estimators in bike parking assessment.
1 paper · 0 benchmarks
A dataset consisting of high-quality, synthetic chest X-rays from the CheXGenBench-benchmark leading model, Sana (0.6B).
1 paper · 0 benchmarks
A public open dataset of synthetic chest X-ray images of COVID-19.
1 paper · 0 benchmarks
This dataset is a large-scale synthetic dataset to simulate the attack scenario for a keystroke inference attack.
1 paper · 0 benchmarks
Synthetic dataset comprising three different environments for multi-camera dynamic novel view synthesis for soccer.
1 paper · 0 benchmarks
SyntheticFur is a dataset for neural rendering.
1 paper · 0 benchmarks
T2 Guiding is a dataset of 1000 images, each with six image labels.
1 paper · 0 benchmarks
TADAC (Text Annotated Distortion, Appearance and Content Dataset)
We have developed a systematic method for constructing large text annotated image databases designed for exploiting vision-language modeling for image quality assessment and present the Text Annotated Distortion, Appearance and Content…
1 paper · 0 benchmarks
TAMPAR is a real-world dataset of parcel photos for tampering detection with annotations in COCO format.
1 paper · 0 benchmarks
Our dataset augments the TAO dataset with amodal bounding box annotations for fully invisible, out-of-frame, and occluded objects.
1 paper · 0 benchmarks
TCB-DS (Toxigenic Cyanobacteria Dataset)
The TCB-DS dataset is a specialized collection of microscopic images focusing on the automatic recognition of cyanobacteria genera.
1 paper · 0 benchmarks
The TED VCR Video Retrieval Dataset is a multimodal collection derived from publicly available TED Talks.
1 paper · 0 benchmarks
A collection of photographic and synthetic images intended for analysis of image processing techniques and quality assessment of displays.
1 paper · 0 benchmarks
THRawS (Thermal Hotspots in Raw Sentinel-2 data)
THRawS is a new dataset of raw Sentinel-2 (S-2) satellite data containing warm temperature hotspots such as wildfires and volcanic eruptions from around the world.
1 paper · 0 benchmarks
TLFM dataset (TLFM dataset for microscopy image sequence generation)
TLFM dataset structured in sequences of at least nine timesteps.
1 paper · 0 benchmarks
TMBuD is a dataset for building recognition and 3D reconstruction of human made structures in urban scenarios.
1 paper · 0 benchmarks
TRR360D is based on the ICDAR2019MTD modern table detection dataset, it refers to the annotation format of the DOTA dataset.
1 paper · 1 benchmark
TUMTraffic-VideoQA is a novel dataset designed to understand spatiotemporal video in complex roadside traffic scenarios.
1 paper · 0 benchmarks
TXL-PBC dataset (a freely accessible labeled peripheral blood cell dataset)
The TXL-PBC Dataset is a comprehensive collection of re-annotated and integrated cell images from multiple cell datasets.
1 paper · 0 benchmarks
TYC Dataset (The TYC Dataset for Understanding Instance-Level Semantics and Motions of Cells in Microstructures)
We introduce the trapped yeast cell (TYC) dataset, a novel dataset for understanding instance-level semantics and motions of cells in microstructures.
1 paper · 0 benchmarks
In ICDAR-17, a Page-Object Detection (POD) competition was organized where the task was to identify page objects in documents which includes tables, figures and equations in document.
1 paper · 0 benchmarks
Tecnalia Hyperspectral Dataset contains different non-ferreous fractions of Waste from Electric and Electronic Equipment (WEEE) of Copper, Brass, Aluminum, Stainless Steel and White Copper.
1 paper · 0 benchmarks
Green family of datasets for emergent communications on relations.
1 paper · 0 benchmarks
We introduce TextAtlas5M, a dataset specifically designed for training and evaluating multimodal generation models on dense-text image generation.
1 paper · 0 benchmarks
The Benchmark is a collection of datasets for Monocular Height Estimation.
1 paper · 0 benchmarks
In this paper, we introduce a victim dataset for the RoboCup Rescue competitions.
1 paper · 0 benchmarks
The ULS23 training dataset contains 38,693 diverse lesions from chest-abdomen-pelvis CT examinations.
1 paper · 0 benchmarks
High-resolution thermal infrared face database with extensive manual annotations, introduced by Kopaczka et al, 2018.
1 paper · 0 benchmarks
ThermalWORLD-C is an evaluation set that consists of algorithmically generated corruptions applied to the ThermalWORLD test-set, and especially to both the visible and the thermal data.
1 paper · 0 benchmarks
Dataset of paired thermal and RGB images comprising ten diverse scenes—six indoor and four outdoor scenes— for 3D scene reconstruction and novel view synthesis (e.g.
1 paper · 0 benchmarks
TiROD (Tiny Robotics Object Detection)
Dataset to benchmark Continual Learning for Object Detection in a Tiny Robotics settings.
1 paper · 1 benchmark
The TimberVision dataset consists of more than 2k annotated RGB images and contains a total of 51k trunk components including cut and lateral surfaces, thereby surpassing any existing dataset in this domain in terms of both quantity and…
1 paper · 0 benchmarks
Tiny ImageNet-A is a subset of the Tiny ImageNet test set consisting of 3,374 images comprising real-world, unmodified, and naturally occurring examples that are misclassified by ResNet-18.
1 paper · 0 benchmarks
TinySocial is a dataset to enable research on Social Visual Question Answering.
1 paper · 0 benchmarks
A dataset made of 3D image data and their embeddings to test TomoSAM
1 paper · 0 benchmarks
Toulouse Vanishing Points Dataset is a public photographs database of Manhattan scenes taken with an iPad Air 1.
1 paper · 0 benchmarks
Trailers12k is a movie trailer dataset comprised of 12,000 titles associated to ten genres.
1 paper · 0 benchmarks
This repository contains data for the NeurIPS conference paper titled "Harnessing Machine Learning for Single-Shot Measurement of Free Electron Laser Pulse Power".
1 paper · 0 benchmarks
The dataset contains procedurally generated images of transparent vessels containing liquid and objects .
1 paper · 1 benchmark
Tsinghua Dogs is a fine-grained classification dataset for dogs, over 65% of whose images are collected from people's real life.
1 paper · 0 benchmarks
Turath-150K is a database of images of the Arab world that reflect objects, activities, and scenarios commonly found there.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.