Home › Datasets › task › Image-to-Image Translation
Image-to-Image Translation datasets
archive 2025-07-28
32 datasets carry the task tag "Image-to-Image Translation" (the task itself: Image-to-Image Translation), ordered by the archive's paper count. Page 1 of 1: 32 shown of 32. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Image-to-Image Translation datasets 1–32 of 32
Cityscapes is a large-scale database which focuses on semantic understanding of urban street scenes.
3,702 papers · 51 benchmarks
KITTI (Karlsruhe Institute of Technology and Toyota Technological Institute) is one of the most popular datasets for use in mobile robotics and autonomous driving.
3,661 papers · 137 benchmarks
The ADE20K semantic segmentation dataset contains more than 20K scene-centric images exhaustively annotated with pixel-level objects and object parts labels.
1,213 papers · 32 benchmarks
The CelebA-HQ dataset is a high-quality version of CelebA that consists of 30,000 images at 1024×1024 resolution.
954 papers · 13 benchmarks
SYNTHIA (SYNTHetic Collection of Imagery and Annotations)
The SYNTHIA dataset is a synthetic dataset that consists of 9400 multi-viewpoint photo-realistic frames rendered from a virtual city and comes with pixel-level semantic annotations for 13 classes.
538 papers · 10 benchmarks
Perceptual Similarity is a dataset of human perceptual similarity judgments.
452 papers · 0 benchmarks
GTA5 (Grand Theft Auto 5)
The GTA5 dataset contains 24966 synthetic images with pixel level semantic annotation.
412 papers · 7 benchmarks
DeepFashion is a dataset containing around 800K diverse fashion images with their rich annotations (46 categories, 1,000 descriptive attributes, bounding boxes and landmark information) ranging from well-posed product images to…
397 papers · 5 benchmarks
The Common Objects in COntext-stuff (COCO-stuff) dataset is a dataset for scene understanding tasks like semantic segmentation, object detection and image captioning.
338 papers · 17 benchmarks
Animal FacesHQ (AFHQ) is a dataset of animal faces consisting of 15,000 high-quality images at 512 × 512 resolution.
327 papers · 6 benchmarks
Foggy Cityscapes is a synthetic foggy dataset which simulates fog on real scenes.
249 papers · 7 benchmarks
CelebAMask-HQ is a large-scale face image dataset that has 30,000 high-resolution face images selected from the CelebA dataset by following CelebA-HQ.
164 papers · 5 benchmarks
LLVIP (A Visible-infrared Paired Dataset for Low-light Vision)
Visible-infrared Paired Dataset for Low-light Vision 30976 images (15488 pairs) 24 dark scenes, 2 daytime scenes Support for image-to-image translation (visible to infrared, or infrared to visible), visible and infrared image fusion,…
116 papers · 6 benchmarks
RaFD (Radboud Faces Database)
The Radboud Faces Database (RaFD) is a set of pictures of 67 models (both adult and children, males and females) displaying 8 emotional expressions.
81 papers · 2 benchmarks
Synscapes is a synthetic dataset for street scene parsing created using photorealistic rendering techniques, and show state-of-the-art results for training and validation as well as new types of analysis.
46 papers · 1 benchmark
Enables detailed human body model reconstruction in clothing from a single monocular RGB video without requiring a pre scanned template or manually clicked points.
36 papers · 0 benchmarks
UT Zappos50K is a large shoe dataset consisting of 50,025 catalog images collected from Zappos.com.
32 papers · 2 benchmarks
IXI (IXI Brain Development Dataset)
IXI Dataset is a collection of 600 MR brain images from normal, healthy subjects.
23 papers · 4 benchmarks
VIDIT (Virtual Image Dataset for Illumination Transfer)
VIDIT is a reference evaluation benchmark and to push forward the development of illumination manipulation methods.
20 papers · 1 benchmark
BCI (Breast Cancer Immunohistochemical Image Generation)
The evaluation of human epidermal growth factor receptor 2 (HER2) expression is essential to formulate a precise treatment for breast cancer.
19 papers · 1 benchmark
An annotated image memorability dataset to date (with 60,000 labeled images from a diverse array of sources).
18 papers · 0 benchmarks
BCNB (Early Breast Cancer Core-Needle Biopsy WSI)
Breast cancer (BC) has become the greatest threat to women’s health worldwide.
11 papers · 0 benchmarks
The selfie dataset contains 46,836 selfie images annotated with 36 different attributes.
10 papers · 1 benchmark
SEN12MS-CR-TS is a multi-modal and multi-temporal data set for cloud removal.
7 papers · 1 benchmark
FFHQ-Aging is a Dataset of human faces designed for benchmarking age transformation algorithms as well as many other possible vision tasks.
6 papers · 0 benchmarks
This dataset provides the VCIP 2020 Grand Challenge on the NIR Image Colorization dataset.
3 papers · 1 benchmark
Mila Simulated Floods Dataset is a 1.5 square km virtual world using the Unity3D game engine including urban, suburban and rural areas.
2 papers · 1 benchmark
OADAT (OADAT: Experimental and Synthetic Clinical Optoacoustic Data for Standardized Image Processing)
An experimental and synthetic (simulated) OA raw signals and reconstructed image domain datasets rendered with different experimental parameters and tomographic acquisition geometries.
2 papers · 0 benchmarks
This dataset contains synthetic images extracted from the CARLA simulator along with rich information extracted from the deferred rendering pipeline of Unreal Engine 4.
1 paper · 0 benchmarks
LISA Gaze is a dataset for driver gaze estimation comprising of 11 long drives, driven by 10 subjects in two different cars.
1 paper · 0 benchmarks
RASMD (RASMD: RGB And SWIR Multispectral Driving Dataset for Robust Perception in Adverse Conditions)
Current autonomous driving algorithms heavily rely on the visible spectrum, which is prone to performance degradation in adverse conditions like fog, rain, snow, glare, and high contrast.
1 paper · 0 benchmarks
UDA-CH (Unsupervised Domain Adaptation on Cultural Heritage)
UDA-CH contains 16 objects that cover a variety of artworks which can be found in a museum like sculptures, paintings and books.
1 paper · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.