Home › Datasets › task › Image Generation
Image Generation datasets
archive 2025-07-28
79 datasets carry the task tag "Image Generation" (the task itself: Image Generation), ordered by the archive's paper count. Page 2 of 2: 31 shown of 79. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Image Generation datasets 49–79 of 79
We construct a style-balanced dataset, called StyleGallery, covering several open source datasets.
7 papers · 0 benchmarks
ChineseFoodNet aims to automatically recognizing pictured Chinese dishes.
6 papers · 0 benchmarks
YouTube Driving Dataset contains a massive amount of real-world driving frames with various conditions, from different weather, different regions, to diverse scene types
6 papers · 1 benchmark
MatSynth MatSynth is a Physically Based Rendering (PBR) materials dataset designed for modern AI applications.
5 papers · 0 benchmarks
The ObjectsRoom dataset is based on the MuJoCo environment used by the Generative Query Network [4] and is a multi-object extension of the 3d-shapes dataset.
5 papers · 2 benchmarks
RC-49 is a benchmark dataset for generating images conditional on a continuous scalar variable.
5 papers · 1 benchmark
StreetStyle is a large-scale dataset of photos of people annotated with clothing attributes, and use this dataset to train attribute classifiers via deep learning.
5 papers · 0 benchmarks
A Dense-text Image Benchmark to evaluate large generation model's ability on text generation.
5 papers · 1 benchmark
The FIGR-8 database is a dataset containing 17,375 classes of 1,548,256 images representing pictograms, ideograms, icons, emoticons or object or conception depictions.
4 papers · 0 benchmarks
The Kvasir-VQA dataset is an extended dataset derived from the HyperKvasir and Kvasir-Instrument datasets, augmented with question-and-answer annotations.
3 papers · 0 benchmarks
QST contains 1,167 video clips that are cut out from 216 time-lapse 4K videos collected from YouTube, which can be used for a variety of tasks, such as (high-resolution) video generation, (high-resolution) video prediction,…
3 papers · 0 benchmarks
SCapRepo (Google Play Screenshot Caption)
A screenshot-caption dataset containing 135k pairs of screenshots and captions extracted from Google Play.
3 papers · 0 benchmarks
SketchHairSalon is a dataset for hair generation containing thousands of annotated hair sketch-image pairs and corresponding hair mattes.
3 papers · 0 benchmarks
ENTIGEN (Ethical NaTural Language Interventions in Text-to-Image GENeration)
ENTIGEN is a benchmark dataset to evaluate the change in image generations conditional on ethical interventions across three social axes -- gender, skin color, and culture.
2 papers · 0 benchmarks
Recently, Text-to-Image (T2I) generation models have achieved significant advancements.
2 papers · 0 benchmarks
Boombox is a multi-modal dataset for visual reconstruction from acoustic vibrations.
1 paper · 0 benchmarks
ColorSVG-100K contains: - 100K samples - 500 categories Project website
1 paper · 0 benchmarks
Dataset information (e.g., google drive link) is attached in the GitHub repo: https://github.com/YY-GX/Annotated-Hands-Dataset Please find the description of the dataset in our paper: http://arxiv.org/abs/2401.15075
1 paper · 0 benchmarks
HRI (High-resolution Rainy Image)
The HRI Dataset comprises a total of 3,200 image pairs.
1 paper · 0 benchmarks
A dataset for image editing containing >450k samples of: 1.
1 paper · 0 benchmarks
It is composed of around 770k of color 256x256 RGB images extracted from the European Union Intellectual Property Office (EUIPO) open registry.
1 paper · 1 benchmark
Samples from NASA Perseverance and set of GAN generated synthetic images from Neural Mars.
1 paper · 1 benchmark
A synthetic dataset with 206K pressure images with 3D human poses and shapes.
1 paper · 0 benchmarks
This is a dataset of 306,006 galaxies whose coordinates are taken from the Sloan Digital Sky Survey Data Release 7 and a modified catalogue from Brinchmann+2003 and Wilman+2010.
1 paper · 1 benchmark
SPOT-10 (Animal Pattern Benchmark Dataset for Machine Learning Algorithms)
The SPOTS-10 dataset is an extensive collection of grayscale images showcasing diverse patterns found in ten animal species.
1 paper · 1 benchmark
We introduce TextAtlas5M, a dataset specifically designed for training and evaluating multimodal generation models on dense-text image generation.
1 paper · 0 benchmarks
A high-resolution version of VGGFace2 for academic face editing purposes.
1 paper · 0 benchmarks
WiFiCam dataset for through-wall imaging based on WiFi channel state information.
1 paper · 0 benchmarks
The dataset was curated from the 1% data sample file of the Wikipedia-based Image Text (WIT) Dataset.
1 paper · 0 benchmarks
This dataset is the images of corn seeds considering the top and bottom view independently (two images for one corn seed: top and bottom).
0 papers · 0 benchmarks
Lemon dataset has been prepared to investigate the possibilities to tackle the issue of fruit quality control.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.