Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 57 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2689–2736 of 3,239
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
PhotoMatte85 contains 85 protrait images.
1 paper · 0 benchmarks
Photozilla is a large-scale dataset which includes over 990k images belonging to 10 different photographic styles.
1 paper · 0 benchmarks
Physical concept understanding benchmark.
1 paper · 0 benchmarks
Pick-a-Filter is a semi-synthetic dataset constructed from Pick-a-Pic v1 to measure the capability of text-to-image models of adapting to heterogeneous preferences.
1 paper · 0 benchmarks
The Pinterest Complete the Look dataset consists of over 1 million outfits and 4 million objects.
1 paper · 0 benchmarks
Advanced pixel shift technology is employed to perform a full color sampling of the image.
1 paper · 0 benchmarks
Placepedia contains 240K places with 35M images from all over the world.
1 paper · 0 benchmarks
An evaluation dataset for planning with LLM agents
1 paper · 0 benchmarks
We established a large-scale plant disease segmentation dataset named PlantSeg.
1 paper · 0 benchmarks
A dataset containing four sets of playing card images.
1 paper · 0 benchmarks
PolarRR is a new dataset with more than 100 types of glass in which obtained transmission images are perfectly aligned with input mixed images.
1 paper · 0 benchmarks
The dataset includes polarimetric, RGB and depth automotive (on the road) data.
1 paper · 0 benchmarks
The current industrial pipeline includes 315 dynamic industrial scenarios, which can be categorized into three types: QR codes, text, and products.
1 paper · 0 benchmarks
PolyMATH, a challenging benchmark aimed at evaluating the general cognitive reasoning abilities of MLLMs.
1 paper · 0 benchmarks
PolyU-BPCoMa: A Dataset and Benchmark Towards Mobile Colorized Mapping Using a Backpack Multisensorial System
1 paper · 0 benchmarks
This dataset was built with data acquired at the Hospital Clinic of Barcelona, Spain.
1 paper · 0 benchmarks
This dataset contains annual Sentinel-2 MSI composites (wet and dry season) for Kigali for the period 2016-2020.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
In this Pre-Contest Workshop Slidedeck.pdf: Instructional materials delivered for the seven pre-contest workshops
1 paper · 0 benchmarks
Processed Twitter is a dataset that is used for Twitter topic recognition.
1 paper · 0 benchmarks
Product Page is a large-scale and realistic dataset of webpages.
1 paper · 0 benchmarks
PsOCR (Pashto OCR Dataset)
PsOCR is a large-scale synthetic dataset for Optical Character Recognition in low-resource Pashto language.
1 paper · 0 benchmarks
QDSD (Quantum Dots Stability Diagrams)
This Quantum Dots Stability Diagrams (QDSD) Dataset aggregates experimental stability diagrams of quantum dots from different research groups.
1 paper · 0 benchmarks
Synthetic datasets have successfully been used to probe visual question-answering datasets for their reasoning abilities.
1 paper · 1 benchmark
The R1-Onevision dataset is a meticulously crafted resource designed to empower models with advanced multimodal reasoning capabilities.
1 paper · 0 benchmarks
RAOS (Rethinking Abdominal Organ Segmentation)
Rethinking Abdominal Organ Segmentation (RAOS) in the clinical scenario: A robustness evaluation benchmark with challenging cases.
1 paper · 0 benchmarks
RASMD (RASMD: RGB And SWIR Multispectral Driving Dataset for Robust Perception in Adverse Conditions)
Current autonomous driving algorithms heavily rely on the visible spectrum, which is prone to performance degradation in adverse conditions like fog, rain, snow, glare, and high contrast.
1 paper · 0 benchmarks
RB-Dust (RB-Dust: Real-world Industrial Dust Dehazing Dataset)
A small-scale real-world dataset containing hazy/dusty industrial images and their clean ground truth counterparts.
1 paper · 1 benchmark
We conducted a large crowdsourcing study of click patterns in an interactive segmentation scenario and collected 475K real-user clicks.
1 paper · 0 benchmarks
REAP is a digital benchmark that allows the user to evaluate patch attacks on real images, and under real-world conditions.
1 paper · 0 benchmarks
REBUS (A Robust Evaluation Benchmark of Understanding Symbols)
Recent advances in large language models have led to the development of multimodal LLMs (MLLMs), which take both image data and text as an input.
1 paper · 1 benchmark
Relations in Captions (REC-COCO) is a new dataset that contains associations between caption tokens and bounding boxes in images.
1 paper · 0 benchmarks
RGB Arabic Alphabet Sign Language (AASL) dataset
1 paper · 1 benchmark
The data used in - "Radio Galaxy Zoo EMU: Towards a Semantic Radio Galaxy Morphology Taxonomy" (Bowles et al.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
In this dataset, various objects are arranged on a white table.
1 paper · 0 benchmarks
The RMRC 2014 indoor dataset is a dataset for indoor semantic segmentation.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
RPCD (Reddit Photo Critique Dataset)
The Reddit Photo Critique Dataset (RPCD) contains tuples of image and photo critiques.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
RTB (Robot Tracking Benchmark)
The Robot Tracking Benchmark (RTB) is a synthetic dataset that facilitates the quantitative evaluation of 3D tracking algorithms for multi-body objects.
1 paper · 1 benchmark
RVL-CDIPMP is our first contribution to retrieve the original documents of the IIT-CDIP test collection which were used to create RVL-CDIP.
1 paper · 0 benchmarks
RVL-CDIPMP-N can serve its original goal as a covariate shift test set, now for multi-page document classification.
1 paper · 0 benchmarks
Rogue Wave Dataset-10K dataset consists of 10191 rogue wave images.
1 paper · 0 benchmarks
A total of 227 cross sectional images (20 x 54 mm with a resolution of 289 x 648 pixels) of hind-leg xenograft tumors from 29 mice were obtained with 1mm step-wise movement of the array mounted on a manual positioning device.
1 paper · 0 benchmarks
Automating the creation of catalogues for radio galaxies in next-generation deep surveys necessitates the identification of components within extended sources and their respective infrared hosts.
1 paper · 1 benchmark
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.