Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 58 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2737–2784 of 3,239

RaindropsOnWindshield is a dataset for training and assessing vision algorithms' performance for different tasks of image artifacts detection on either camera lens or windshield.
1 paper · 0 benchmarks
Raw-Microscopy: 940 raw bright-field microscopy images of human blood smear slides for leukocyte classification (microscopy/images/rawscale100) with corresponding labels (microscopy/labels).
1 paper · 0 benchmarks
realfred is an embodied instruction following benchmark.
1 paper · 0 benchmarks
Synthetic humans generated by the RePoGen method.
1 paper · 0 benchmarks
Reactive Diffusion Policy-Dataset (Dataset of Reactive Diffusion Policy)
Two versions of the dataset are offered: one is the full dataset used to train the models in our paper, and the other is a mini dataset for easier examination.
1 paper · 0 benchmarks
A genomics dataset for OOD detection that allows other researchers to benchmark progress on this important problem.
1 paper · 0 benchmarks
A total of 80 real material samples were captured in a dark room.
1 paper · 0 benchmarks
RealHDRTV dataset is the first real-world paired SDRTV-HDRTV dataset, which includes SDRTV-HDRTV pairs with 8K resolutions captured by a smartphone camera with the “SDR” and “HDR10” modes.
1 paper · 0 benchmarks
RegDB-C is an evaluation set that consists of algorithmically generated corruptions applied to the RegDB test-set, and especially to both the visible and the thermal data.
1 paper · 0 benchmarks
This dataset was acquired in a retrospective study from a cohort of pediatric patients admitted with abdominal pain to Children’s Hospital St.
1 paper · 0 benchmarks
Replay is a collection of multi-view, multi-modal videos of humans interacting socially.
1 paper · 0 benchmarks
Risholme-2021 contains >3.5K images of strawberries at various growth stages along with anomalous instances.
1 paper · 0 benchmarks
Risk-Aware Planning is a dataset that contains the overhead images and their semantic segmentation captured by a drone from the CityEnviron environment in AirSim simulator.
1 paper · 0 benchmarks
Robot@Home2 (Robot@Home2, a robotic dataset of home environments)
Robot@Home2, is an enhanced version aimed at improving usability and functionality for developing and testing mobile robotics and computer vision algorithms.
1 paper · 0 benchmarks
RoomSpace: a new benchmark designed to evaluate language models on spatial reasoning tasks demanding spatial relation knowledge and multi-hop reasoning.
1 paper · 0 benchmarks
The RoseBlooming dataset is a stage-specific flower dataset for detection.
1 paper · 0 benchmarks
S-BIAD843 (Individual 3D cell shapes of Drosophila Wing Disc)
Late third instar wing imaginal discs were cultured in Shields and Sang M3 media (Sigma) supplemented with 2% FBS (Sigma), 1% pen/strep (Gibco), 3ng/ml ecdysone (Sigma) and 2ng/ml insulin (Sigma).
1 paper · 0 benchmarks
S-ODv2 (SeaDronesSee-Object Detection v2)
SeaDronesSee-Object Detection v2 (S-ODv2) dataset contains 14,227 RGB images (training: 8,930; validation: 1,547; testing: 3,750).
1 paper · 0 benchmarks
S-SOD (Surveillance Salient Object Detection)
To validate the generalization abilities of SOD models, we create a small-scale dataset by collecting the most challenging images with varying brightness and contrast, background and foreground colors overlap, among many others.
1 paper · 0 benchmarks
SACID (Saliency Aware Compressed Images Dataset)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
SAT-MTB-VSR is a large-scale dataset for satellite video super-resolution made from original videos of Jilin-1, which is a subset of the satellite video multitasking dataset SAT-MTB.
1 paper · 1 benchmark
SBA (Sequentail Brick Assembly Dataset)
The RAD (Randomly Assembled Object Construction) dataset is a synthetic 3D LEGO dataset designed for the task of Sequential Brick Assembly (SBA).
1 paper · 0 benchmarks
SBCoseg (SBCoseg Dataset)
The SBCoseg dataset includes 889 groups of images and each group consists of 18 images with a common object, leading to 16002 images in total.
1 paper · 1 benchmark
SCARED-C (SCARED-Corrupted)
The dataset SCARED-C is introduced in the context of assessing robustness in endoscopic depth prediction models.
1 paper · 1 benchmark
SCI (Self-Contradictory Instructions)
Large multimodal models (LMMs) excel in adhering to human instructions.
1 paper · 0 benchmarks
SCIAN (SCIAN Gold-standard for Morphological Sperm Analysis)
Dataset of sperm head images with expert-classification labels.
1 paper · 0 benchmarks
SD7K (Shadow Document 7K)
SD7K is the only large-scale high-resolution dataset that satisfies all important data features about document shadow currently, which covers a large number of document shadow images.
1 paper · 0 benchmarks
SESIV (SEmantic Salient Instance Video)
SEmantic Salient Instance Video (SESIV) dataset is obtained by augmenting the DAVIS-2017 benchmark dataset by assigning semantic ground-truth for salient instance labels.
1 paper · 0 benchmarks
SF-MASK (Small Face MASK)
SF-MASK is a collection made from 20k low-resolution images exported from diverse and heterogeneous datasets, ranging from 7 x 7 to 64 x 64 pixel resolution.
1 paper · 0 benchmarks
This work contributes a large, complex, and realistic high-quality safety clothing and helmet detection (SFCHD) dataset.
1 paper · 1 benchmark
SICKLE (Satellite Imagery for Cropping annotated with Keyparameter LabEls)
The availability of well-curated datasets has driven the success of Machine Learning (ML) models.
1 paper · 1 benchmark
Smartphone cameras are ubiquitous in daily life, yet their performance can be severely impacted by dirty lenses, leading to degraded image quality.
1 paper · 0 benchmarks
SIDOD is a new, publicly-available image dataset generated by the NVIDIA Deep Learning Data Synthesizer intended for use in object detection, pose estimation, and tracking applications.
1 paper · 0 benchmarks
SIRST-UAVB (Single frame infrared small target dataset - UAV and birds.)
Infrared dim-small target detection has gained increasing importance in both military and civilian applications due to its ability to detect thermal radiation, operate effectively at night, passively sense radiation, and offer strong…
1 paper · 0 benchmarks
We present the SJTU Multispectral Object Detection (SMOD) dataset for detection.
1 paper · 0 benchmarks
SMOT (Single sequence-Multi Objects Training)
The SMOT dataset, Single sequence-Multi Objects Training, is collected to represent a practical scenario of collecting training images of new objects in the real world, i.e.
1 paper · 0 benchmarks
SMR IU X-Ray (Simplified Medical Reports)
This paper introduces CPIR-MR (Chained Prompting for Improved Readability of Medical Reports), a method designed to simplify complex chest X-ray reports for better patient understanding.
1 paper · 0 benchmarks
SOMPT22 (Surveillance Oriented Multi-Pedestrian Tracking Dataset (SOMPT22))
SOMPT22 is a multi-object tracking (MOT) benchmark focused on surveillance-style pedestrian tracking.
1 paper · 0 benchmarks
SOTIF-PCOD (SOTIF-related Use Case Dataset)
SOTIF-PCOD is a dataset generated using the CARLA simulator, specifically designed for Safety of the Intended Functionality (SOTIF) research.
1 paper · 0 benchmarks
SOTVerse is a user-defined task space of single object tracking.
1 paper · 0 benchmarks
Dataset for Land Cover segmentation from sparse labels, using Sentinel-2 as source imagery.
1 paper · 0 benchmarks
Contains three types of 2D-3D reasoning tasks on view consistency, camera pose, and shape generation, with increasing difficulty.
1 paper · 0 benchmarks
The dataset contains both RGB and depth images, and the data from two accelerometers, together with ground truth calorie values from a calorimeter for calorie expenditure estimation in home environments.
1 paper · 0 benchmarks
SPIQA Dataset Card Dataset Details Dataset Name: SPIQA (Scientific Paper Image Question Answering) Paper: SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers Github: SPIQA eval and metrics code repo Dataset Summary:…
1 paper · 0 benchmarks
SPKL (Seasonal Parking Lot Dataset)
The SPKL dataset contains 1203 images of parking lots divided into 11 categories regarding vision conditions (including the 'winter' category absent in other datasets at the time of publishing).
1 paper · 1 benchmark
SPOT-10 (Animal Pattern Benchmark Dataset for Machine Learning Algorithms)
The SPOTS-10 dataset is an extensive collection of grayscale images showcasing diverse patterns found in ten animal species.
1 paper · 1 benchmark
Confocal fluorescence microscopy is one of the most accessible and widely used imaging techniques for the study of biological processes at the cellular and subcellular levels.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.