Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 59 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2785–2832 of 3,239
The Synthetic Signature Bankcheck Images (SSBI) Dataset is the first publicly available dataset of bank check images with annotations for detecting handwritten components, including names, amounts, dates, and signatures.
1 paper · 0 benchmarks
This is a dataset to benchmark real-time embedded object detection models for RoboCup SSL (Small Size League).
1 paper · 0 benchmarks
STDW is a diverse large-scale dataset for table detection with more than seven thousand samples containing a wide variety of table structures collected from many diverse sources.
1 paper · 1 benchmark
STN PLAD (STN Power Line Assets Dataset)
STN PLAD is a high-resolution and real-world image dataset of multiple high-voltage power line components.
1 paper · 1 benchmark
A RGB-D dataset converted from SUN-RGBD into COCO-style instance segmentation format.
1 paper · 2 benchmarks
SVLD (Social Vision and Language Dataset)
The social vision and language dataset is a large-scale multimodal dataset designed for research into social contextual learning.
1 paper · 0 benchmarks
SVRT (Synthetic Visual Reasoning Task)
The Synthetic Visual Reasoning Test (SVRT) is a series of 23 classification problems involving images of randomly generated shapes.
1 paper · 0 benchmarks
SYNTHIA-PANO is the panoramic version of SYNTHIA dataset.
1 paper · 0 benchmarks
SYSU-MM01-C is an evaluation set that consists of algorithmically generated corruptions applied to the SYSU-MM01 test-set, and especially to both the visible and the thermal data.
1 paper · 0 benchmarks
SaRNet is a single class dataset consisting of tiles of satellite imagery labeled with potential 'targets'.
1 paper · 0 benchmarks
This dataset contains nine video sequences captured by a webcam for salient closed boundary tracking evaluation.
1 paper · 0 benchmarks
Salient-KITTI is a saliency map prediction dataset based on KITTI.
1 paper · 0 benchmarks
The Satellite dataset forms a practical VFL scenario for location identification based on satellite imagery.
1 paper · 0 benchmarks
The satire dataset is a new multi-modal dataset of satirical and regular news articles.
1 paper · 0 benchmarks
SceneNet-RGBD is a synthetic dataset containing large-scale photorealistic renderings of indoor scene trajectories with pixel-level annotations.
1 paper · 0 benchmarks
This dataset is the image stimulus pool of 50 deepfake and 50 real images, used for the experiment in the study titled "Testing Human Ability To Detect Deepfake Images of Human Faces".
1 paper · 0 benchmarks
Test dataset for Semantic Segmentation.
1 paper · 0 benchmarks
This dataset contains 547 social media images taken in the aftermath of various earthquakes.
1 paper · 0 benchmarks
SemanticSugarBeets, a novel and high-quality dataset containing 953 monocular RGB images and 2920 annotations of sugar beets, enables a wide range of learning tasks including object detection, semantic segmentation, instance segmentation…
1 paper · 0 benchmarks
SemanticUSL is a dataset for domain adaptation for LiDAR point cloud semantic segmentation.
1 paper · 0 benchmarks
Sen4AgriNet (A Sentinel-2 multi-year, multi-country benchmark dataset for crop classification and segmentation with deep learning)
A Sentinel-2 based time series multi country benchmark dataset, tailored for agricultural monitoring applications with Machine and Deep Learning.
1 paper · 0 benchmarks
Sequence Consistency Evaluation (SCE) consists of a benchmark task for sequence consistency evaluation (SCE).
1 paper · 0 benchmarks
This dataset accompanies the linked SerialTrack paper and provides test case data (2D/3D, varying particle density) across a range of synthetic and experimental imaging modalities.
1 paper · 0 benchmarks
The synthetic ShapeNet intrinsic image decomposition dataset of 90,000 images.
1 paper · 0 benchmarks
This repository contains documentation for the dataset that accompanies our ICPE 2025 paper, "Shaved Ice: Optimal Compute Resource Commitments for Dynamic Multi-Cloud Workloads".
1 paper · 0 benchmarks
To construct such a dataset, a straightforward approach was scraping images from the web.
1 paper · 1 benchmark
Short BBC Pose contains five one-hour-long videos with sign language signers each with different sleeve length (in contrast to the BBC pose and Extended BBC Pose, which only contain signers with moderately long sleeves).
1 paper · 0 benchmarks
In this Adjudicator ScoresShort Stories and Written Reflections folder: Four files from four student participants of the contest.
1 paper · 0 benchmarks
The proposed dataset includes 1,309 short text instances from Adobe Spark.
1 paper · 0 benchmarks
The SimBEV dataset is a collection of 320 scenes spread across all 11 CARLA maps and contains data from a variety of sensors, including five camera types (RGB, semantic segmentation, instance segmentation, depth, and optical flow), lidar,…
1 paper · 3 benchmarks
It consists of 32x32 pixel images of shapes with multiple attributes (size, location, rotation, color).
1 paper · 0 benchmarks
SinGAN-Seg-polyps is a synthetic dataset for polyp segmentation consisting of 10,000 synthetic polyps and masks.
1 paper · 0 benchmarks
Dataset of 374 photos of hand-drawn sketches of App Inventor apps used for development of the Sketch2aia model for automatic generation of App Inventor wireframes from hand-drawn sketches.
1 paper · 0 benchmarks
This dataset consists of virtual scenes rendered in MuJoCo with multiple views each presented in multiple modalities: image, and synthetic or natural language descriptions.
1 paper · 0 benchmarks
We introduce a large-scale video dataset Slovo for Russian Sign Language task.
1 paper · 1 benchmark
SolarDK is a dataset for the detection and localization of solar.
1 paper · 0 benchmarks
Songdo Traffic (Songdo Traffic: High Accuracy Georeferenced Vehicle Trajectories from a Large-Scale Study in a Smart City)
The Songdo Traffic dataset delivers precisely georeferenced vehicle trajectories captured through high-altitude bird's-eye view (BeV) drone footage over Songdo International Business District, South Korea.
1 paper · 0 benchmarks
Songdo Vision (Songdo Vision: Vehicle Annotations from High-Altitude BeV Drone Imagery in a Smart City)
The Songdo Vision dataset provides high-resolution (4K, 3840×2160 pixels) RGB images annotated with categorized axis-aligned bounding boxes (BBs) for vehicle detection from a high-altitude bird’s-eye view (BeV) perspective.
1 paper · 1 benchmark
Scene Graph Generation (SGG) converts visual scenes into structured graph representations, providing deeper scene understanding for complex vision tasks.
1 paper · 0 benchmarks
Dataset built from partial reconstructions of real-world indoor scenes using RGB-D sequences from ScanNet, aimed at estimating the unknown position of an object (e.g.
1 paper · 0 benchmarks
SpectroVision is a dataset of 14,400 high resolution texture images and spectral measurements collected from a PR2 mobile manipulator that interacted with 144 household objects from eight material categories.
1 paper · 0 benchmarks
Synthetic soccer players rendered on top of real world stadium images in 4K covering half a pitch each.
1 paper · 1 benchmark
The datasets used in our WACV paper High-Fidelity Document Stain Removal via A Large-Scale Real-World Dataset and A Memory-Augmented Transformer.
1 paper · 0 benchmarks
3D confocal stacks with corresponding 2D Light-field microscope images Confocal: -Single volume dimension: 1287x1287x64.
1 paper · 0 benchmarks
Steredo Waterdrop is a real-world dataset for research on stereo waterdrop removal.
1 paper · 0 benchmarks
Stickers is a dataset consisting of 577 high-quality sticker images with alpha channel.
1 paper · 0 benchmarks
This dataset is a "part I" extension of the "Engineered cardiac microbundle time-lapse microscopy image dataset" and contains 732 experimental time-lapse image sequences of beating hiPSC-based cardiac microbundles using microbundle strain…
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.