Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 51 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2401–2448 of 3,239

GF-PA66 3D XCT (Glass fiber-reinforced polyamide 66 (GF-PA66) 3D X-ray Computed Tomography (XCT)))
Stack of 2D gray images of glass fiber-reinforced polyamide 66 (GF-PA66) 3D X-ray Computed Tomography (XCT) specimen.
1 paper · 1 benchmark
GF-PA66 3D XCT (latest) (Glass fiber-reinforced polyamide 66 3D X-ray Computed Tomography)
Stack of 2D gray images of glass fiber-reinforced polyamide 66 (GF-PA66) 3D X-ray Computed Tomography (XCT) specimen.
1 paper · 0 benchmarks
The released GIF Reply dataset contains 1,562,701 real text-GIF conversation turns on Twitter.
1 paper · 1 benchmark
GLAMI-1M (A Multilingual Image-Text Fashion Dataset)
We introduce GLAMI-1M: the largest multilingual image-text classification dataset and benchmark.
1 paper · 1 benchmark
GOD (Generic Object Decoding)
The Generic Object Decoding (GOD) Dataset is a specialized resource developed for fMRI-based decoding.
1 paper · 1 benchmark
GOZ (Generic Object ZSL Dataset)
The Generix Object Zero-shot Learning (GOZ) dataset is a benchmark dataset for zero-shot learning.
1 paper · 0 benchmarks
GQN rooms-ring-camera consist of scenes of a variable number of random objects captured in a square room of size 7x7 units.
1 paper · 0 benchmarks
A new dataset containing over 550K pairs (covering 143 km^2 area) of RGB and aerial LIDAR depth images.
1 paper · 0 benchmarks
GTA-UAV dataset provides a large continuous area dataset (covering 81.3km2) for UAV visual geo-localization, expanding the previously aligned drone-satellite pairs to arbitrary drone-satellite pairs to better align with real-world…
1 paper · 0 benchmarks
GUISS dataset (Meshes, textures, Blend files, stereo datasets, depth maps, depth estimations))
We provide all the expected data inputs to GUISS such as meshes, texture images, and blend files.
1 paper · 0 benchmarks
GameQA (GameQA-140K)
GameQA is a large-scale, diverse, and challenging multimodal reasoning dataset designed to enhance the general reasoning capabilities of Vision Language Models (VLMs).
1 paper · 0 benchmarks
We construct Gaze-CIFAR-10, a gaze-augmented image dataset based on the standard CIFAR-10 benchmark, enhanced with human eye-tracking annotations collected using the HTC VIVE Pro Eye headset.
1 paper · 1 benchmark
GeBiD (Geometric shapes Bimodal Dataset)
We provide a custom synthetic bimodal dataset, called GeBiD, designed specifically for the comparison of the joint- and cross-generative capabilities of Multimodal Variational Autoencoders.
1 paper · 0 benchmarks
GelSight Young's Modulus Dataset ============== by Michael Burgess Dataset of tactile images collected over grasping common objects labelled with the objects' Young's Moduli.
1 paper · 0 benchmarks
GenPlot (GenPlot: 500k pre-generated plots)
This dataset contains the pre-generated dataset referenced in the GenPlot Paper.
1 paper · 0 benchmarks
GeoJEPAD (GeoJEPA Dataset)
GeoJEPAD is a multimodal dataset combining OpenStreetMap (OSM) data (attributes and geometries) with high-resolution aerial imagery from diverse urban areas.
1 paper · 0 benchmarks
GeoMNIST (Geometric Shapes MNIST)
A simple dataset consisting of three geometric shapes (Triangle, Rectangle, Ellipsoid) of similar sizes but different orientations.
1 paper · 0 benchmarks
Google Local review (Google Local Data)
Description This Dataset contains review information on Google map (ratings, text, images, etc.), business metadata (address, geographical info, descriptions, category information, price, open hours, and MISC info), and links (relative…
1 paper · 0 benchmarks
Robotic grasp dataset for multi-object multi-grasp evaluation with RGB-D data.
1 paper · 0 benchmarks
GraspClutter6D is a large-scale real-world dataset for robust object perception and robotic grasping in cluttered environments.
1 paper · 0 benchmarks
A large-scale grasp pose detection dataset with a unified evaluation system.
1 paper · 0 benchmarks
GroundCap is a novel grounded image captioning dataset derived from MovieNet, containing 52,350 movie frames with detailed grounded captions.
1 paper · 0 benchmarks
The Human-to-Human-or-Object Interaction Dataset (H2O) dataset is a dataset for Human-Object Interaction (HOI) detection.
1 paper · 0 benchmarks
HA-ViD (HA-ViD: A Human Assembly Video Dataset)
Understanding comprehensive assembly knowledge from videos is critical for futuristic ultra-intelligent industry.
1 paper · 0 benchmarks
HAC (Hybrid Adverse Conditions)
HAC is a dataset for learning and benchmarking arbitrary Hybrid Adverse Conditions restoration.
1 paper · 0 benchmarks
HARRISON dataset is a benchmark on hashtag recommendation for real world images in social networks.
1 paper · 0 benchmarks
HDRT (HDRT Dataset)
The HDRT dataset is a large-scale dataset designed for infrared-guided high dynamic range (HDR) imaging.
1 paper · 0 benchmarks
HGP (Hands Guns and Phones Dataset)
Hands Guns and Phones (HGP) dataset contains 2199 images (1989 for training an 210 for testing) of people using guns or phones in real-world scenarios (people making phones reviews, shooting drills, or making calls).
1 paper · 0 benchmarks
HICRD (Heron Island Coral Reef Dataset)
HICRD (Heron Island Coral Reef Dataset) is a large-scale real underwater image dataset for underwater image restoration.
1 paper · 0 benchmarks
HM3D-Semantics (Habitat-Matterport 3D Semantics)
Habitat-Matterport 3D Semantics Dataset (HM3D-Semantics v0.1) is the largest-ever dataset of semantically-annotated 3D indoor spaces.
1 paper · 0 benchmarks
HOWS (HOWS-CL-25)
HOWS-CL-25 (Household Objects Within Simulation dataset for Continual Learning) is a synthetic dataset especially designed for object classification on mobile robots operating in a changing environment (like a household), where it is…
1 paper · 2 benchmarks
HR-Crime (Human-Related Crime)
HR-Crime is a subset of the UCF-Crime dataset suitable for human-related anomaly detection tasks.
1 paper · 0 benchmarks
HRI (High-resolution Rainy Image)
The HRI Dataset comprises a total of 3,200 image pairs.
1 paper · 0 benchmarks
The dataset concerns toy tasks that a human should teach to a robot.
1 paper · 0 benchmarks
HRPlanesV2 (HRPlanesv2 - High Resolution Satellite Imagery for Aircraft Detection)
The HRPlanesv2 dataset contains 2120 VHR Google Earth images.
1 paper · 0 benchmarks
HSIRS (High-quality Spectral Image Resonstruction and Segmentation Dataset)
We introduce HSIRS, a large scale dataset of hyper-spectral images along with corresponding manually annotated segmentation maps for material characterization and classification based on spectral signature.
1 paper · 0 benchmarks
HYPERVIEW (Seeing Beyond the Visible)
The dataset comprises 2886 patches in total (2 m GSD), of which 1732 patches for training and 1154 patches for testing.
1 paper · 1 benchmark
HaSPeR (Hand Shadow Puppet Image Repository)
TODO
1 paper · 0 benchmarks
Halpe-FullBody is a full body keypoints dataset where each person has annotated 136 keypoints, including 20 for body, 6 for feet, 42 for hands and 68 for face.
1 paper · 0 benchmarks
Multi-Modal Hate Speech Detection with Graph Context.
1 paper · 0 benchmarks
This dataset is composed of 7,753 pairs of whole slide images and their corresponding diagnostic reports, extracted from the TCGA platform and refined with large language models.
1 paper · 1 benchmark
Honeycombs in Concrete (Honeycombs in Concrete Instance Segmentation)
The directory HiCIS contains two datasets for instance segmentation of honeycombs in concrete in COCO Format.
1 paper · 0 benchmarks
This dataset is used for predicting house prices from both images and textual information.
1 paper · 0 benchmarks
HuSHeM (Human Sperm Head Morphology Dataset)
At the Isfahan Fertility and Infertility Center, semen samples were collected from fifteen patients.
1 paper · 0 benchmarks
HuTu 80 (HuTu 80 cell populations)
The image set contains 180 high-resolution color microscopic images of human duodenum adenocarcinoma HuTu 80 cell populations obtained in an in vitro scratch assay (for the details of the experimental protocol, we refer to (Liang et al.,…
1 paper · 1 benchmark
The HumanoidRobotPose dataset is a dataset for real-time pose estimation of humanoid robots.
1 paper · 0 benchmarks
IAPR TC-12 (IAPR TC-12 Benchmark)
The image collection of the IAPR TC-12 Benchmark consists of 20,000 still natural images taken from locations around the world and comprising an assorted cross-section of still natural images.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.