Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 67 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 3169–3216 of 3,239
Dataset contains annotated photographs of pear fruitlets for object detection tasks using YOLO architecture.
0 papers · 0 benchmarks
POET (Pascal Objects Eye Tracking)
The POET (Pascal Objects Eye Tracking) is a dataset that consists of eye tracking data for the complete trainval set of ten objects classes (cat, dog, bicycle, motorbike, boat, aeroplane, horse, cow, sofa, dining table) from Pascal VOC…
0 papers · 0 benchmarks
Post-Spraying Image Evaluation This dataset is for the paper Deep Learning for Precision Agriculture: Post-Spraying Evaluation and Deposition Estimation (https://arxiv.org/abs/2409.16213).
0 papers · 0 benchmarks
The Panoramic Image Database is a panoramic image dataset.
0 papers · 0 benchmarks
The Parzival dataset consists of 47 pages by three writers.
0 papers · 0 benchmarks
Pathfinder and Pathfinder-X have proven to be instrumental in training and testing Large Language Models with long-range dependencies.
0 papers · 0 benchmarks
Dataset contains annotated photographs of pear orchard for object detection tasks using YOLO architecture.
0 papers · 0 benchmarks
The penguin dataset is a collection of images of penguin colonies in Antarctica coming from the larger penguin watch project, which was setup with the purpose of monitoring their changes in population.
0 papers · 0 benchmarks
Plant Centroids is a dataset for stem emerging points (SEP) detection in RGB and NIR image data.
0 papers · 0 benchmarks
Pose Estimation Lunar Robot (Dataset for camera pose estimation research using computer simulated images from rovers on the lunar surface)
Overview The goal: using simulation data to train neural networks to estimate the pose of a rover's camera with respect to a known target object The mission context: A simulated lunar surface, with lunar landers and lunar rovers.
0 papers · 0 benchmarks
RAF-ML (Real-world Affective Faces Multi Label)
Real-world Affective Faces Multi Label (RAF-ML) is a multi-label facial expression dataset with around 5K great-diverse facial images downloaded from the Internet with blended emotions and variability in subjects' identity, head poses,…
0 papers · 0 benchmarks
Laser powder bed fusion (LBPF) is the additive manufacturing (3D printing) process for metals.
0 papers · 0 benchmarks
RF100-VL is a multi-domain benchmark for object detection.
0 papers · 0 benchmarks
RSPECT (The RSNA Pulmonary Embolism CT)
The RSNA Pulmonary Embolism CT (RSPECT) Dataset is composed of CT pulmonary angiogram images and annotations related to pulmonary embolism.
0 papers · 0 benchmarks
RTI International (RTI) generated 2,611 labeled point locations representing 19 different land cover types, clustered in 5 distinct agroecological zones within Rwanda.
0 papers · 0 benchmarks
The dataset has railway track images of two types: normal and defective The task is to classify a given image into normal or defective.
0 papers · 1 benchmark
This dataset consists of two categories.
0 papers · 0 benchmarks
Citation Request: See the articles for more detailed information on the data.
0 papers · 0 benchmarks
RuFa (Ruqaa-Farsi) dataset contains images of text written in one of two Arabic fonts: Ruqaa and Nastaliq (Farsi).
0 papers · 0 benchmarks
Dataset Card for SENTINEL: Mitigating Object Hallucinations via Sentence-Level Early Intervention For the details of this dataset, please refer to the documentation of the GitHub repo.
0 papers · 0 benchmarks
SESYD "Systems Evaluation SYnthetic Documents" is a database of synthetical documents with groundtruth.
0 papers · 0 benchmarks
Facial landmark detection is a cornerstone in many facial analysis tasks such as face recognition, drowsiness detection, and facial expression recognition.
0 papers · 0 benchmarks
SMDG (Standardized Multi-Channel Dataset for Glaucoma)
Standardized Multi-Channel Dataset for Glaucoma (SMDG-19) is a collection and standardization of 19 public datasets, comprised of full-fundus glaucoma images, associated image metadata like, optic disc segmentation, optic cup segmentation,…
0 papers · 0 benchmarks
Sakha-TB (400+400 CXR images for TB diagnosis)
Sakha-TB is a de-identified image dataset of frontal chest X-rays (CXR), collected through collaboration with several medical institutions in the Republic of Sakha (Yakutia, Russia).
0 papers · 0 benchmarks
Semeion (Semeion Handwritten Digit Data Set)
1593 handwritten digits from around 80 persons were scanned, stretched in a rectangular box 16x16 in a gray scale of 256 values.
0 papers · 0 benchmarks
It is released by the Shanghai Central Meteorological Observatory (SCMO) in 2020, records serval years of historical precipitation events in the Yangtze River delta area.
0 papers · 0 benchmarks
This dataset contains images and annotations for scene text detection and recognition.
0 papers · 0 benchmarks
SimNICT is the first dataset for training universal non-ideal measurement CT (NICT) enhancement models.
0 papers · 0 benchmarks
Simulacra Aesthetic Captions is a dataset of over 238000 synthetic images generated with AI models such as CompVis latent GLIDE and Stable Diffusion from over forty thousand user submitted prompts.
0 papers · 0 benchmarks
A novel 360◦ fisheye panoramas dataset, i.e., the Spherical-Navi image dataset is collected, with a unique labeling strategy enabling automatic generation of an arbitrary number of negative samples (wrong heading direction).
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 3000+ originally Stair images captured and crowdsourced from over 500+ urban and rural areas, where each image is manually reviewed and verified by computer vision professionals at…
0 papers · 0 benchmarks
Sugar Beets 2016 is a robot dataset for plant classification as well as localization and mapping that covers the relevant stages for robotic intervention and weed control.
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 7000+ original Suitcase/Luggage images captured and crowdsourced from over 800+ urban and rural areas, where each image is manually reviewed and verified by computer vision professionals…
0 papers · 0 benchmarks
Malaria, a mosquito-borne infectious disease affecting humans and other animals, is widespread in the tropical and subtropical regions.
0 papers · 0 benchmarks
TCMP-300 (Traditional Chinese Medicinal Plant Dataset)
Traditional Chinese medicinal plants are often used to prevent and treat diseases for the human body.
0 papers · 1 benchmark
Face detection and subsequent localization of facial landmarks are the primary steps in many face applications.
0 papers · 0 benchmarks
TS-TR (Turkish Scene Text Recognition Dataset)
The Turkish Scene Text Recognition (TS-TR) dataset was primarily developed to fill the gap in non-English text recognition resources, specifically addressing the unique challenges presented by the Turkish language, such as special…
0 papers · 0 benchmarks
The Temporal Logic Video (TLV) Dataset addresses the scarcity of state-of-the-art video datasets for long-horizon, temporally extended activity and object detection.
0 papers · 0 benchmarks
The Tornado Network (TorNet) dataset is a large, high-resolution benchmark dataset developed to support machine learning research in tornado detection and prediction.
0 papers · 0 benchmarks
Toronto NeuroFace Dataset: A New Dataset for Facial Motion Analysis in Individuals with Neurological Disorders Toronto NeuroFace Dataset is a public dataset with videos of oro-facial gestures performed by individuals with oro-facial…
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 3000+ original Transparent object images such as glasses and mirrors are captured and crowdsourced from over 500+ urban and rural areas, where each image is manually reviewed and…
0 papers · 0 benchmarks
UAS-based Multispectral othomosaics of vineyards from central Portugal - 2 distinct vineyards - Multispectral and HD orthomosaics
0 papers · 0 benchmarks
The UBIRIS.v2 iris dataset contains 11,102 iris images from 261 subjects with 10 images each subject.
0 papers · 0 benchmarks
UFO Cherry Tree Point Clouds consists of a collection of 82 scanned Upright Fruiting Offshoot (UFO) cherry tree point clouds.
0 papers · 0 benchmarks
USYD CAMPUS is a driving dataset collected by Zhou et al at the University of Sydney (USyd) campus and surroundings.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.