Home › Datasets › modality › Images
Images datasets
archive 2025-07-28
3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 53 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Images datasets 2497–2544 of 3,239
The Iranis Dataset is a Large-scale dataset of Farsi license plate characters containing a large-scale dataset with more than 83,000 images of Farsi numbers and letters collected from real-world license plate images captured by various…
1 paper · 0 benchmarks
JAMBO (A Multi-Annotator Image Dataset for Benthic Habitat Classification)
The JAMBO dataset contains 3290 underwater images of the seabed captured by an ROV in temperate waters in the Jammer Bay area off the North West coast of Jutland, Denmark.
1 paper · 0 benchmarks
Involves data where a robot interacts with 5.1 cm colored blocks to complete an order-fulfillment style block stacking task.
1 paper · 0 benchmarks
KHATT (KFUPM Handwritten Arabic TexT Database)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Extension of the official KITTI'15 dataset.
1 paper · 0 benchmarks
KITTI-6DoF is a dataset that contains annotations for the 6DoF estimation task for 5 object categories on 7,481 frames.
1 paper · 0 benchmarks
KTI Multiview Football II consists of images of professional footballers during a match of the Allsvenskan league.
1 paper · 0 benchmarks
Explicitly created for Human Computer Interaction (HCI).
1 paper · 0 benchmarks
This dataset comprises over 9,000 images captured in the AI2-THOR simulation environment, featuring 69 distinct object classes.
1 paper · 0 benchmarks
The Kite database is a multi-modal dataset for the control of unmanned aerial vehicles (UAVs).
1 paper · 0 benchmarks
The Sentinel-2 satellite carries 12 CMOS detectors for the VNIR bands, with adjacent detectors having overlapping fields of view that result in overlapping regions in level-1 B (L1B) images.
1 paper · 0 benchmarks
LADI v2 (Low Altitude Disaster Imagery v2)
LADI Overview The Low Altitude Disaster Imagery (LADI) dataset was created to address the relative lack of annotated post-disaster aerial imagery in the computer vision community.
1 paper · 0 benchmarks
Cosmic rays in the LCO CR dataset are labeled accurately and consistently across many diverse observations from various instruments.
1 paper · 0 benchmarks
LDD (LDD: A Grape Diseases Dataset Detection and Instance Segmentation)
The Instance Segmentation task, an extension of the well-known Object Detection task, is of great help in many areas, such as precision agriculture: being able to automatically identify plant organs and the possible diseases associated…
1 paper · 2 benchmarks
LDDRS (LWIR DoFP Dataset of Road Scene)
The LWIR DoFP Dataset of Road Scene (LDDRS) is a road detection dataset with 2,113 annotated images.
1 paper · 0 benchmarks
The LEMMA dataset aims to explore the essence of complex human activities in a goal-directed, multi-agent, multi-task setting with ground-truth labels of compositional atomic-actions and their associated tasks.
1 paper · 0 benchmarks
LIB-HSI (RGB and Hyperspectral images of Building Facades)
The LIB-HSI dataset contains hyperspectral reflectance images and their corresponding RGB images of building façades in a light industrial environment.
1 paper · 0 benchmarks
The National Institute of Informatics provides LIFULL HOME'S Dataset to researchers, which was offered by LIFULL Co., Ltd.
1 paper · 0 benchmarks
LIRCAD (Inria Liver vessels subbranch anotomical nomenclature labels - "LIRCAD")
The structure for the dataset is as follows : 3DLiverVasculatureProject/ ├── CT/ │ Contains the CT scans ├── Labels/ │ Contains the dual labels for the vessel tree annotations 0 background, 1 Portal vein, 2 hepatic vein ├──…
1 paper · 0 benchmarks
LISA Gaze is a dataset for driver gaze estimation comprising of 11 long drives, driven by 10 subjects in two different cars.
1 paper · 0 benchmarks
LIV360SV (Liverpool 360 degree Street View)
The dataset contains 26,645, 360 degree, street-level images collected via cycling with a GoPro Fusion camera, recorded Jan 14th -- 18th 2020.
1 paper · 0 benchmarks
LLNeRF Dataset is a real-world dataset as a benchmark for model learning and evaluation.
1 paper · 0 benchmarks
LLaVA-Rad MIMIC-CXR features more accurate section extractions from MIMIC-CXR free-text radiology reports.
1 paper · 0 benchmarks
LPBA40 (LONI Probabilistic Brain Atlas)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
LSDBench (Long-video Sampling Dilemma Benchmark)
A benchmark that focuses on the sampling dilemma in long-video tasks.
1 paper · 0 benchmarks
LSLF (Large-scale Labeled Face)
Consists of a large number of unconstrained multi-view and partially occluded faces.
1 paper · 0 benchmarks
LVVO (Lecture Video Visual Objects)
The Lecture Video Visual Objects (LVVO) dataset is a benchmark designed for object detection in lecture video frames.
1 paper · 0 benchmarks
This dataset consists of more than 16,000 retinal OCT B-scans from 441 cases (Normal: 120, Drusen: 160, CNV: 161) and is acquired at Noor Eye Hospital, Tehran, Iran.
1 paper · 0 benchmarks
This data set contains weekly scans of cauliflower and broccoli covering a ten week growth cycle from transplant to harvest.
1 paper · 0 benchmarks
It is composed of around 770k of color 256x256 RGB images extracted from the European Union Intellectual Property Office (EUIPO) open registry.
1 paper · 1 benchmark
Dataset containing 9372 RGB images of weeds with the number of leaves counted.
1 paper · 0 benchmarks
This dataset includes sharp-blur pairs of Leishmania image, which is a protozoan parasite microscopy image dataset of Leishmania, obtained from the preserved slides stained with Giemsa.
1 paper · 0 benchmarks
We created this robust and custom light field dataset in order to assist light field researchers in using SOTA machine learning algorithms for a variety of light field tasks such as depth estimation, synthetic aperture imaging, and more.
1 paper · 0 benchmarks
The Lincolnbeet dataset is an object detection dataset designed to encourage research in the identification of items in environments with high levels of occlusion, and in the development of better approaches to evaluate object detection…
1 paper · 0 benchmarks
Liver-US (Liver Ultrasound Dataset for Medical Image Classification)
The Liver-US dataset is a comprehensive collection of high-quality ultrasound images of the liver, including both normal and abnormal cases.
1 paper · 1 benchmark
Loucount is a retail object detection and and counting dataset with rich annotations in retail stores, which consists of 50, 394 images with more than 1.9 million object instances in 140 categories
1 paper · 0 benchmarks
M3LS (Multi-Lingual Multi-Modal Summarization Dataset)
Significant developments in techniques such as encoder-decoder models have enabled us to represent information comprising multiple modalities.
1 paper · 0 benchmarks
In this project, we tried to make malaria detection easily possible at a low cost.
1 paper · 1 benchmark
MAI (Multi-scene Aerial Image)
MAI is a dataset for multi-scene recognition in single aerial images.
1 paper · 0 benchmarks
MAKED (MultiModal MultiLingual Summarization and Keyword Extraction Dataset)
Keyword extraction is an integral task for many downstream problems like clustering, recommendation, search and classification.
1 paper · 0 benchmarks
MAOMaps is a dataset for evaluation of Visual SLAM, RGB-D SLAM and Map Merging algorithms.
1 paper · 0 benchmarks
MARIO (Monitoring Age-related Macular Degeneration Progression In Optical Coherence Tomography)
MICCAI Challenge 2024
1 paper · 0 benchmarks
Analogical reasoning is fundamental to human cognition and holds an important place in various fields.
1 paper · 1 benchmark
MARS dataset processed with our re-Detect and Link (DL) module.
1 paper · 0 benchmarks
MAST (Multi-Attributed Structured Text-to-face Dataset)
A new data consolidation called Multi-Attributed and Structured Text-to-face (MAST) dataset.
1 paper · 0 benchmarks
Manually vAlidated Vq2a Examples fRom Image/Caption datasetS (MAVERICS) is a suite of test-only visual question answering datasets.
1 paper · 0 benchmarks
a dataset of reading pointer meter
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.