Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 66 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 3121–3168 of 3,239

KID-F (K-pop Idol Dataset - Female)
Description K-pop Idol Dataset - Female (KID-F) is the first dataset of K-pop idol high quality face images.
0 papers · 0 benchmarks
The KIT Whole-Body Human Motion Database is a large-scale dataset of whole-body human motion with methods and tools, which allows a unifying representation of captured human motion, and efficient search in the database, as well as the…
0 papers · 0 benchmarks
The KITTI-Motion dataset contains pixel-wise semantic class labels and moving object annotations for 255 images taken from the KITTI Raw dataset.
0 papers · 0 benchmarks
KTI Multiview Football I is a dataset of football players with annotated joints that can be used for multi-view reconstruction.
0 papers · 0 benchmarks
KinFaceW consists of two kinship datasets: KinFaceW-I and KinFaceW-II.
0 papers · 0 benchmarks
The Dataset consists of the multimodal facial images of 52 people (14 females, 38 males) obtained by Kinect.
0 papers · 0 benchmarks
LASIESTA (Labeled and Annotated Sequences for Integral Evaluation of SegmenTation Algorithms) is a segmentation and detection dataset composed by many real indoor and outdoor sequences organized into categories, each of one covering a…
0 papers · 0 benchmarks
LSFM (Large Scale Facial Model (LSFM))
The Large Scale Facial Model (LSFM) is a 3D statistical model of facial shape built from nearly 10,000 individuals.
0 papers · 0 benchmarks
LTIR (Linköping Thermal InfraRed)
The LTIR dataset is a thermal infrared dataset for evaluation of Short-Term Single-Object (STSO) tracking.
0 papers · 0 benchmarks
Lemon dataset has been prepared to investigate the possibilities to tackle the issue of fruit quality control.
0 papers · 0 benchmarks
LfED-6D (Learning from Experience and Demonstration for 6-DOF Grasping Dataset)
The LfED-6D dataset is a collection of 6D grasp annotations acquired through experience (with a robot platform) or by human demonstration.
0 papers · 0 benchmarks
LoLI-Street (Low-Light Images of Streets)
We introduce low-light image enhancement benchmark dataset “Low-light Images of Streets (LoLI-Street),” which contains three subsets: train, validation, and test.
0 papers · 0 benchmarks
Low Light Dataset (Dataset with ill-lighting conditions DILCOD)
Introduced by Khan.
0 papers · 0 benchmarks
The Lusitano dataset was collected over a 3-month period, spanning from January to March, from Paulo de Oliveira, S.A., a prominent textile company, based in Covilhã, Portugal, renowned for its innovative contributions to the textile…
0 papers · 0 benchmarks
A dataset for multi-context visual grounding.
0 papers · 0 benchmarks
MHRI dataset (Multimodal Human-Robot Interaction dataset)
The dataset includes recordings from 10 different users teaching the robot different common kitchen objects, that consists of synchronized recordings from three cameras and a microphone mounted on the robot: An RGB-d camera covers the user…
0 papers · 0 benchmarks
MIMIC Meme Dataset (Misogyny Identification in Multimodal Internet Content in Hindi-English Code-Mix Language)
This dataset endeavors to fill the research void by presenting a meticulously curated collection of misogynistic memes in a code-mixed language of Hindi and English.
0 papers · 0 benchmarks
IQ testing has served as a foundational methodology for evaluating human cognitive capabilities, deliberately decoupling assessment from linguistic background, language proficiency, or domain-specific knowledge to isolate core competencies…
0 papers · 0 benchmarks
MS-EVS Dataset (Multispectral Event-based Face detection dataset)
The MS-EVS Dataset is the first large-scale event-based dataset for face detection.
0 papers · 0 benchmarks
MVP-24K (Multi-grained Vehicle Parsing dataset)
Multi-grained Vehicle Parsing (MVP) is a large-scale dataset for semantic analysis of vehicles in the wild, which has several featured properties.
0 papers · 0 benchmarks
Multi-domain Image Editing Benchmark
0 papers · 0 benchmarks
MapAI: Precision in Building Segmentation Dataset The dataset comprises 7500 training images and 1500 validation images from Denmark.
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 7000+ original Masks images captured and crowdsourced from over 1200+ urban and rural areas, where each image is manually reviewed and verified by computer vision professionals at DC…
0 papers · 0 benchmarks
Media-Text (MediaText: a media industry-based dataset for scene text detetcion)
Media-Text dataset comprising images of banners, posters, covers and another images characterised for media industry.
0 papers · 0 benchmarks
MoCap (CMU Graphics Lab Motion Capture Database)
Collection of various motion capture recordings (walking, dancing, sports, and others) performed by over 140 subjects.
0 papers · 0 benchmarks
This dataset is collected by DataCluster Labs, India.
0 papers · 0 benchmarks
The Mouse Embryo Tracking Database is a dataset for tracking mouse embryos.
0 papers · 0 benchmarks
MuS2 (A Real-World Benchmark for Sentinel-2 Multi-Image Super-Resolution)
A real-world dataset for multi-image super-resolution that matches low-resolution Sentinel-2 images with high-resolution WorldView-2 images.
0 papers · 0 benchmarks
Mudestreda (Mudestreda Multimodal Device State Recognition Dataset)
Mudestreda Multimodal Device State Recognition Dataset obtained from real industrial milling device with Time Series and Image Data for Classification, Regression, Anomaly Detection, Remaining Useful Life (RUL) estimation, Signal Drift…
0 papers · 0 benchmarks
Multi-Spectral Leaf Segmentation (Multi-Spectral Leaf Segmentation For Crop/Weed Identification)
This dataset were acquired with the Airphen (Hyphen, Avignon, France) six-band multi-spectral camera configured using the 450/570/675/710/730/850 nm bands with a 10 nm FWHM.
0 papers · 0 benchmarks
Abstract: We introduce the multi-spectral stereo (MS2) outdoor dataset, including stereo RGB, stereo NIR, stereo thermal, stereo LiDAR data, and GPS/IMU information.
0 papers · 0 benchmarks
we propose the augmented KITTI dataset with fog for both camera and LiDAR sensors with different visibility ranges from 20 to 80 meters to best match realistic fog environment.
0 papers · 0 benchmarks
Multimodal Humor Dataset (Multimodal Humor Dataset: Predicting Laughter Tracks for Sitcoms)
A great number of situational comedies (sitcoms) are being regularly made and the task of adding laughter tracks to these is a critical task.
0 papers · 0 benchmarks
Multimodal Large Language Models (MLLMs) have shown significant promise in various applications, leading to broad interest from researchers and practitioners alike.
0 papers · 0 benchmarks
NEMO (NEMO: A Database for Emotion Analysis Using Functional Near-Infrared Spectroscopy)
We present a dataset for the analysis of human affective states using functional near-infrared spectroscopy (fNIRS).
0 papers · 0 benchmarks
NHR-Edit (NoHumansRequired Edit Dataset)
NHR-Edit is a training dataset for instruction-based image editing.
0 papers · 0 benchmarks
NTLNP (wildlife image dataset)
This is an image dataset for object detection of wildlife in the mixed coniferous broad-leaved forest.
0 papers · 0 benchmarks
A vehicle detection database for vision tasks set in the real world.
0 papers · 0 benchmarks
NeuB1 is a microscopic neuronal image dataset for retinal vessel segmentation, which contains 112 images of size 512 x 152.
0 papers · 0 benchmarks
The thickness and appearance of retinal layers are essential markers for diagnosing and studying eye diseases.
0 papers · 0 benchmarks
OCTCBVS is a benchmark dataset for testing and evaluating novel and state-of-the-art computer vision algorithms.
0 papers · 0 benchmarks
One of the most important aspects of robot scene understanding is semantic segmentation of external environments.
0 papers · 0 benchmarks
OpenSurfaces is a large database of annotated surfaces created from real-world consumer photographs.
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 2000+ original Oximeter images captured and crowdsourced from over 300+ urban and rural areas, where each image is manually reviewed and verified by computer vision professionals at…
0 papers · 0 benchmarks
Data in this study come from western Ecuador's Choco tropical forest, including \textit{Fundación para la Conservación de los Andes Tropicales Reserve and adjacent Reserva Ecológica Mache-Chindul park} (FCAT; 00°23'28'' N, 79°41'05'' W),…
0 papers · 0 benchmarks
PAVIS RGB-D is a dataset for person re-identification using depth information.
0 papers · 0 benchmarks
PCN (Pedestrian Color Naming)
Pedestrian Color Naming (PCN) is a dataset for pedestrian color naming, which contains 14,213 images, each of which hand-labeled with color label for each pixel.
0 papers · 0 benchmarks
The PEARL dataset comprises with 30K pedestrian images, each annotated with 25 attribute categories, spanning over 146 sub-attributes.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.