Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 46 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2161–2208 of 3,239

Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
AcousticRooms is a large-scale synthetic room impulse response (RIR) dataset designed for cross-room RIR prediction tasks.
1 paper · 0 benchmarks
We filter and match the landmarks in the Google Landmarks dataset with their OpenStreetMap polygons and filter for those located in the United States, resulting in 602 landmarks.
1 paper · 0 benchmarks
AerialMPT is a dataset for pedestrian tracking in aerial image sequences and presents real-world challenges for MOT algorithms such as low frame rate, small moving objects, and complex backgrounds.
1 paper · 0 benchmarks
AesVQA is a dataset that contains 72168 high-quality images and 324756 pairs of aesthetic questions.
1 paper · 0 benchmarks
We introduce the novel task of multimodal puzzle solving, framed within the context of visual question-answering.
1 paper · 1 benchmark
The Algonauts 2023 Challenge focuses on predicting responses in the human brain as participants perceive complex natural visual scenes.
1 paper · 0 benchmarks
The Ambiguous VQA dataset is a dataset of ambiguous questions about images.
1 paper · 0 benchmarks
The AneuX morphology database includes data from 3 different data sources: AneuX, @neurIST and Aneurisk.
1 paper · 0 benchmarks
The Inpainting dataset consists of synchronized Labeled image and LiDAR scanned point clouds.
1 paper · 1 benchmark
The Apron Dataset focuses on training and evaluating classification and detection models for airport-apron logistics.
1 paper · 0 benchmarks
This dataset contains 369 images of Trash used for deep learning.
1 paper · 1 benchmark
ArtDL is a novel painting data set for iconography classification composed of images collected from online sources.
1 paper · 1 benchmark
These images consist of a series of bacteria of the type Bacillus Subtilis that are suspended and captured by a digital microscope.
1 paper · 0 benchmarks
AtyPict is a dataset of atypical sketch content designed for atypical sketch content detection tasks.
1 paper · 0 benchmarks
Temporal Dataset for Indoor and In-Vehicle Thermal Comfort Estimation Abstract Thermal comfort estimation is essential for enhancing user experience in static indoor environments and dynamic in-vehicle scenarios.
1 paper · 0 benchmarks
Dataset Overview: 998 images and 4,208 annotations focusing on interaction with in-vehicle infotainment (IVI) systems.
1 paper · 0 benchmarks
The Autonomous-driving StreAming Perception (ASAP) benchmark is a benchmark to evaluate the online performance of vision-centric perception in autonomous driving.
1 paper · 0 benchmarks
Since robust foreground/background separation and segmentation of cellular objects (i.e.,identification of which pixels below to which objects) strongly depends on image quality, focus artifacts are detrimental to data quality.
1 paper · 0 benchmarks
BBBC041 (P. vivax (malaria) infected human blood smears)
P.
1 paper · 0 benchmarks
BCSD (Bank Check Segmentation Dataset)
The dataset consists of images of 158 filled out bank checks containing various complex backgrounds, and handwritten text and signatures in the respective fields, along with both pixel-level and patch-level segmentation masks for the…
1 paper · 0 benchmarks
BCSS (Breast Cancer Semantic Segmentation)
The BCSS dataset contains over 20,000 segmentation annotations of tissue regions from breast cancer images from The Cancer Genome Atlas (TCGA).
1 paper · 0 benchmarks
BD-TypoSAT (Building Damage Typology Satellite Dataset)
On Sunday, August 29, 2021, Hurricane Ida struck parts of Louisiana and Mississippi with wind gusts reaching up to 172 mph, leaving more than a million customers without electricity, including the entire New Orleans area.
1 paper · 0 benchmarks
BDD100K-weather is a dataset which is inherited from BDD100K using image attribute labels for Out-of-Distribution object detection.
1 paper · 0 benchmarks
BFN (Backdoored Face-Networks Dataset)
This database is a database of backdoored neural networks intended for face recognition.
1 paper · 0 benchmarks
BH-rPPG dataset (stands for Beihang University Remote PhotoPlethysmoGraphy) is a dataset consists of 3 lighting conditions with uneven distribution which collected in indoor environment.
1 paper · 0 benchmarks
BIDCD (Bosch Industrial Depth Completion Dataset)
Bosch Industrial Depth Completion Dataset (BIDCD) is an RGBD dataset for of static table-top scenes with industrial objects.
1 paper · 0 benchmarks
BIRDeep (BIRDeep_AudioAnnotations)
The BIRDeep Audio Annotations dataset is a collection of bird vocalizations from Doñana National Park, Spain.
1 paper · 0 benchmarks
BLN600 (BLN600: A Parallel Corpus of Machine/Human Transcribed Nineteenth Century Newspaper Texts)
A publicly available corpus of nineteenth-century newspaper text focused on crime in London, derived from the Gale British Library Newspapers corpus parts 1 and 2.
1 paper · 0 benchmarks
This is a dataset for Bengali Captioning from Images.
1 paper · 0 benchmarks
BPCIS (Bacterial Phase Contrast for Instance Segementation)
BPCIS is collection of 364 bacterial phase contrast images and corresponding label matrices for instance segmentation.
1 paper · 0 benchmarks
Brown Pedestrian Odometry Dataset (BPOD) is a dataset for benchmarking visual odometry algorithms in head-mounted pedestrian settings.
1 paper · 0 benchmarks
The dataset consists of images of bananas and apples.
1 paper · 0 benchmarks
In this dataset two robots, Baxter and UR5, perform 8 behaviors (look, grasp, pick, hold, shake, lower, drop, and push) on 95 objects that vary by 5 color (blue, green, red, white, and yellow), 6 contents (wooden button, plastic dices,…
1 paper · 0 benchmarks
BeautyFace is a dataset containing 3,000 high-quality face images with a higher resolution of 512512, covering more recent makeup styles and more diverse face poses, backgrounds, expressions, races, illumination.
1 paper · 0 benchmarks
Belfort (The Belfort dataset: Handwritten Text Recognition from Crowdsourced Annotations)
The Belfort dataset This dataset includes minutes of Belfort municipal council drawn up between 1790 and 1946.
1 paper · 1 benchmark
BigBIRD (Big Berkeley Instance Recognition Dataset)
BigBIRD is a 3D dataset of 125 objects, with the following data for each object: 600 12 megapixel images, sampling the viewing hemisphere 600 registered RGB-D point clouds from a Carmine 1.09 sensor Pose information for each of the above…
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
A large-scale hand pose dataset, collected using a novel capture method.
1 paper · 0 benchmarks
BioDrone is the first bionic drone-based single object tracking benchmark, it features videos captured from a flapping-wing UAV system with a major camera shake due to its aerodynamics.
1 paper · 0 benchmarks
Overview This is a dataset of blood cells photos.
1 paper · 0 benchmarks
A real-world low-light camera motion blur dataset for evaluating deblurring radiance fields methods.
1 paper · 0 benchmarks
The first large-scale dataset for training and evaluating novel-view synthesis from blurred images.
1 paper · 0 benchmarks
Boombox is a multi-modal dataset for visual reconstruction from acoustic vibrations.
1 paper · 0 benchmarks
Boreal Forest Fire (Boreal Forest Fire: UAV-collected Wildfire Detection and Smoke Segmentation Dataset)
This dataset consists of annotated images and videos of smoke resulting from prescribed burning events in Finnish boreal forests.
1 paper · 0 benchmarks
RGB-D instance segmentation box dataset.
1 paper · 2 benchmarks
Bramble flower image dataset (BRAMBLE FLOWER DETECTION AND CLASSIFICATION DATASET FOR PRECISION POLLINATION)
This dataset contains both the artificial and real flower images of bramble flowers.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.