Home › Datasets › modality › Images

Images datasets

archive 2025-07-28

3,239 datasets carry the modality tag "Images", ordered by the archive's paper count. Page 50 of 68: 48 shown of 3,239. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Images datasets 2353–2400 of 3,239

A Zero-Shot Sketch-based Inter-Modal Object Retrieval Scheme for Remote Sensing Images WITH the advancement in sensor technology, huge amounts of data are being collected from various satellites.
1 paper · 0 benchmarks
Edge-Map-345C is a large-scale edge-map dataset including 290,281 edge-maps corresponding to 345 object categories of QuickDraw dataset.
1 paper · 0 benchmarks
EgoMon (Egomon Gaze & Video dataset)
EgoMon Gaze & Video Dataset is an Egocentric (first person) Dataset that consists of 7 videos of 30 minutes, more or less, each one of them.
1 paper · 0 benchmarks
EleThermal (EleThermal: Infrared Elephant Images Dataset)
This is the Infrared Elephant Images Dataset (named 'EleThermal dataset') collected from here and annotated by our project, released under GPLv3.
1 paper · 0 benchmarks
Each HDF5 file has the following structure: energy Dataset {100000, 1} layer0 Dataset {100000, 3, 96} layer1 Dataset {100000, 12, 12} layer2 Dataset {100000, 12, 6} overflow Dataset {100000, 3} In practice, each file is a collection of…
1 paper · 0 benchmarks
A configurable synthetic dataset of simple shapes with ground truth concepts and known causal relationships between concepts and classes.
1 paper · 0 benchmarks
Embrapa ADD 256 (Embrapa Apples by Drones Detection Dataset)
This is a detailed description of the dataset, a data sheet for the dataset as proposed by Gebru et al.
1 paper · 0 benchmarks
The "Microbundle Time-lapse Dataset" contains 24 experimental time-lapse images of cardiac microbundles using three distinct types of experimental testbed of beating lab grown hiPSC-based cardiac microbundles.
1 paper · 0 benchmarks
EuroSAT-C is an open-source data set comprising algorithmically generated corruptions applied to the EuroSAT test set following the concept of ImageNet-C.
1 paper · 0 benchmarks
Event-Stream Dataset is a robotic grasping dataset with 91 objects.
1 paper · 0 benchmarks
A dataset of illusions generated by the AI model EIGen.
1 paper · 0 benchmarks
Minecraft Corpus dataset with builder utterance annotations
1 paper · 0 benchmarks
Extended Task10_Colon Medical Decathlon (Extended Task10_Colon of Medical Segmentation Decathlon dataset)
A dataset of abdominal CT studies in NifTi format from the open-source medical data repository Medical Decathlon was utilized.
1 paper · 1 benchmark
The proposed Extended-YouTube Faces (E-YTF) is an extension of the famous YouTube Faces (YTF) dataset and is specifically designed to further push the challenges of face recognition by addressing the problem of open-set face identification…
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
FAD (Face Attributes Dataset)
FAD is a dataset that have roughly 200,000 attribute labels for the above traits, for over 10,000 facial images.
1 paper · 0 benchmarks
FAS100K is a large-scale visual localization dataset.
1 paper · 0 benchmarks
FBIS-22M (Field Boundary Instance Segmentation - 22M)
FBIS-22M is the largest field boundary instance segmentation dataset to date, featuring over 22 million labeled field instances across more than 672 000 high-resolution satellite image patches.
1 paper · 0 benchmarks
FCoT (Foreground Chain-of-Thought)
FCoT (Chain‑of‑Thought Segmentation) is replicate the step-by-step reasoning process a human annotator follows when using SAM2 to generate masks.
1 paper · 0 benchmarks
FER2013 Blendshapes (FER2013 blendshapes dataset example (Partial))
Tables of the blendshapes from a group of the images of the FER2013 dataset, generated using MediaPipe library, based on the ARKit face blendshapes.
1 paper · 0 benchmarks
FETA Car-Manuals (FETA Car-Manuals dataset, image-text retrieval for foundation models' expert data performance.)
FETA benchmark focuses on text-to-image and image-to-text retrieval in public car manuals and sales catalogue brochures.
1 paper · 2 benchmarks
FETA benchmark focuses on text-to-image and image-to-text retrieval in public car manuals and sales catalogue brochures.
1 paper · 0 benchmarks
FFHQH (Flickr-Faces-HQ-Harmonization)
A new dataset for portrait harmonization based on the FFHQ.
1 paper · 1 benchmark
FGVD (Fine-Grained Vehicle Detection)
Fine-Grained Vehicle Detection (FGVD) is a dataset for fine-grained vehicle detection captured from a moving camera mounted on a car.
1 paper · 0 benchmarks
Optical images of printed circuit boards as well as detailed annotations of any text, logos, and surface-mount devices (SMDs).
1 paper · 0 benchmarks
FIGRIM (FIne-GRained Image Memorability)
This is a dataset of 9428 images, 1754 of which are target images with memorability scores.
1 paper · 0 benchmarks
FIND (Fused Image dataset for convolutional neural Network-based crack Detection)
The “Fused Image dataset for convolutional neural Network-based crack Detection” (FIND) is a large-scale image dataset with pixel-level ground truth crack data for deep learning-based crack segmentation analysis.
1 paper · 0 benchmarks
FMARS (Foundation Models Annotation in Remote Sensing)
FMARS is a large-scale dataset of Very High Resolution (VHR) remote sensing images with annotations generated using Vision Foundation Models.
1 paper · 0 benchmarks
FP4S (Floor plan image segmentation via scribble-based semi-weakly-supervised learning)
We introduce a new style- and category-agnostic floor plan image parsing benchmark developed in collaboration with professional architectural designers.
1 paper · 1 benchmark
FSOCO is a collaborative dataset for vision-based cone detection systems in Formula Student Driverless competitions.
1 paper · 0 benchmarks
Faces Through Time (FTT) features 26,247 images of notable people from the 19th to 21st centuries, with roughly 1,900 images per decade on average.
1 paper · 0 benchmarks
A Racial Fairness Benchmark Dataset for Face Forgery Detection.
1 paper · 0 benchmarks
The FashionFail dataset comprises 2,495 high-resolution images (2400x2400 pixels) of products found on e-commerce websites.
1 paper · 0 benchmarks
FashionRec (Fashion Recommendation Dataset)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
The FeatherV1 dataset is a dataset for fine-grained visual classification.
1 paper · 0 benchmarks
This dataset is a "part II" extension of the "Engineered cardiac microbundle time-lapse microscopy image dataset" and contains 808 experimental time-lapse image sequences of beating hiPSC-based cardiac microbundles using FibroTUG platforms…
1 paper · 0 benchmarks
Synthetic training set: This set is constructed in the following two steps and will be used for estimation/training purposes.
1 paper · 0 benchmarks
This dataset is collected by DataCluster Labs, India.
1 paper · 1 benchmark
The Foggy KITTI dataset extends the KITTI dataset to include challenging weather conditions, aiming to support research in real-world applications such as autonomous driving.
1 paper · 0 benchmarks
FooDI-ML (Food Drinks and groceries Images Multi Lingual)
Food Drinks and groceries Images Multi Lingual (FooDI-ML) is a dataset that contains over 1.5M unique images and over 9.5M store names, product names descriptions, and collection sections gathered from the Glovo application.
1 paper · 2 benchmarks
The 'Me 163' was a Second World War fighter airplane and a result of the German air force secret developments.
1 paper · 0 benchmarks
The Fraunhofer Portugal AICOS EDoF Dataset was produced within the TAMI project and is composed of images of microscopic fields of view (FOV) of Liquid-based Cervical Cytology (LBC) samples.
1 paper · 0 benchmarks
The Freiburg Street Crossing dataset consists of data collected from three different street crossings in Freiburg, Germany; ; two of which were traffic light regulated intersections and one a zebra crossing without traffic lights.
1 paper · 0 benchmarks
Freiburg Terrains consist of three parts: 3.7 hours of audio recordings of the microphone pointed at the robot wheels.
1 paper · 0 benchmarks
FunKPoint is a dataset for finding correspondences in visual data that has ground truth correspondences for 10 tasks and 20 object categories.
1 paper · 0 benchmarks
Demonstration data for 4 FurnitureBench tasks collected with a SpaceMouse using a DiffIK Controller.
1 paper · 0 benchmarks
This dataset was created to test whether it's possible to build a general-purpose detector that can tell real images apart from fake ones generated by convolutional neural networks (CNNs), no matter which model or dataset was used to…
1 paper · 0 benchmarks
The GDIT Aerial Airport dataset consists of aerial images containing instances of parked airplanes.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.