Home › Datasets › task › Image Classification

Image Classification datasets

archive 2025-07-28

283 datasets carry the task tag "Image Classification" (the task itself: Image Classification), ordered by the archive's paper count. Page 6 of 6: 43 shown of 283. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Image Classification datasets 241–283 of 283

This is the small version of the MuMiN dataset.
1 paper · 1 benchmark
The AASL-Clear dataset is a collection of RGB images featuring Arabic alphabet sign Language gestures with backgrounds removed.
1 paper · 1 benchmark
Number of images: 1,657 images during or after the fire If you use the dataset, please cite the following works: > Padilha, Rafael and Andaló, Fernanda A.
1 paper · 0 benchmarks
OLID I (An Open Leaf Image Dataset of Bangladesh's Major Crops)
The success of any AI-driven system relies heavily on vast amounts of training data.
1 paper · 0 benchmarks
Omni-Image is built as a challenging but tractable dataset for continual learning and few-shot learning.
1 paper · 0 benchmarks
A high-resolution multi-sensor remote sensing scene classification dataset, appropriate for training and evaluating image classification models in the remote sensing domain.
1 paper · 0 benchmarks
Orchid2024 is a fine-grained classification dataset specifically designed for Chinese Cymbidium orchid cultivars.
1 paper · 0 benchmarks
The Prima head pose dataset consists of 2790 images of 15 persons recorded twice.
1 paper · 1 benchmark
Photozilla is a large-scale dataset which includes over 990k images belonging to 10 different photographic styles.
1 paper · 0 benchmarks
PoTATO is a dataset designed to enhance the detection of floating plastic waste in aquatic environments by leveraging polarimetric imaging.
1 paper · 0 benchmarks
RGB Arabic Alphabet Sign Language (AASL) dataset
1 paper · 1 benchmark
SPOT-10 (Animal Pattern Benchmark Dataset for Machine Learning Algorithms)
The SPOTS-10 dataset is an extensive collection of grayscale images showcasing diverse patterns found in ten animal species.
1 paper · 1 benchmark
SSBI Dataset (Synthetic Signature Bankcheck Images)
The Synthetic Signature Bankcheck Images (SSBI) Dataset is the first publicly available dataset of bank check images with annotations for detecting handwritten components, including names, amounts, dates, and signatures.
1 paper · 0 benchmarks
SVLD (Social Vision and Language Dataset)
The social vision and language dataset is a large-scale multimodal dataset designed for research into social contextual learning.
1 paper · 0 benchmarks
To construct such a dataset, a straightforward approach was scraping images from the web.
1 paper · 1 benchmark
SolarDK is a dataset for the detection and localization of solar.
1 paper · 0 benchmarks
The SuSy Dataset combines authentic photographs and AI-generated images designed for training and evaluating synthetic image detection models.
1 paper · 0 benchmarks
A public open dataset of synthetic chest X-ray images of COVID-19.
1 paper · 0 benchmarks
TCB-DS (Toxigenic Cyanobacteria Dataset)
The TCB-DS dataset is a specialized collection of microscopic images focusing on the automatic recognition of cyanobacteria genera.
1 paper · 0 benchmarks
TEM nanowire morphologies for classification and segmentation (Transmission electron microscopy (TEM) image datasets of peptide / protein nanowire morphologies)
TEM image dataset containing four nanowire morphologies of bio-derived protein nanowires and synthetic peptide nanowires.
1 paper · 0 benchmarks
Tsinghua Dogs is a fine-grained classification dataset for dogs, over 65% of whose images are collected from people's real life.
1 paper · 0 benchmarks
The TwinSynths dataset is a novel benchmark designed to overcome common limitations found in earlier synthetic image datasets, such as low image quality, inadequate content preservation, and limited class diversity.
1 paper · 0 benchmarks
About Dataset Context With increasing unrest, security cams should be armed with advance technology.
1 paper · 0 benchmarks
The YFCC100M Fine-Grained Geolocation dataset is a subset of 100 a set of 36,146 YFCC100M images that had Flickr tags that could be identified as corresponding to one of the labels in the iNaturalist 2017 dataset.
1 paper · 0 benchmarks
The iNaturalist Fine-Grained Geolocation dataset is an extension of the iNaturalist dataset with complementary geolocation information.
1 paper · 0 benchmarks
idsprites (Infinite dSprites)
Easily generate simple continual learning benchmarks.
1 paper · 0 benchmarks
topex-printer is a dataset containing 102 machine parts of a label printing machine.
1 paper · 0 benchmarks
ADFI (Anomaly Detection Datasets for Visual Inspection)
ADFI Dataset is an image dataset for anomaly detection methods with a focus on industrial inspection.
0 papers · 0 benchmarks
Dataset contains images with apples infected by scab.
0 papers · 0 benchmarks
Dataset contains images with apple leaves infected by scab.
0 papers · 0 benchmarks
CCIC (Concrete Crack Images for Classification)
The dataset contains concrete images having cracks.
0 papers · 0 benchmarks
CEAHB2021-5 (Chinese Ethnic Ancient Handwritten Books database)
Ancient books script identification of Chinese ethnic minorities with deep convolutional neural networks via multi-branch and spatial pyramid pooling Automatic classification of ancient books is an important component of the digital…
0 papers · 0 benchmarks
This dataset is the images of corn seeds considering the top and bottom view independently (two images for one corn seed: top and bottom).
0 papers · 0 benchmarks
This dataset contains 16,000 images of four shapes; square, star, circle, and triangle.
0 papers · 0 benchmarks
The Lusitano dataset was collected over a 3-month period, spanning from January to March, from Paulo de Oliveira, S.A., a prominent textile company, based in Covilhã, Portugal, renowned for its innovative contributions to the textile…
0 papers · 0 benchmarks
A dataset of all Moroccan money
0 papers · 0 benchmarks
Mudestreda (Mudestreda Multimodal Device State Recognition Dataset)
Mudestreda Multimodal Device State Recognition Dataset obtained from real industrial milling device with Time Series and Image Data for Classification, Regression, Anomaly Detection, Remaining Useful Life (RUL) estimation, Signal Drift…
0 papers · 0 benchmarks
Sakha-TB (400+400 CXR images for TB diagnosis)
Sakha-TB is a de-identified image dataset of frontal chest X-rays (CXR), collected through collaboration with several medical institutions in the Republic of Sakha (Yakutia, Russia).
0 papers · 0 benchmarks
TCMP-300 (Traditional Chinese Medicinal Plant Dataset)
Traditional Chinese medicinal plants are often used to prevent and treat diseases for the human body.
0 papers · 1 benchmark
eAppleScab (Apple Scab in the Early Stage of Development)
The study showed that the apple scab can be detected in the high-resolution RGB images in an early stage of its development.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.