Home › Datasets › task › Image Classification

Image Classification datasets

archive 2025-07-28

283 datasets carry the task tag "Image Classification" (the task itself: Image Classification), ordered by the archive's paper count. Page 4 of 6: 48 shown of 283. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Image Classification datasets 145–192 of 283

HErlev (HErlev Pap Smear Dataset)
5 papers · 1 benchmark
ImageNet-Hard is a new benchmark that comprises 10,980 images collected from various existing ImageNet-scale benchmarks (ImageNet, ImageNet-V2, ImageNet-Sketch, ImageNet-C, ImageNet-R, ImageNet-ReaL, ImageNet-A, and ObjectNet).
5 papers · 1 benchmark
ImageNet-Patch: A Dataset for Benchmarking Machine Learning Robustness against Adversarial Patches Adversarial patches are optimized contiguous pixel blocks in an input image that cause a machine-learning model to misclassify it.
5 papers · 0 benchmarks
InsPLAD (Inspection Power Line Asset Dataset)
InsPLAD is a Dataset for Power Line Asset Inspection containing 10,607 high-resolution Unmanned Aerial Vehicles colour images.
5 papers · 1 benchmark
Context This is image data of Natural Scenes around the world.
5 papers · 2 benchmarks
MuMiN is a misinformation graph dataset containing rich social media data (tweets, replies, users, images, articles, hashtags), spanning 21 million tweets belonging to 26 thousand Twitter threads, each of which have been semantically…
5 papers · 0 benchmarks
NIH-CXR-LT (Long-tailed (LT) NIH ChestXRay14)
NIH-CXR-LT.
5 papers · 1 benchmark
RF100 (Roboflow 100)
The evaluation of object detection models is usually performed by optimizing a single metric, e.g.
5 papers · 1 benchmark
SIPaKMeD (SIPaKMeD Pap Smear dataset)
a high-level explanation of the dataset characteristics explain motivations and summary of its content potential use cases of the dataset
5 papers · 1 benchmark
Sewer-ML is a sewer defect dataset.
5 papers · 0 benchmarks
- Games dataset containing 100,000 Gameplay Images of 175 Video Games across 10 Sports Genres - AMERICAN FOOTBALL, BASKETBALL, BIKE RACING, CAR RACING, FIGHTING, HOCKEY, SOCCER, TABLE TENNIS, TENNIS.
5 papers · 2 benchmarks
Tencent ML-Images is a large open-source multi-label image database, including 17,609,752 training and 88,739 validation image URLs, which are annotated with up to 11,166 categories.
5 papers · 0 benchmarks
CI-MNIST (Correlated and Imbalanced MNIST)
CI-MNIST (Correlated and Imbalanced MNIST) is a variant of MNIST dataset with introduced different types of correlations between attributes, dataset features, and an artificial eligibility criterion.
4 papers · 0 benchmarks
DiagSet is a histopathological dataset for prostate cancer detection.
4 papers · 0 benchmarks
ETHEC (ETH Entomological Collection (ETHEC) Dataset)
It includes 47,978 butterfly images with a 4-level label-hierarchy.
4 papers · 0 benchmarks
A SAR version of the EuroSAT dataset.
4 papers · 1 benchmark
The KTH-TIPS (Textures under varying Illumination, Pose and Scale) image database was created to extend the CUReT database in two directions, by providing variations in scale as well as pose and illumination, and by imaging other samples…
4 papers · 1 benchmark
Kuzushiji-Kanji is an imbalanced dataset of total 3832 Kanji characters (64x64 grayscale, 140,426 images), ranging from 1,766 examples to only a single example per class.
4 papers · 0 benchmarks
LIMUC (Labeled Images for Ulcerative Colitis)
The LIMUC dataset is the largest publicly available labeled ulcerative colitis dataset that compromises 11276 images from 564 patients and 1043 colonoscopy procedures.
4 papers · 1 benchmark
MIMIC-CXR-LT (long-tailed version of MIMIC-CXR)
MIMIC-CXR-LT.
4 papers · 1 benchmark
The MNIST Large Scale dataset is based on the classic MNIST dataset, but contains large scale variations up to a factor of 16.
4 papers · 1 benchmark
Oracle-MNIST (Oracle-MNIST: a Realistic Image Dataset for Benchmarking Machine Learning Algorithms)
We introduce the Oracle-MNIST dataset, comprising of 2828 grayscale images of 30,222 ancient characters from 10 categories, for benchmarking pattern classification, with particular challenges on image noise and distortion.
4 papers · 1 benchmark
WHU-Hi (Wuhan UAV-borne hyperspectral image)
WHU-Hi dataset (Wuhan UAV-borne hyperspectral image) is collected and shared by the RSIDEA research group of Wuhan University, and it could serve as a benchmark dataset for precise crop classification and hyperspectral image classification…
4 papers · 0 benchmarks
ACL-Fig is a large-scale automatically annotated corpus consisting of 112,052 scientific figures extracted from 56K research papers in the ACL Anthology.
3 papers · 0 benchmarks
It contains about 28K medium quality animal images belonging to 10 categories: dog, cat, horse, spyder, butterfly, chicken, sheep, cow, squirrel, and elephant.
3 papers · 1 benchmark
Atlas is a dataset for e-commerce clothing product categorization.
3 papers · 0 benchmarks
DEIC Benchmark (Data-Efficient Image Classification Benchmark)
DEIC is a benchmark for measuring the data efficiency of models in the context of image classification.
3 papers · 3 benchmarks
Human face Deepfake dataset sampled from large datasets - High Quality Dataset - Diverse Dataset - Challenging Dataset - Large Dataset - Text prompts
3 papers · 0 benchmarks
DirtyMNIST is a concatenation of MNIST + AmbiguousMNIST, with 60k samples each in the training set.
3 papers · 0 benchmarks
FMD (materials) (Flickr Material Dataset)
Sharan, Lavanya, Ruth Rosenholtz, and Edward Adelson.
3 papers · 1 benchmark
FireRisk (FireRisk: A Remote Sensing Dataset for Fire Risk Assessment)
In this work, we propose a novel remote sensing dataset, FireRisk, consisting of 7 fire risk classes with a total of 91 872 labelled images for fire risk assessment.
3 papers · 1 benchmark
The Image and Video Advertisements collection consists of an image dataset of 64,832 image ads, and a video dataset of 3,477 ads.
3 papers · 0 benchmarks
LKS (Liver Kidney Stomach)
LKS is a dataset of 684 Liver-Kidney-Stomach immunofluorescence whole slide images (WSIs) used in the investigation of autoimmune liver disease.
3 papers · 0 benchmarks
LSA16 (Lengua de Señas Argentina - 16 Handshapes classes)
This database contains images of 16 handshapes of the Argentinian Sign Language (LSA), each performed 5 times by 10 different subjects, for a total of 800 images.
3 papers · 1 benchmark
OFDIW (OnFocus Detection In the Wild)
OnFocus Detection In the Wild (OFDIW) is an onfocus detection dataset.
3 papers · 0 benchmarks
This paper introduces the RGB Arabic Alphabet Sign Language (AASL) dataset.
3 papers · 0 benchmarks
SIDD-Image (Segmented Intrusion Detection Dataset)
This is the first image-based network intrusion detection dataset.
3 papers · 1 benchmark
A new dataset for streaming classification consisting of temporally correlated images from 51 distinct object categories and additional evaluation classes outside of the training distribution to test novelty recognition.
3 papers · 0 benchmarks
ASIRRA ((Animal Species Image Recognition for Restricting Access)
Web services are often protected with a challenge that's supposed to be easy for people to solve, but difficult for computers.
2 papers · 0 benchmarks
AdvNet is a dataset of traffic signs images.
2 papers · 0 benchmarks
CARBEN (Composite Adversarial Robustness Benchmark)
Prior literature on adversarial attack methods has mainly focused on attacking with and defending against a single threat model, e.g., perturbations bounded in Lp ball.
2 papers · 0 benchmarks
CHAMMI (CHAMMI: A benchmark for channel-adaptive models in microscopy imaging)
We present a cellular microscopic image dataset for investigating channel-adaptive models.
2 papers · 0 benchmarks
The appearance of the world varies dramatically not only from place to place but also from hour to hour and month to month.
2 papers · 1 benchmark
DF20 - Mini (Danish Fungi 2020 - Mini)
Danish Fungi 2020 (DF20) is a novel fine-grained dataset and benchmark.
2 papers · 1 benchmark
DIB-10K (DongNiao International Birds 10000)
Is a challenging image dataset which has more than 10 thousand different types of birds.
2 papers · 1 benchmark
Deep PCB (Deep Printed Circuit Board)
DeepPCB Dataset Link : A dataset contains 1,500 image pairs, each of which consists of a defect-free template image and an aligned tested image with annotations including positions of 6 most common types of PCB defects: open, short,…
2 papers · 1 benchmark
HiAML Computational Graph (CG) family introduced in "GENNAPE: Towards Generalized Neural Architecture Performance Estimators", accepted to AAAI-23.
2 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.