Home › Datasets › task › Classification

Classification datasets

archive 2025-07-28

222 datasets carry the task tag "Classification" (the task itself: Classification), ordered by the archive's paper count. Page 2 of 5: 48 shown of 222. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Classification datasets 49–96 of 222

As part of an ongoing worldwide effort to comprehend and monitor insect biodiversity, we present the BIOSCAN-5M Insect dataset to the machine learning community.
4 papers · 0 benchmarks
This brain tumor dataset contains 3064 T1-weighted contrast-enhanced images with three kinds of brain tumor.
4 papers · 0 benchmarks
The quality of AI-generated images has rapidly increased, leading to concerns of authenticity and trustworthiness.
4 papers · 2 benchmarks
The data set covers recordings of ripening fruit with labels of destructive measurements (fruit flesh firmness, sugar content and overall ripeness).
4 papers · 1 benchmark
DiaMOS Plant (A Dataset for Diagnosis and Monitoring Plant Disease)
Abstract The classification and recognition of foliar diseases is an increasingly developing field of research, where the concepts of machine and deep learning are used to support agricultural stakeholders.
4 papers · 0 benchmarks
ECG-Image-Database (Digitization and Classification of ECG Images: The George B. Moody PhysioNet Challenge 2024)
The George B.
4 papers · 1 benchmark
Oracle-MNIST (Oracle-MNIST: a Realistic Image Dataset for Benchmarking Machine Learning Algorithms)
We introduce the Oracle-MNIST dataset, comprising of 2828 grayscale images of 30,222 ancient characters from 10 categories, for benchmarking pattern classification, with particular challenges on image noise and distortion.
4 papers · 1 benchmark
Sentiment140 is a dataset that allows you to discover the sentiment of a brand, product, or topic on Twitter.
4 papers · 1 benchmark
ACL-Fig is a large-scale automatically annotated corpus consisting of 112,052 scientific figures extracted from 56K research papers in the ACL Anthology.
3 papers · 0 benchmarks
Attention Deficit Hyperactivity Disorder (ADHD) affects at least 5-10% of school-age children and is associated with substantial lifelong impairment, with annual direct costs exceeding $36 billion/year in the US.
3 papers · 0 benchmarks
For a detailed description, we refer to Section 3 in our research article.
3 papers · 0 benchmarks
In an effort to catalog insect biodiversity, we propose a new large dataset of hand-labelled insect images, the BIOSCAN-1M Insect Dataset.
3 papers · 1 benchmark
The normal chest X-ray (left panel) depicts clear lungs without any areas of abnormal opacification in the image.
3 papers · 2 benchmarks
The dermatology differential diagnoses (ddx) dataset for skin condition classification includes expert annotations and model predictions for 1947 cases.
3 papers · 0 benchmarks
Hephaestus (Hephaestus: A large scale multitask dataset towards InSAR understanding)
Hephaestus is the first large-scale InSAR dataset.
3 papers · 0 benchmarks
LLeQA (Long-form Legal Question Answering)
LLeQA is a French native dataset for studying information retrieval and long-form question answering in the legal domain.
3 papers · 0 benchmarks
MS-FIMU (Mobility Scenario FIMU)
Open Dataset: Mobility Scenario FIMU An open, multidimensional (6 categorical attributes), and synthetic dataset of faked virtual humans generated by an optimization approach applied to a real-life call-detail-records-based anonymized…
3 papers · 0 benchmarks
MUTE (Multimodal Bengali Hateful Memes Dataset)
MUTE This is the first open-source Bengali Hateful Meme dataset, consisting of around 4200 memes annotated with two labels: hate and not hate.
3 papers · 0 benchmarks
Description 12-channel lung sounds for each patient Multi-channel Analysis opportunity 5 COPD severities (COPD0, COPD1, COPD2, COPD3, COPD4) Short-term recordings (At least 17s) The details: https://doi.org/10.28978/nesciences.349282 This…
3 papers · 1 benchmark
SupplyGraph (SupplyGraph: A Benchmark Dataset for Supply Chain Planning using Graph Neural Networks)
Graph Neural Networks (GNNs) have gained traction across different domains such as transportation, bio-informatics, language processing, and computer vision.
3 papers · 0 benchmarks
Huggingface Datasets is a great library, but it lacks standardization, and datasets require preprocessing work to be used interchangeably.
3 papers · 0 benchmarks
Tiny ImageNetv2 is a subset of the ImageNetV2 (matched frequency) dataset by Recht et al.
3 papers · 0 benchmarks
Urban is one of the most widely used hyperspectral data used in the hyperspectral unmixing study.
3 papers · 1 benchmark
In our benchmark WHYSHIFT, we explore distribution shifts on 5 real-world tabular datasets from the economic and traffic sectors with natural spatiotemporal distribution shifts.We only pick 7 typical settings out of 22 settings and select…
3 papers · 0 benchmarks
We present XHate-999, a multi-domain and multilingual evaluation data set for abusive language detection.
3 papers · 0 benchmarks
BreastClassifications4 ([MIMBCD-UI] UTA4: Severity & Pathology Classifications Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
2 papers · 0 benchmarks
The researchers of Qatar University have compiled the COVID-QU-Ex dataset, which consists of 33,920 chest X-ray (CXR) images including: 11,956 COVID-19 11,263 Non-COVID infections (Viral or Bacterial Pneumonia) 10,701 Normal Ground-truth…
2 papers · 0 benchmarks
CRC100K (100,000 histological images of human colorectal cancer and healthy tissue)
This is a set of 100,000 non-overlapping image patches from hematoxylin & eosin (H&E) stained histological images of human colorectal cancer (CRC) and normal tissue.
2 papers · 0 benchmarks
CWD30 (Crop Weed Dataset 30 species)
CWD30 comprises over 219,770 high-resolution images of 20 weed species and 10 crop species, encompassing various growth stages, multiple viewing angles, and environmental conditions.
2 papers · 0 benchmarks
Deep PCB (Deep Printed Circuit Board)
DeepPCB Dataset Link : A dataset contains 1,500 image pairs, each of which consists of a defect-free template image and an aligned tested image with annotations including positions of 6 most common types of PCB defects: open, short,…
2 papers · 1 benchmark
This is a dataset used to test deep learning-supported deep learning for fault diagnosis: - A digital twin model for a robot.
2 papers · 1 benchmark
Extended Agriculture-Vision dataset comprises two parts: 1.
2 papers · 0 benchmarks
We provide multiple human annotations for each test image in Fashion-MNIST.
2 papers · 0 benchmarks
FinBench is a benchmark for evaluating the performance of machine learning models with both tabular data inputs and profile text inputs.
2 papers · 0 benchmarks
FracAtlas (A Dataset for Fracture Classification, Localization and Segmentation of Musculoskeletal Radiographs)
FractureAtlas is a musculoskeletal bone fracture dataset with annotations for deep learning tasks like classification, localization, and segmentation.
2 papers · 0 benchmarks
HCP Aging (Lifespan Human Connectome Project Aging)
Lifespan HCP Release 2.0 includes cross-sectional visit 1 (V1) preprocessed structural and functional imaging data, unprocessed V1 imaging data for all included modalities (structural, high-res hippocampal T2, resting state fMRI, task…
2 papers · 1 benchmark
The IRFL dataset consists of idioms, similes, and metaphors with matching figurative and literal images, as well as two novel tasks of multimodal figurative understanding and preference.
2 papers · 2 benchmarks
The Insider Threat Test Dataset is a collection of synthetic insider threat test datasets that provide both background and malicious actor synthetic data.
2 papers · 1 benchmark
LEPISZCZE is an open-source comprehensive benchmark for Polish NLP and a continuous-submission leaderboard, concentrating public Polish datasets (existing and new) in specific tasks.
2 papers · 0 benchmarks
LUMA (Learning from Uncertain and Multimodal Data)
LUMA is a multimodal dataset that consists of audio, image, and text modalities.
2 papers · 0 benchmarks
"We built a large lung CT scan dataset for COVID-19 by curating data from 7 public datasets listed in the acknowledgements.
2 papers · 2 benchmarks
This data set contains 775 video sequences, captured in the wildlife park Lindenthal (Cologne, Germany) as part of the AMMOD project, using an Intel RealSense D435 stereo camera.
2 papers · 0 benchmarks
LymphoMNIST is a comprehensive dataset designed for the nuanced classification of lymphocyte images.
2 papers · 0 benchmarks
MCSI (Mpox Close Skin Images)
The Mpox Close Skin Images dataset (MCSI) is a collection of skin images obtained from diverse public sources, that we accurately pre-processed (i.e., cropped and zoomed) in order to focus the skin lesion (if present), and to evaluate…
2 papers · 0 benchmarks
This repository provides a cleaned dataset, which is intended to be used for text classification, language modeling, and AI-generated content detection tasks.
2 papers · 0 benchmarks
MedMNIST-C is an open-source data set collection comprising algorithmically generated corruptions applied to the test sets of the MedMNIST collection following the concept of ImageNet-C.
2 papers · 0 benchmarks
The process by which sections in a document are demarcated and labeled is known as section identification.
2 papers · 2 benchmarks
The dataset consists of titles and abstracts from NLP-related papers.
2 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.