Home › Datasets › task › Multi-Label Classification
Multi-Label Classification datasets
archive 2025-07-28
30 datasets carry the task tag "Multi-Label Classification" (the task itself: Multi-Label Classification), ordered by the archive's paper count. Page 1 of 1: 30 shown of 30. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Multi-Label Classification datasets 1–30 of 30
The COCO (Common Objects in Context) dataset is a large-scale object detection, segmentation, and captioning dataset.
11,922 papers · 77 benchmarks
The CheXpert dataset contains 224,316 chest radiographs of 65,240 patients with both frontal and lateral views available.
628 papers · 3 benchmarks
The NUS-WIDE dataset contains 269,648 images with a total of 5,018 tags collected from Flickr.
348 papers · 3 benchmarks
MIMIC-CXR from Massachusetts Institute of Technology presents 371,920 chest X-rays associated with 227,943 imaging studies from 65,079 patients.
240 papers · 3 benchmarks
ChestX-ray14 is a medical imaging dataset which comprises 112,120 frontal-view X-ray images of 30,805 (collected from the year of 1992 to 2015) unique patients with the text-mined fourteen common disease labels, mined from the text…
237 papers · 6 benchmarks
PASCAL VOC 2007 is a dataset for image recognition.
126 papers · 13 benchmarks
ECHR is an English legal judgment prediction dataset of cases from the European Court of Human Rights (ECHR).
47 papers · 1 benchmark
The friedman1 data set is commonly used to test semi-supervised regression methods.
31 papers · 0 benchmarks
The MRNet dataset consists of 1,370 knee MRI exams performed at Stanford University Medical Center.
29 papers · 1 benchmark
OpenImages V6 is a large-scale dataset , consists of 9 million training images, 41,620 validation samples, and 125,456 test samples.
23 papers · 2 benchmarks
Our goal is to improve upon the status quo for designing image classification models trained in one domain that perform well on images from another domain.
22 papers · 3 benchmarks
CAL500 (Computer Audition Lab 500)
CAL500 (Computer Audition Lab 500) is a dataset aimed for evaluation of music information retrieval systems.
21 papers · 0 benchmarks
The objective in extreme multi-label classification is to learn feature architectures and classifiers that can automatically tag a data point with the most relevant subset of labels from an extremely large label set.
18 papers · 0 benchmarks
LSHTC is a dataset for large-scale text classification.
18 papers · 0 benchmarks
MLRSNet is a a multi-label high spatial resolution remote sensing dataset for semantic scene understanding.
17 papers · 2 benchmarks
Arxiv ASTRO-PH (Astro Physics) collaboration network is from the e-print arXiv and covers scientific collaborations between authors papers submitted to Astro Physics category.
10 papers · 0 benchmarks
Moviescope is a large-scale dataset of 5,000 movies with corresponding video trailers, posters, plots and metadata.
6 papers · 0 benchmarks
CHiME-Home is a dataset for sound source recognition in a domestic environment.
5 papers · 0 benchmarks
Sewer-ML is a sewer defect dataset.
5 papers · 0 benchmarks
For each dataset we provide a short description as well as some characterization metrics.
4 papers · 0 benchmarks
Aims to help V-NLIs recognize analytic tasks from free-form natural language by training and evaluating cutting-edge multi-label classification models.
4 papers · 0 benchmarks
This dataset contains Bangla handwritten numerals, basic characters and compound characters.
3 papers · 2 benchmarks
CAVES (A Dataset to facilitate Explainable Classification and Summarization of Concerns towards COVID Vaccines)
CAVES is the first large-scale dataset containing about 10k COVID-19 anti-vaccine tweets labelled into various specific anti-vaccine concerns in a multi-label setting.
2 papers · 0 benchmarks
The dataset consists of titles and abstracts from NLP-related papers.
2 papers · 0 benchmarks
This data is for the Mis2-KDD 2021 under review paper: Dataset of Propaganda Techniques of the State-Sponsored Information Operation of the People’s Republic of China We present our dataset that focuses on propaganda techniques in Mandarin…
1 paper · 1 benchmark
MTC is a financial-domain dataset of the multi-label topic classification task.
1 paper · 0 benchmarks
ScienceExamCER is a collection of resources for studying explanation-centered inference, including explanation graphs for 1,680 questions, with 4,950 tablestore rows, and other analyses of the knowledge required to answer elementary and…
1 paper · 0 benchmarks
Trailers12k is a movie trailer dataset comprised of 12,000 titles associated to ten genres.
1 paper · 0 benchmarks
This is a dataset of scientific documents derived from arXiv.
1 paper · 0 benchmarks
The KIT Whole-Body Human Motion Database is a large-scale dataset of whole-body human motion with methods and tools, which allows a unifying representation of captured human motion, and efficient search in the database, as well as the…
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.