Browse State-of-the-Art › Multi-Label Text Classification

Multi-Label Text Classification

78 papers with code · 20 benchmarks · 13 datasets archive 2025-07-28

MethodologyNatural Language Processing

According to Wikipedia "In machine learning, multi-label classification and the strongly related problem of multi-output classification are variants of the classification problem where multiple labels may be assigned to each instance. Multi-label classification is a generalization of multiclass classification, which is the single-label problem of categorizing instances into precisely one of more than two classes; in the multi-label problem there is no constraint on how many of the classes the instance can be assigned to."

Description from the archive archive 2025-07-28.

Benchmarks archive 2025-07-28

20 leaderboard tables shown for this task, 20 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 20 until expanded.

DatasetBest model (first row in archive order)PaperCodeSyntologyCompare
Reuters-21578 (7 rows) TIACBM Task-Informed Anti-Curriculum by Masking Improves Downstream... code — Compare
CC3M-TagMask (6 rows) TTD (w/ fine-tuning) TTD: Text-Tag Self-Distillation Enhancing Image-Text Alignment in... code Syntology ran 1 of 1 samples · 0 unverified Compare
AAPD (5 rows) LSAN Label-Specific Document Representation for Multi-Label Text Classification code — Compare
Freecode (5 rows) TagBERT Tag Recommendation for Online Q&A Communities based on BERT... code — Compare
BVICTOR (3 rows) XGBoost VICTOR: a Dataset for Brazilian Legal Documents Classification code — Compare
EUR-Lex (3 rows) bert-base Large-Scale Multi-Label Text Classification on EU Legislation code — Compare
MIMIC-III (3 rows) HLAN Explainable Automated Coding of Clinical Notes using Hierarchical... code — Compare
MVICTOR (theme) (3 rows) XGBoost VICTOR: a Dataset for Brazilian Legal Documents Classification code — Compare
SVICTOR (theme) (3 rows) XGBoost VICTOR: a Dataset for Brazilian Legal Documents Classification code — Compare
MIMIC-III-50 (2 rows) D2SBERT using Sequence Attention Medical Code Prediction from Discharge Summary: Document to... code — Compare
Amazon-12K (1 row) LAHA Label-aware Document Representation via Hybrid Attention for... code — Compare
Dataset of Propaganda Techniques of the State-Sponsored Information Operation of the People's Republic of China (1 row) Bert Dataset of Propaganda Techniques of the State-Sponsored... code — Compare
Kan-Shan Cup (1 row) LAHA Label-aware Document Representation via Hybrid Attention for... code — Compare
LF-AmzonTitles-131K (1 row) ECLARE ECLARE: Extreme Classification with Label Graph Correlations code Syntology ran 1 of 1 samples · 0 unverified Compare
LF-AmazonTitles-131K (1 row) DECAF ECLARE: Extreme Classification with Label Graph Correlations code Syntology ran 1 of 1 samples · 0 unverified Compare
RCV1 (1 row) HiddeN Joint Learning of Hyperbolic Label Embeddings for Hierarchical... code — Compare
RCV1-v2 (1 row) MAGNET MAGNET: Multi-Label Text Classification using Attention-based... code — Compare
Slashdot (1 row) MAGNET MAGNET: Multi-Label Text Classification using Attention-based... code — Compare
USPTO-3M (1 row) BERT PatentBERT: Patent Classification with Fine-Tuning a pre-trained BERT Model code — Compare
Wiki-30K (1 row) LAHA Label-aware Document Representation via Hybrid Attention for... code — Compare

Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.

Libraries

Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.

Datasets archive 2025-07-28

13 datasets whose archive record lists this task, ordered by the archive's paper count.

Subtasks archive 2025-07-28

No subtask under this task in the archive's task tree.

Parent tasks archive 2025-07-28

Most implemented papers archive 2025-07-28

30 shown of 78 papers with code (171 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.

Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections