Browse State-of-the-Art › Medical Code Prediction
Medical Code Prediction
16 papers with code · 7 benchmarks · 7 datasets archive 2025-07-28
Context: Prediction of medical codes from clinical notes is both a practical and essential need for every healthcare delivery organization within current medical systems. Automating annotation will save significant time and excessive effort by human coders today. A new milestone will mark a meaningful step toward fully Autonomous Medical Coding in machines reaching parity with human coders' performance in medical code prediction.
Question: What exactly is the medical code prediction problem?
Answer: Clinical notes contain much information about what precisely happened during the patient's entire stay. And those clinical notes (e.g., discharge summary) is typically long, loosely structured, consists of medical domain language, and sometimes riddled with spelling errors. So, it's a highly multi-label classification problem, and the forthcoming ICD-11 standard will add more complexity to the problem! The medical code prediction problem is to annotate this clinical note with multiple codes subset from nearly 70K total codes (in the current ICD-10 system, for example).
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
7 leaderboard tables shown for this task, 7 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| MIMIC-III (18 rows) | GKI-ICD | A General Knowledge Injection Framework for ICD Coding | code | — | Compare |
| MIMIC-IV ICD-10 (6 rows) | PLM-ICD | Automated Medical Coding on MIMIC-III and MIMIC-IV: A Critical... | code | — | Compare |
| MIMIC-IV ICD-9 (6 rows) | PLM-ICD | Automated Medical Coding on MIMIC-III and MIMIC-IV: A Critical... | code | — | Compare |
| MIMIC-IV-ICD-10-full (5 rows) | MSMN | Mimic-IV-ICD: A new benchmark for eXtreme MultiLabel Classification | code | — | Compare |
| MIMIC-IV-ICD10-top50 (5 rows) | MSMN | Mimic-IV-ICD: A new benchmark for eXtreme MultiLabel Classification | code | — | Compare |
| MIMIC-IV-ICD9-top50 (5 rows) | MSMN | Mimic-IV-ICD: A new benchmark for eXtreme MultiLabel Classification | code | — | Compare |
| MIMIC-IV-ICD9-full (5 rows) | MSMN | Mimic-IV-ICD: A new benchmark for eXtreme MultiLabel Classification | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
7 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
16 shown of 16 papers with code (27 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
25 Nov 2019 3 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedThe innovations of our model are two-folds: it utilizes a multi-filter convolutional layer to capture various text patterns with different lengths and a residual convolutional layer to enlarge the receptive field.
-
15 Feb 2018 3 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedOur method aggregates information across the document using a convolutional neural network, and uses an attention mechanism to select the most relevant segments for each of the thousands of possible codes.
-
13 Jun 2024 2 repositories listedElectronic healthcare records are vital for patient safety as they document conditions, plans, and procedures in both free text and medical codes.
-
6 Sep 2021 2 repositories listedNevertheless, automated medical coding is still challenging because of the imbalanced class problem, complex code association, and noise in lengthy documents.
-
14 Jan 2021 2 repositories listedOur approach substantially outperforms previous results on top-50 medical code prediction on MIMIC-III dataset.
-
29 Oct 2020 2 repositories listedLE initialisation consistently boosted most deep learning models for automated medical coding.
-
13 Jul 2020 2 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 1 pointer-only (licence)In this paper, we propose a new label attention model for automatic ICD coding, which can handle both the various lengths and the interdependence of the ICD code related text fragments.
-
24 May 2016 2 repositories listedMIMIC-III (‘Medical Information Mart for Intensive Care’) is a large, single-center database comprising information relating to patients admitted to critical care units at a large tertiary care hospital.
-
21 Apr 2023 1 repository listedMedical coding is the task of assigning medical codes to clinical free-text documentation.
-
7 Oct 2022 1 repository listed Syntology ran 2 of 2 samples · 0 unverifiedAutomatic International Classification of Diseases (ICD) coding aims to assign multiple ICD codes to a medical note with average length of 3, 000+ tokens.
-
1 Oct 2022 1 repository listedHowever, existing studies did not exploit the discourse structure of clinical notes, which provides rich contextual information for code assignment.
-
3 Aug 2022 1 repository listedOne of the challenges in curriculum learning is the design of curricula -- i.
-
3 Mar 2022 1 repository listedAutomatic ICD coding is defined as assigning disease codes to electronic medical records (EMRs).
-
24 Jun 2021 1 repository listedTo address this problem, we propose a two-stage framework to improve automatic ICD coding by capturing the label correlation.
-
15 Jun 2021 1 repository listedSo we propose a model based on bidirectional encoder representations from transformers (BERT) using the sequence attention method for automatic ICD code assignment.
-
2 Apr 2021 1 repository listedMedical coding translates professionally written medical reports into standardized codes, which is an essential part of medical information systems and health insurance reimbursement.
Syntology lines on 4 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections