Home › Datasets › modality › Medical

Medical datasets

archive 2025-07-28

394 datasets carry the modality tag "Medical", ordered by the archive's paper count. Page 5 of 9: 48 shown of 394. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Medical datasets 193–240 of 394

CMMD (The Chinese Mammography Database)
Breast carcinoma is the second largest cancer in the world among women.
3 papers · 1 benchmark
ExtMarker (3D motion of chest external markers)
Three-dimensional position of external markers placed on the chest and abdomen of healthy individuals breathing during intervals from 73s to 222s.
3 papers · 1 benchmark
Fetoscopic Placental Vessel Segmentation and Registration (FetReg) is a large-scale multi-centre dataset for the development of generalized and robust semantic segmentation and video mosaicking algorithms for the fetal environment with a…
3 papers · 0 benchmarks
GraSP (Holistic and Multi-Granular Surgical Scene Understanding of Prostatectomies)
Holistic and Multi-Granular Surgical Scene Understanding of Prostatectomies (GraSP) dataset, a curated benchmark that models surgical scene understanding as a hierarchy of complementary tasks with varying levels of granularity.
3 papers · 1 benchmark
HiRID is a freely accessible critical care dataset containing data relating to almost 34 thousand patient admissions to the Department of Intensive Care Medicine of the Bern University Hospital, Switzerland (ICU), an interdisciplinary…
3 papers · 6 benchmarks
IBC (Individual Brain Charting)
The Individual Brain Charting (IBC) project aims at providing a new generation of functional-brain atlases.
3 papers · 0 benchmarks
Kvasir-VQA (A Text-Image Pair GI Tract Dataset)
The Kvasir-VQA dataset is an extended dataset derived from the HyperKvasir and Kvasir-Instrument datasets, augmented with question-and-answer annotations.
3 papers · 0 benchmarks
The dataset contains a Video capsule endoscopy dataset for polyp segmentation.
3 papers · 1 benchmark
LKS (Liver Kidney Stomach)
LKS is a dataset of 684 Liver-Kidney-Stomach immunofluorescence whole slide images (WSIs) used in the investigation of autoimmune liver disease.
3 papers · 0 benchmarks
MIMIC-IV-ECG (MIMIC-IV-ECG: Diagnostic Electrocardiogram Matched Subset)
The MIMIC-IV-ECG module contains approximately 800,000 diagnostic electrocardiograms across nearly 160,000 unique patients.
3 papers · 0 benchmarks
Question Answering (QA) is a widely-used framework for developing and evaluating an intelligent machine.
3 papers · 0 benchmarks
MIT-BIH AFDB (MIT-BIH Atrial Fibrilation Database)
This database includes 25 long-term ECG recordings of human subjects with atrial fibrillation (mostly paroxysmal).
3 papers · 0 benchmarks
Operating rooms (ORs) are complex, high-stakes environments requiring precise understanding of interactions among medical staff, tools, and equipment for enhancing surgical assistance, situational awareness, and patient safety.
3 papers · 3 benchmarks
The “Medico automatic polyp segmentation challenge” aims to develop computer-aided diagnosis systems for automatic polyp segmentation to detect all types of polyps (for example, irregular polyp, smaller or flat polyps) with high efficiency…
3 papers · 1 benchmark
NinaPro DB2 (DB2 - 40 Intact Subjects - Delsys Trigno electrodes)
The second Ninapro database includes 40 intact subjects and it is thoroughly described in the paper: "Manfredo Atzori, Arjan Gijsberts, Claudio Castellini, Barbara Caputo, Anne-Gabrielle Mittaz Hager, Simone Elsig, Giorgio Giatsidis,…
3 papers · 0 benchmarks
OCTAGON (OCTAGON Dataset)
The OCTAGON dataset is a set of Angiography by Octical Coherence Tomography images (OCT-A) used to the segmentation of the Foveal Avascular Zone (FAZ).
3 papers · 0 benchmarks
ORVS (Online Retinal image for Vessel Segmentation (ORVS))
The ORVS dataset has been newly established as a collaboration between the computer science and visual-science departments at the University of Calgary.
3 papers · 0 benchmarks
PWDB (Pulse Wave Database)
Overview This database of simulated arterial pulse waves is designed to be representative of a sample of pulse waves measured from healthy adults.
3 papers · 0 benchmarks
Placenta is a benchmark dataset for node classification in an underexplored domain: predicting microanatomical tissue structures from cell graphs in placenta histology whole slide images.
3 papers · 1 benchmark
This prostate MRI segmentation dataset is collected from six different data sources.
3 papers · 0 benchmarks
SMILE-UHURA (Small Vessel Segmentation at Mesoscopic Scale from Ultra-High Resolution 7T Magnetic Resonance Angiogram)
The human brain receives nutrients and oxygen through an intricate network of blood vessels.
3 papers · 0 benchmarks
SkinCon (SKIN Concepts Dataset)
SkinCon is a skin disease dataset densely annotated by dermatologists.
3 papers · 0 benchmarks
SurgT is a dataset for benchmarking 2D Trackers in Minimally Invasive Surgery (MIS).
3 papers · 0 benchmarks
A vast amount of information in the biomedical domain is available as natural language free text.
3 papers · 0 benchmarks
The US-4 is a dataset of Ultrasound (US) images.
3 papers · 0 benchmarks
BioVid (BioVid Heat Pain Database)
To advance methods for pain assessment, in particular automatic assessment methods, the BioVid Heat Pain Database was collected in a collaboration of the Neuro-Information Technology group of the University of Magdeburg and the Medical…
2 papers · 0 benchmarks
The breast lesion detection in ultrasound videos dataset uses a clip-level and video-level feature aggregated network (CVA-Net) and consists of 188 ultrasound videos, of which 113 are labeled malignant and 75 benign.
2 papers · 0 benchmarks
BreastClassifications4 ([MIMBCD-UI] UTA4: Severity & Pathology Classifications Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
2 papers · 0 benchmarks
CENTER-TBI (Collaborative European NeuroTrauma Effectiveness Research in TBI)
The CENTER-TBI database contains prospectively collected data of more than 4,500 patients with TBI in Europe.
2 papers · 0 benchmarks
Annotated audio files (separate combined annotation file) of lung sounds as recorded from various vantage points of the chest wall.
2 papers · 1 benchmark
Set of landmark annotations for JSRT, Montgomery, Shenzhen and a subset of Padchest datasets
2 papers · 0 benchmarks
Colorectal Adenoma contains 177 whole slide images (156 contain adenoma) gathered and labelled by pathologists from the Department of Pathology, The Chinese PLA General Hospital.
2 papers · 0 benchmarks
This collection contains data and code associated with the IPCAI/IJCARS 2020 paper “Automatic Annotation of Hip Anatomy in Fluoroscopy for Robust and Efficient 2D/3D Registration.” The data hosted here consists of annotated datasets of…
2 papers · 0 benchmarks
DisKnE (Disease Knowledge Evaluation)
DisKnE is a benchmark for Disease Knowledge Evaluation built from MedNLI and MEDIQA-NLI.
2 papers · 0 benchmarks
Estimating camera motion in deformable scenes poses a complex and open research challenge.
2 papers · 1 benchmark
Background: Lung cancer risk classification is an increasingly important area of research as low-dose thoracic CT screening programs have become standard of care for patients at high risk for lung cancer.
2 papers · 1 benchmark
EPISURG (EPISURG: a dataset of postoperative MRI for quantitative analysis of resection neurosurgery for refractory epilepsy)
EPISURG is a clinical dataset of T1-weighted magnetic resonance images (MRI) from 430 epileptic patients who underwent resective brain surgery at the National Hospital of Neurology and Neurosurgery (Queen Square, London, United Kingdom)…
2 papers · 0 benchmarks
EchoCP is an echocardiography dataset in cTTE targeting PFO (Patent foramen ovale) diagnosis.
2 papers · 0 benchmarks
HCP Aging (Lifespan Human Connectome Project Aging)
Lifespan HCP Release 2.0 includes cross-sectional visit 1 (V1) preprocessed structural and functional imaging data, unprocessed V1 imaging data for all included modalities (structural, high-res hippocampal T2, resting state fMRI, task…
2 papers · 1 benchmark
The ISIC 2017 dataset was published by the International Skin Imaging Collaboration (ISIC) as a large-scale dataset of dermoscopy images.
2 papers · 0 benchmarks
ImDrug is a comprehensive benchmark with an open-source Python library which consists of 4 imbalance settings, 11 AI-ready datasets, 54 learning tasks and 16 baseline algorithms tailored for imbalanced learning.
2 papers · 0 benchmarks
Kvasir-Capsule dataset is the largest publicly released VCE dataset.
2 papers · 0 benchmarks
The LeukemiaAttri dataset is a large-scale, multi-domain collection of microscopy images derived from leukemia patient samples, enriched with detailed morphological information.
2 papers · 2 benchmarks
MCSCSet is a large-scale specialist-annotated dataset, designed for the task of Medical-domain Chinese Spelling Correction that contains about 200k samples.
2 papers · 0 benchmarks
MHSMA (The Modified Human Sperm Morphology Analysis)
The MHSMA dataset is a collection of human sperm images from 235 patients with male factor infertility.
2 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.