Home › Datasets › modality › Biomedical

Biomedical datasets

archive 2025-07-28

122 datasets carry the modality tag "Biomedical", ordered by the archive's paper count. Page 2 of 3: 48 shown of 122. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Biomedical datasets 49–96 of 122

The “Medico automatic polyp segmentation challenge” aims to develop computer-aided diagnosis systems for automatic polyp segmentation to detect all types of polyps (for example, irregular polyp, smaller or flat polyps) with high efficiency…
3 papers · 1 benchmark
NinaPro DB2 (DB2 - 40 Intact Subjects - Delsys Trigno electrodes)
The second Ninapro database includes 40 intact subjects and it is thoroughly described in the paper: "Manfredo Atzori, Arjan Gijsberts, Claudio Castellini, Barbara Caputo, Anne-Gabrielle Mittaz Hager, Simone Elsig, Giorgio Giatsidis,…
3 papers · 0 benchmarks
PETRAW (PEg TRAnsfer Workflow recognition by different modalities)
PETRAW data set was composed of 150 sequences of peg transfer training sessions.
3 papers · 6 benchmarks
PWDB (Pulse Wave Database)
Overview This database of simulated arterial pulse waves is designed to be representative of a sample of pulse waves measured from healthy adults.
3 papers · 0 benchmarks
Placenta is a benchmark dataset for node classification in an underexplored domain: predicting microanatomical tissue structures from cell graphs in placenta histology whole slide images.
3 papers · 1 benchmark
PubMedCite is a domain-specific dataset with about 192K biomedical scientific papers and a large citation graph preserving 917K citation relationships between them.
3 papers · 0 benchmarks
fluocells (Fluorescent Neuronal Cells)
By releasing this dataset, we aim at providing a new testbed for computer vision techniques using Deep Learning.
3 papers · 0 benchmarks
BC7 NLM-Chem (BioCreative VII NLM-Chem)
Full-text chemical identification and indexing in PubMed articles.
2 papers · 3 benchmarks
BioVid (BioVid Heat Pain Database)
To advance methods for pain assessment, in particular automatic assessment methods, the BioVid Heat Pain Database was collected in a collaboration of the Neuro-Information Technology group of the University of Magdeburg and the Medical…
2 papers · 0 benchmarks
BreastClassifications4 ([MIMBCD-UI] UTA4: Severity & Pathology Classifications Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
2 papers · 0 benchmarks
CRC100K (100,000 histological images of human colorectal cancer and healthy tissue)
This is a set of 100,000 non-overlapping image patches from hematoxylin & eosin (H&E) stained histological images of human colorectal cancer (CRC) and normal tissue.
2 papers · 0 benchmarks
This collection contains data and code associated with the IPCAI/IJCARS 2020 paper “Automatic Annotation of Hip Anatomy in Fluoroscopy for Robust and Efficient 2D/3D Registration.” The data hosted here consists of annotated datasets of…
2 papers · 0 benchmarks
Background: Lung cancer risk classification is an increasingly important area of research as low-dose thoracic CT screening programs have become standard of care for patients at high risk for lung cancer.
2 papers · 1 benchmark
Fetoscopic Placental Vessel Segmentation and Registration (FetReg2021) challenge was organized as part of the MICCAI2021 Endoscopic Vision (EndoVis) challenge.
2 papers · 0 benchmarks
Kvasir-Capsule dataset is the largest publicly released VCE dataset.
2 papers · 0 benchmarks
The LeukemiaAttri dataset is a large-scale, multi-domain collection of microscopy images derived from leukemia patient samples, enriched with detailed morphological information.
2 papers · 2 benchmarks
The MIMIC PERform Testing dataset contains the following physiological signals recorded from 200 critically-ill patients during routine clinical care: - electrocardiogram (ECG) - photoplethysmogram (PPG) - impedance pneumography (imp),…
2 papers · 2 benchmarks
the MTHS dataset contains 30Hz PPG signals obtained from 62 patients, including 35 men and 27 women.
2 papers · 2 benchmarks
MedMNIST-C is an open-source data set collection comprising algorithmically generated corruptions applied to the test sets of the MedMNIST collection following the concept of ImageNet-C.
2 papers · 0 benchmarks
Abstract The Norwegian Endurance Athlete ECG Database contains 12-lead ECG recordings from 28 elite athletes from various sports in Norway.
2 papers · 0 benchmarks
PAX-Ray++ (Projected Anatomy in X-Ray Dataset ++)
The PAX-Ray++ dataset uses pseudo-labeled thorax CTs to enable the segmentation of anatomy in Chest X-Rays.
2 papers · 0 benchmarks
To take advantage of the ever-increasing amount of structural data now available, we also trained Paragraph on a larger dataset.
2 papers · 1 benchmark
Electrophysiological data from implanted electrodes in the human brain are rare, and therefore scientific access to it has remained somewhat exclusive.
2 papers · 1 benchmark
Tc1 Mouse cerebellum atlas (Tc1 Mouse cerebellum atlas with Purkinje layer segmentation)
This mouse cerebellar atlas can be used for mouse cerebellar morphometry.
2 papers · 0 benchmarks
WildPPG (WildPPG: A Real-World PPG Dataset of Long Continuous Recordings)
a dataset of multi-modal signals from wearable devices at four sites on the body.
2 papers · 1 benchmark
The datasets used and analysed from the glucose clamp study are available in this DIF file.
1 paper · 0 benchmarks
The datasets used and analysed from the glucose clamp study are available in this Excel file.
1 paper · 0 benchmarks
A Dataset for Relation Extraction of Natural-Products (A curated evaluation dataset for end-to-end Relation Extraction of relationships between organisms and natural-products)
A curated evaluation dataset for end-to-end Relation Extraction of relationships between organisms and natural-products.
1 paper · 0 benchmarks
ACCT Data Repository (ACCT is a fast and accessible automatic cell counting tool using machine learning for 2D image segmentation)
This dataset is a collection of fluorescent images from mice in order to test an automatic cell counting tool that we developed.
1 paper · 0 benchmarks
ATC-GRAPH is the most extensive ATC benchmark dataset.
1 paper · 1 benchmark
BCSS (Breast Cancer Semantic Segmentation)
The BCSS dataset contains over 20,000 segmentation annotations of tissue regions from breast cancer images from The Cancer Genome Atlas (TCGA).
1 paper · 0 benchmarks
BioLeaflets is a biomedical dataset for Data2Text generation.
1 paper · 0 benchmarks
BreastDICOM4 ([MIMBCD-UI] UTA4: Medical Imaging DICOM Files Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
1 paper · 1 benchmark
BreastRates4 ([MIMBCD-UI] UTA4: Rates Dataset)
Several datasets are fostering innovation in higher-level functions for everyone, everywhere.
1 paper · 0 benchmarks
CoCaHis (Colon Cancer Histology Dataset)
Highlights • Publicly available dataset with 82 H&E stained images of frozen sections.
1 paper · 0 benchmarks
DARai (Daily Activity Recordings for AI and ML applications)
Daily Activity Recordings for Artificial Intelligence (DARai, pronounced "Dahr-ree") is a multimodal, hierarchically annotated dataset constructed to understand human activities in real-world settings.
1 paper · 0 benchmarks
The data presented here was extracted from a larger dataset collected through a collaboration between the Embedded Systems Laboratory (ESL) of the Swiss Federal Institute of Technology in Lausanne (EPFL), Switzerland and the Institute of…
1 paper · 0 benchmarks
This is the supplemental data for our paper on how to benchmark registrations of serial sections with ground truths.
1 paper · 0 benchmarks
The dataset X of this work is an extension of the heartSeg dataset.
1 paper · 1 benchmark
The Fraunhofer Portugal AICOS EDoF Dataset was produced within the TAMI project and is composed of images of microscopic fields of view (FOV) of Liquid-based Cervical Cytology (LBC) samples.
1 paper · 0 benchmarks
- The dataset contains full-spectral autofluorescence lifetime microscopic images (FS-FLIM) acquired on unstained ex-vivo human lung tissue, where 100 4D hypercubes of 256x256 (spatial resolution) x 32 (time bins) x 512 (spectral channels…
1 paper · 0 benchmarks
The National Health and Nutrition Examination Survey (NHANES) provides data on the health and environmental exposure of the non-institutionalized US population.
1 paper · 0 benchmarks
Prediction of a speaker's height is of interest in fields such as voice forensics, surveillance, and automatic speaker profiling.
1 paper · 0 benchmarks
HuTu 80 (HuTu 80 cell populations)
The image set contains 180 high-resolution color microscopic images of human duodenum adenocarcinoma HuTu 80 cell populations obtained in an in vitro scratch assay (for the details of the experimental protocol, we refer to (Liang et al.,…
1 paper · 1 benchmark
Liver-US (Liver Ultrasound Dataset for Medical Image Classification)
The Liver-US dataset is a comprehensive collection of high-quality ultrasound images of the liver, including both normal and abnormal cases.
1 paper · 1 benchmark
This dataset contains pre-processed versions of datasets introduced in prior works.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.