Home › Datasets › modality › Biomedical

Biomedical datasets

archive 2025-07-28

122 datasets carry the modality tag "Biomedical", ordered by the archive's paper count. Page 3 of 3: 26 shown of 122. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Biomedical datasets 97–122 of 122

MediConfusion is a challenging medical Visual Question Answering (VQA) benchmark dataset, that probes the failure modes of medical Multimodal Large Language Models (MLLMs) from a vision perspective.
1 paper · 0 benchmarks
Microscopy Image Dataset of Pulmonary Vascular Changes (Microscopy Image Dataset for Deep Learning-Based Quantitative Assessment of Pulmonary Vascular Changes)
Pulmonary hypertension (PH) is a syndrome complex that accompanies a number of diseases of different etiologies, associated with basic mechanisms of structural and functional changes of the pulmonary circulation vessels and revealed…
1 paper · 0 benchmarks
MuCeD, a dataset that is carefully curated and validated by expert pathologists from the All India Institute of Medical Science (AIIMS), Delhi, India.
1 paper · 0 benchmarks
Multi-template MRI mouse brain atlas (Multi-template MRI mouse brain atlas for both in vivo and ex vivo analysis)
Mouse Brain MRI atlas (both in-vivo and ex-vivo) (repository relocated from the original webpage) List of atlases - FVBNCrl: Brain MRI atlas of the wild-type FVBNCrl mouse strain (used as the background strain for the rTg4510 which is a…
1 paper · 0 benchmarks
RAOS (Rethinking Abdominal Organ Segmentation)
Rethinking Abdominal Organ Segmentation (RAOS) in the clinical scenario: A robustness evaluation benchmark with challenging cases.
1 paper · 0 benchmarks
A total of 227 cross sectional images (20 x 54 mm with a resolution of 289 x 648 pixels) of hind-leg xenograft tumors from 29 mice were obtained with 1mm step-wise movement of the array mounted on a manual positioning device.
1 paper · 0 benchmarks
SCARED-C (SCARED-Corrupted)
The dataset SCARED-C is introduced in the context of assessing robustness in endoscopic depth prediction models.
1 paper · 1 benchmark
Confocal fluorescence microscopy is one of the most accessible and widely used imaging techniques for the study of biological processes at the cellular and subcellular levels.
1 paper · 0 benchmarks
SinGAN-Seg-polyps is a synthetic dataset for polyp segmentation consisting of 10,000 synthetic polyps and masks.
1 paper · 0 benchmarks
A public open dataset of synthetic chest X-ray images of COVID-19.
1 paper · 0 benchmarks
TUAC (Temple University Artifact Corpus)
A new subset of the popular open source electroencephalogram (EEG) corpus – TUH EEG: - The Temple University Artifact Corpus (TUAR) consists of high yield artifact files annotated using a five-way classification system: 1.
1 paper · 0 benchmarks
The EMBO SourceData-NLP dataset (The SourceData-NLP dataset: integrating curation into scientific publishing for training large language models)
We present the SourceData-NLP dataset produced through the routine curation of papers during the publication process.
1 paper · 1 benchmark
Wearanize+ includes overnight sleep data from 130 participants (one night each) using three different wearable devices: Zmax headband, Empatica E4 wristband, and ActivPAL leg patch, alongside full-scale PSG recorded with SomnoScreen Plus…
1 paper · 0 benchmarks
uBench (MicroBench)
Microscopy is a cornerstone of biomedical research, enabling detailed study of biological structures at multiple scales.
1 paper · 0 benchmarks
Objective This study introduces the BlendedICU dataset, a massive dataset of international intensive care data.
0 papers · 0 benchmarks
Colorectal-Liver-Metastases (Colorectal-Liver-Metastases | Preoperative CT and Survival Data for Patients Undergoing Resection of Colorectal Liver Metastases)
This collection consists of DICOM images and DICOM Segmentation Objects (DSOs) for 197 patients with Colorectal Liver Metastases (CRLM).
0 papers · 0 benchmarks
DREAMING Inpainting Dataset (Diminished Reality for Emerging Applications in Medicine through Inpainting Dataset)
Dataset for the DREAMING - Diminished Reality for Emerging Applications in Medicine through Inpainting Challenge!
0 papers · 0 benchmarks
FHRMA dataset for FS detection (FHRMA dataset for fetal heart rate false signal detection)
FHRMA is an open-source project for Fetal Heart Rate Morphological Analysis containing Matlab source code and datasets.
0 papers · 0 benchmarks
Genome-wide miRNA detection (Genome-wide hairpins datasets of animals and plants for novel miRNA prediction)
We've made available several genome-wide datasets, which can be used for training microRNA (miRNA) classifiers.
0 papers · 0 benchmarks
The medaka (Oryzias latipes) and the zebrafish (Danio rerio) are used as a model organism for a variety of subjects in biomedical research.
0 papers · 0 benchmarks
InfiniteRep is a synthetic, open-source dataset for fitness and physical therapy (PT) applications.
0 papers · 0 benchmarks
Moh (SeyedMohammad Kashani)
We introduce an open-source physical-layer dataset of Bluetooth Low Energy (BLE) IoT sensor devices recorded in an anechoic chamber using USRP x310.
0 papers · 0 benchmarks
RSPECT (The RSNA Pulmonary Embolism CT)
The RSNA Pulmonary Embolism CT (RSPECT) Dataset is composed of CT pulmonary angiogram images and annotations related to pulmonary embolism.
0 papers · 0 benchmarks
SourceData-NLP (The SourceData-NLP dataset: integrating curation into scientific publishing for training large language models)
Introduction: The scientific publishing landscape is expanding rapidly, creating challenges for researchers to stay up-to-date with the evolution of the literature.
0 papers · 0 benchmarks
Human activity recognition and clinical biomechanics are challenging problems in physical telerehabilitation medicine.
0 papers · 0 benchmarks
Volumetric CMR Cartesian Datasets (Free-running self-gated 3D cine, 4D Flow and stress 4D Flow Undersampled Datasets)
Datasets at https://zenodo.org/record/8105485 for Motion Robust CMR Reconstruction Code in https://github.com/syedmurtazaarshad/motion-robust-CMR
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.