Home › Datasets › task › Segmentation

Segmentation datasets

archive 2025-07-28

45 datasets carry the task tag "Segmentation" (the task itself: Segmentation), ordered by the archive's paper count. Page 1 of 1: 45 shown of 45. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Segmentation datasets 1–45 of 45

SA-1B consists of 11M diverse, high resolution, licensed, and privacy protecting images and 1.1B high-quality segmentation masks.
183 papers · 1 benchmark
For the details of the work, the readers are refer to the paper "Feature Pyramid and Hierarchical Boosting Network for Pavement Crack Detection" (FPHB), T-ITS 2019.
34 papers · 0 benchmarks
UIIS (General Underwater Image Instance Segmentation dataset)
This is the first general Underwater Image Instance Segmentation (UIIS) dataset containing 4,628 images for 7 categories with pixel-level annotations for underwater instance segmentation task
9 papers · 1 benchmark
ROBUST-MIS (Robust Medical Instrument Segmentation Challenge 2019)
The ROBUST-MIS dataset was made available to support the Robust Medical Instrument Segmentation (ROBUST-MIS) Challenge 2019, part of the Endoscopic Vision Challenge associated with MICCAI.
8 papers · 1 benchmark
25kTrees (Individual Tree Crown Annotations)
Manual crown delineation of individual trees in two countries: Denmark and Finland.
5 papers · 0 benchmarks
BHSD (A 3D Multi-class Brain Hemorrhage Segmentation Dataset)
Intracranial hemorrhage (ICH) is a pathological condition characterized by bleeding inside the skull or brain, which can be attributed to various factors.
5 papers · 0 benchmarks
The dataset offers tag and mask annotations for image-text pairs from the CC3M validation set.
5 papers · 2 benchmarks
MMFlood is remote sensing dataset derived from Sentinel-1 (VV-VH), MapZen (DEM) and OpenStreetMap (Hydrography).
5 papers · 1 benchmark
CFD (CrackForestDataset)
CrackForest Dataset is an annotated road crack image database which can reflect urban road surface condition in general.
4 papers · 0 benchmarks
We present the CrackVision12k dataset, a collection of 12,000 crack images derived from 13 publicly available crack datasets.
4 papers · 1 benchmark
ECG-Image-Database (Digitization and Classification of ECG Images: The George B. Moody PhysioNet Challenge 2024)
The George B.
4 papers · 1 benchmark
From my knowledge, the dataset used in the project is the largest crack segmentation dataset so far.
4 papers · 2 benchmarks
LLM-Seg40K dataset contains 14K images in total.
4 papers · 0 benchmarks
PASTIS-R (Panoptic Segmentation of Radar and Optical Satellite image TIme Series)
Extension of the PASTIS benchmark with radar and optical image time series.
4 papers · 2 benchmarks
USIS10K (Large-scale Underwater Salient Instance Segmentation Dataset)
We construct the first large-scale dataset, USIS10K, for the underwater salient instance segmentation task, which contains 10,632 images and pixel-level annotations of 7 categories.
4 papers · 0 benchmarks
SMILE-UHURA (Small Vessel Segmentation at Mesoscopic Scale from Ultra-High Resolution 7T Magnetic Resonance Angiogram)
The human brain receives nutrients and oxygen through an intricate network of blood vessels.
3 papers · 0 benchmarks
UIIS10K (General Underwater Image Instance Segmentation dataset 10K)
We propose a large-scale underwater instance segmentation dataset, UIIS10K, which includes 10,048 images with pixel-level annotations for 10 categories.
3 papers · 0 benchmarks
The researchers of Qatar University have compiled the COVID-QU-Ex dataset, which consists of 33,920 chest X-ray (CXR) images including: 11,956 COVID-19 11,263 Non-COVID infections (Viral or Bacterial Pneumonia) 10,701 Normal Ground-truth…
2 papers · 0 benchmarks
CaBuAr (CaBuAr: California Burned Areas dataset)
This dataset contains images from Sentinel-2 satellites taken before and after a wildfire.
2 papers · 0 benchmarks
Pre-training is a strong strategy for enhancing visual models to efficiently train them with a limited number of labeled images.
2 papers · 0 benchmarks
The ULS23 test set contains 725 lesions from 284 patients of the Radboudumc and JBZ hospitals in the Netherlands.
2 papers · 1 benchmark
Underwater Trash Detection Dataset Overview The Underwater Trash Detection Dataset is a custom-annotated dataset designed to address the challenges of underwater trash detection caused by varying environmental features.
2 papers · 0 benchmarks
The AneuX morphology database includes data from 3 different data sources: AneuX, @neurIST and Aneurisk.
1 paper · 0 benchmarks
BRISC (BRISC: Annotated Dataset for Brain Tumor Segmentation and Classification)
BRISC is a high-quality, expert-annotated MRI dataset curated for brain tumor segmentation and classification.
1 paper · 1 benchmark
Boreal Forest Fire (Boreal Forest Fire: UAV-collected Wildfire Detection and Smoke Segmentation Dataset)
This dataset consists of annotated images and videos of smoke resulting from prescribed burning events in Finnish boreal forests.
1 paper · 0 benchmarks
RGB-D instance segmentation box dataset.
1 paper · 2 benchmarks
BraTS PEDs 2023 (The Brain Tumor Segmentation (BraTS) Challenge 2023: Focus on Pediatrics (CBTN-CONNECT-DIPGR-ASNR-MICCAI BraTS-PEDs))
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
This is the dataset released along with the publication: CUTS: A Deep Learning and Topological Framework for Multigranular Unsupervised Medical Image Segmentation [[ArXiv]](https://arxiv.org/abs/2209.11359)…
1 paper · 0 benchmarks
This dataset derives from Coil100.
1 paper · 0 benchmarks
There are 537 RGB jpg images of cracks and corresponding png binary segmentation masks of crack: a training set with 300 images and a testing set with 237 images.
1 paper · 0 benchmarks
Extended Task10_Colon Medical Decathlon (Extended Task10_Colon of Medical Segmentation Decathlon dataset)
A dataset of abdominal CT studies in NifTi format from the open-source medical data repository Medical Decathlon was utilized.
1 paper · 1 benchmark
FP4S (Floor plan image segmentation via scribble-based semi-weakly-supervised learning)
We introduce a new style- and category-agnostic floor plan image parsing benchmark developed in collaboration with professional architectural designers.
1 paper · 1 benchmark
This data set comprises 22 fundus images with their corresponding manual annotations for the blood vessels, separated as arteries and veins.
1 paper · 2 benchmarks
MFSD (Masked Face Segmentation Dataset)
During the covid-19 era wearing face masks posed new challenges to face-related tasks, including facial recognition, face inpainting, expression recognition, and object removal.
1 paper · 1 benchmark
Microscopy Image Dataset of Pulmonary Vascular Changes (Microscopy Image Dataset for Deep Learning-Based Quantitative Assessment of Pulmonary Vascular Changes)
Pulmonary hypertension (PH) is a syndrome complex that accompanies a number of diseases of different etiologies, associated with basic mechanisms of structural and functional changes of the pulmonary circulation vessels and revealed…
1 paper · 0 benchmarks
NPO (Negative and Positive Obstacles)
The dataset is recorded with an on-vehicle ZED stereo camera in both urban and rural environments The dataset contains various lighting conditions, such as normal lights, large-area shadows, dim lights, and sun glare.
1 paper · 1 benchmark
A RGB-D dataset converted from NYUDv2 into COCO-style instance segmentation format.
1 paper · 2 benchmarks
S-BIAD843 (Individual 3D cell shapes of Drosophila Wing Disc)
Late third instar wing imaginal discs were cultured in Shields and Sang M3 media (Sigma) supplemented with 2% FBS (Sigma), 1% pen/strep (Gibco), 3ng/ml ecdysone (Sigma) and 2ng/ml insulin (Sigma).
1 paper · 0 benchmarks
A RGB-D dataset converted from SUN-RGBD into COCO-style instance segmentation format.
1 paper · 2 benchmarks
SimGas (Computer Simulated Gas Leakage Segmentation)
This dataset consists of computer-generated images for gas leakage segmentation.
1 paper · 2 benchmarks
A fully synthetic dataset of drones generated using structured domain randomization.
1 paper · 0 benchmarks
​The Aachen-Heerlen annotated steel microstructure dataset comprises 1,705 scanning electron microscopy (SEM) images of bainitic steel samples.
0 papers · 0 benchmarks
Data in this study come from western Ecuador's Choco tropical forest, including \textit{Fundación para la Conservación de los Andes Tropicales Reserve and adjacent Reserva Ecológica Mache-Chindul park} (FCAT; 00°23'28'' N, 79°41'05'' W),…
0 papers · 0 benchmarks
Tornet (Tornado Network)
The Tornado Network (TorNet) dataset is a large, high-resolution benchmark dataset developed to support machine learning research in tornado detection and prediction.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.