Home › Datasets › task › Semantic Segmentation

Semantic Segmentation datasets

archive 2025-07-28

347 datasets carry the task tag "Semantic Segmentation" (the task itself: Semantic Segmentation), ordered by the archive's paper count. Page 4 of 8: 48 shown of 347. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Semantic Segmentation datasets 145–192 of 347

The TrashCan dataset is an instance-segmentation dataset of underwater trash.
12 papers · 0 benchmarks
WildScenes is a bi-modal benchmark dataset consisting of multiple large-scale, sequential traversals in natural environments, including semantic annotations in high-resolution 2D images and dense 3D LiDAR point clouds, and accurate 6-DoF…
12 papers · 2 benchmarks
LandCover.ai (Dataset for Automatic Mapping of Buildings, Woodlands, Water and Roads from Aerial Imagery)
The LandCover.ai (Land Cover from Aerial Imagery) dataset is a dataset for automatic mapping of buildings, woodlands, water and roads from aerial images.
11 papers · 1 benchmark
RailSem19 (RailSem19: A Dataset for Semantic Rail Scene Understanding)
RailSem19 offers 8500 unique images taken from a the ego-perspective of a rail vehicle (trains and trams).
11 papers · 0 benchmarks
The SD-198 dataset contains 198 different diseases from different types of eczema, acne and various cancerous conditions.
11 papers · 0 benchmarks
So2Sat LCZ42 consists of local climate zone (LCZ) labels of about half a million Sentinel-1 and Sentinel-2 image patches in 42 urban agglomerations (plus 10 additional smaller areas) across the globe.
11 papers · 1 benchmark
TACO is a growing image dataset of waste in the wild.
11 papers · 0 benchmarks
2-PM Vessel is an open-source volumetric brain vasculature dataset obtained with two-photon microscopy at Focused Ultrasound Lab, at Sunnybrook Research Institute (affiliated with University of Toronto by Dr.
10 papers · 0 benchmarks
The CropAndWeed dataset is focused on the fine-grained identification of 74 relevant crop and weed species with a strong emphasis on data variability.
10 papers · 0 benchmarks
The dataset was created using high-resolution (8 m) satellite imagery from the Gaofen series (Gaofen-2 and Gaofen-6), captured in 2019 over Maduo County, China, located in the Yellow River source area.
10 papers · 1 benchmark
The increasing incidence of melanoma has recently promoted the development of computer-aided diagnosis systems for the classification of dermoscopic images.
10 papers · 3 benchmarks
SpaceNet 2 (SpaceNet 2: Building Detection v2)
SpaceNet 2: Building Detection v2 - is a dataset for building footprint detection in geographically diverse settings from very high resolution satellite images.
10 papers · 1 benchmark
The AIRS (Aerial Imagery for Roof Segmentation) dataset provides a wide coverage of aerial imagery with 7.5 cm resolution and contains over 220,000 buildings.
9 papers · 1 benchmark
A high-resolution semantic segmentation dataset with 50 validation and 100 test objects.
9 papers · 1 benchmark
BIMCV-COVID19+ dataset is a large dataset with chest X-ray images CXR (CR, DX) and computed tomography (CT) imaging of COVID-19 patients along with their radiographic findings, pathologies, polymerase chain reaction (PCR), immunoglobulin G…
9 papers · 0 benchmarks
EgoHOS (Fine-Grained Egocentric Hand-Object Segmentation Dataset)
EgoHOS is a labeled dataset consisting of 11243 egocentric images with per-pixel segmentation labels of hands and objects being interacted with during a diverse array of daily activities.
9 papers · 0 benchmarks
The EntitySeg dataset contains 33,227 images with high-quality mask annotations.
9 papers · 0 benchmarks
PerSeg is a dataset for personalized segmentation.
9 papers · 1 benchmark
PhenoBench (PhenoBench — A Large Dataset and Benchmarks for Semantic Image Interpretation in the Agricultural Domain)
The PhenoBench dataset contains multiple image segmentation challenges from the agricultural domain.
9 papers · 0 benchmarks
SAMRS is a remote sensing segmentation dataset which provides object category, location, and instance information that can be used for semantic segmentation, instance segmentation, and object detection, either individually or in…
9 papers · 0 benchmarks
SketchyScene is a large-scale dataset of scene sketches to advance research on sketch understanding at both the object and scene level.
9 papers · 0 benchmarks
TSSB (Time Series Segmentation Benchmark)
The time series segmentation benchmark (TSSB) currently contains 75 annotated time series (TS) with 1-9 segments.
9 papers · 1 benchmark
UIIS (General Underwater Image Instance Segmentation dataset)
This is the first general Underwater Image Instance Segmentation (UIIS) dataset containing 4,628 images for 7 categories with pixel-level annotations for underwater instance segmentation task
9 papers · 1 benchmark
Are current 3D object tracking methods truely robust enough for low-fidelity depth sensors like the iPhone LiDAR?
8 papers · 2 benchmarks
Echocardiography, or cardiac ultrasound, is the most widely used and readily available imaging modality to assess cardiac function and structure.
8 papers · 0 benchmarks
X-ray images in this data set have been acquired from the tuberculosis control program of the Department of Health andHuman Services of Montgomery County, MD, USA.
8 papers · 1 benchmark
The RIT-18 dataset was built for the semantic segmentation of remote sensing imagery.
8 papers · 0 benchmarks
SynthCity is a 367.9M point synthetic full colour Mobile Laser Scanning point cloud.
8 papers · 0 benchmarks
Research on semantic segmentation of traffic scenes using color and polarization information (including training and testing sets).
8 papers · 1 benchmark
NDD20 (Northumberland Dolphin Dataset 2020)
Northumberland Dolphin Dataset 2020 (NDD20) is a challenging image dataset annotated for both coarse and fine-grained instance segmentation and categorisation.
7 papers · 0 benchmarks
OST300 is an outdoor scene dataset with 300 test images of outdoor scenes, and a training set of 7 categories of images with rich textures.
7 papers · 0 benchmarks
OpenEDS2020 is a dataset of eye-image sequences captured at a frame rate of 100 Hz under controlled illumination, using a virtual-reality head-mounted display mounted with two synchronized eye-facing cameras.
7 papers · 0 benchmarks
A database of images of approximately 960 unique plants belonging to 12 species at several growth stages is made publicly available.
7 papers · 0 benchmarks
The Sunnybrook Cardiac Data (SCD), also known as the 2009 Cardiac MR Left Ventricle Segmentation Challenge data, consist of 45 cine-MRI images from a mixed of patients and pathologies: healthy, hypertrophy, heart failure with infarction…
7 papers · 0 benchmarks
EarthVQA (A multi-modal multi-task VQA dataset for remote sensing)
Earth vision research typically focuses on extracting geospatial object locations and categories but neglects the exploration of relations between objects and comprehensive reasoning.
6 papers · 1 benchmark
The Freiburg Forest dataset was collected using a Viona autonomous mobile robot platform equipped with cameras for capturing multi-spectral and multi-modal images.
6 papers · 2 benchmarks
GFF (Global Flood Forecasting)
Floods are among the most common and devastating natural hazards, imposing immense costs on our society and economy due to their disastrous consequences.
6 papers · 1 benchmark
MatterportLayout extends the Matterport3D dataset with general Manhattan layout annotations.
6 papers · 0 benchmarks
Open Images is a computer vision dataset covering ~9 million images with labels spanning thousands of object categories.
6 papers · 0 benchmarks
SWIMSEG (Singapore Whole sky IMaging SEGmentation Database)
The SWIMSEG dataset contains 1013 images of sky/cloud patches, along with their corresponding binary segmentation maps.
6 papers · 1 benchmark
VDD (Varied Drone Dataset for Semantic Segmentation)
Semantic segmentation of drone images is critical for various aerial vision tasks as it provides essential seman- tic details to understand scenes on the ground.
6 papers · 1 benchmark
XImageNet-12 (XIMAGENET-12: An Explainable AI Benchmark Dataset for Model Robustness Evaluation)
Enlarge the dataset to understand how image background effect the Computer Vision ML model.
6 papers · 1 benchmark
25kTrees (Individual Tree Crown Annotations)
Manual crown delineation of individual trees in two countries: Denmark and Finland.
5 papers · 0 benchmarks
BUP20 (Sweet Pepper 2020 University of Bonn)
Video sequences from a glasshouse environment in Campus Kleinaltendorf(CKA), University of Bonn, captured by PATHoBot, a glasshouse monitoring robot.
5 papers · 0 benchmarks
BRATS 2014 is a brain tumor segmentation dataset.
5 papers · 1 benchmark
The dataset offers tag and mask annotations for image-text pairs from the CC3M validation set.
5 papers · 2 benchmarks
Cata7 is the first cataract surgical instrument dataset for semantic segmentation.
5 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.