Home › Datasets › task › Semantic Segmentation

Semantic Segmentation datasets

archive 2025-07-28

347 datasets carry the task tag "Semantic Segmentation" (the task itself: Semantic Segmentation), ordered by the archive's paper count. Page 3 of 8: 48 shown of 347. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Semantic Segmentation datasets 97–144 of 347

The ISIC 2018 dataset was published by the International Skin Imaging Collaboration (ISIC) as a large-scale dataset of dermoscopy images.
27 papers · 1 benchmark
ISPRS Potsdam (2D Semantic Labeling Contest - Potsdam)
The data set contains 38 patches (of the same size), each consisting of a true orthophoto (TOP) extracted from a larger TOP mosaic.
27 papers · 2 benchmarks
Nighttime Driving is a dataset of road scenes consisting of 35,000 images ranging from daytime to twilight time and to nighttime.
27 papers · 2 benchmarks
ScribbleSup (PASCAL-Scribble Dataset)
The PASCAL-Scribble Dataset is an extension of the PASCAL dataset with scribble annotations for semantic segmentation.
26 papers · 0 benchmarks
A three million frame, multi-view, furniture assembly video dataset that includes depth, atomic actions, object segmentation, and human pose.
25 papers · 1 benchmark
ROSE (Retinal OCTA SEgmentation dataset)
Retinal OCTA SEgmentation dataset (ROSE) consists of 229 OCTA images with vessel annotations at either centerline-level or pixel level.
25 papers · 4 benchmarks
DADA-seg is a pixel-wise annotated accident dataset, which contains a variety of critical scenarios from traffic accidents.
24 papers · 1 benchmark
Toronto-3D is a large-scale urban outdoor point cloud dataset acquired by an MLS system in Toronto, Canada for semantic segmentation.
24 papers · 2 benchmarks
DSEC (A Stereo Event Camera Dataset for Driving Scenarios)
DSEC is a stereo camera dataset in driving scenarios that contains data from two monochrome event cameras and two global shutter color cameras in favorable and challenging illumination conditions.
23 papers · 2 benchmarks
The DeepWeeds dataset consists of 17,509 images capturing eight different weed species native to Australia in situ with neighbouring flora.
23 papers · 0 benchmarks
A large-scale 4D egocentric dataset with rich annotations, to catalyze the research of category-level human-object interaction.
23 papers · 0 benchmarks
CARRADA is a dataset of synchronized camera and radar recordings with range-angle-Doppler annotations.
22 papers · 0 benchmarks
Satlas is a remote sensing dataset and benchmark that is large in both breadth, featuring all of the aforementioned applications and more, as well as scale, comprising 290M labels under 137 categories and 7 label modalities.
22 papers · 0 benchmarks
The INRIA Aerial Image Labeling dataset is comprised of 360 RGB tiles of 5000×5000px with a spatial resolution of 30cm/px on 10 cities across the globe.
21 papers · 1 benchmark
A composite dataset that unifies semantic segmentation datasets from different domains.
21 papers · 0 benchmarks
AVSBench (Audio −Visual Segmentation)
AVSBench is a pixel-level audio-visual segmentation benchmark that provides ground truth labels for sounding objects.
20 papers · 0 benchmarks
FMB Dataset (Full-time Multi-modality Benchmark Dataset)
FMB contains 1500 well-registered infrared and visible image pairs with 14 annotated pixel-level categories.
20 papers · 1 benchmark
PASCAL VOC 2011 is an image segmentation dataset.
20 papers · 2 benchmarks
DELIVER is an arbitrary-modal segmentation benchmark, covering Depth, LiDAR, multiple Views, Events, and RGB.
19 papers · 3 benchmarks
DeepFish as a benchmark suite with a large-scale dataset to train and test methods for several computer vision tasks.
19 papers · 1 benchmark
FoodSeg103 (lewisnjue)
FoodSeg103 is a new food image dataset containing 7,118 images.
19 papers · 1 benchmark
Hi4D contains 4D textured scans of 20 subject pairs, 100 sequences, and a total of more than 11K frames.
19 papers · 0 benchmarks
ISPRS Vaihingen (2D Semantic Labeling - Vaihingen data)
The data set contains 33 patches (of different sizes), each consisting of a true orthophoto (TOP) extracted from a larger TOP mosaic.
19 papers · 1 benchmark
Consists of annotated frames containing GI procedure tools such as snares, balloons and biopsy forceps, etc.
19 papers · 3 benchmarks
MCubeS (Multimodal Material Segmentation Dataset)
Multimodal material segmentation (MCubeS) dataset contains 500 sets of images from 42 street scenes.
19 papers · 1 benchmark
ModaNet is a street fashion images dataset consisting of annotations related to RGB images.
19 papers · 1 benchmark
AI-TOD (Tiny Object Detection in Aerial Images)
AI-TOD comes with 700,621 object instances for eight categories across 28,036 aerial images.
18 papers · 2 benchmarks
A large-scale aerial farmland image dataset for semantic segmentation of agricultural patterns.
18 papers · 0 benchmarks
CrossMoDA (Cross-Modality Domain Adaptation)
CrossMoDA is a large and multi-class benchmark for unsupervised cross-modality Domain Adaptation.
18 papers · 0 benchmarks
PASTIS (Panoptic Segmentation of satellite image TImes Series)
PASTIS is a benchmark dataset for panoptic and semantic segmentation of agricultural parcels from satellite image time series.
18 papers · 2 benchmarks
The dataset for this challenge was obtained by carefully annotating tissue images of several patients with tumors of different organs and who were diagnosed at multiple hospitals.
17 papers · 2 benchmarks
CryoNuSeg is a fully annotated FS-derived cryosectioned and H&E-stained nuclei instance segmentation dataset.
16 papers · 0 benchmarks
DADA-2000 is a large-scale benchmark with 2000 video sequences (named as DADA-2000) is contributed with laborious annotation for driver attention (fixation, saccade, focusing time), accident objects/intervals, as well as the accident…
16 papers · 0 benchmarks
The database consists of 150 annotated pages of three different medieval manuscripts with challenging layouts.
15 papers · 2 benchmarks
The Paris-Lille-3D is a Benchmark on Point Cloud Classification.
15 papers · 1 benchmark
RoadAnomaly21 is a dataset for anomaly segmentation, the task of identify the image regions containing objects that have never been seen during training.
15 papers · 0 benchmarks
Fashionpedia consists of two parts: (1) an ontology built by fashion experts containing 27 main apparel categories, 19 apparel parts, 294 fine-grained attributes and their relationships; (2) a dataset with everyday and celebrity event…
14 papers · 0 benchmarks
MHP (Multiple-Human Parsing)
The MHP dataset contains multiple persons captured in real-world scenes with pixel-level fine-grained semantic annotations in an instance-aware setting.
14 papers · 3 benchmarks
REFUGE Challenge (Retinal Fundus Glaucoma Challenge)
REFUGE Challenge provides a data set of 1200 fundus images with ground truth segmentations and clinical glaucoma labels, currently the largest existing one.
14 papers · 4 benchmarks
BRATS 2016 is a brain tumor segmentation dataset.
13 papers · 0 benchmarks
Detecting vehicles and representing their position and orientation in the three dimensional space is a key technology for autonomous driving.
13 papers · 3 benchmarks
The Habitat-Matterport 3D Semantics Dataset (HM3DSem) is the largest-ever dataset of 3D real-world and indoor spaces with densely annotated semantics that is available to the academic community.
13 papers · 0 benchmarks
TTPLA (Transmission Towers and Power Lines (TTPLA))
TTPLA is a public dataset which is a collection of aerial images on Transmission Towers (TTs) and Power Lines (PLs).
13 papers · 0 benchmarks
A Multi-Task 4D Radar-Camera Fusion Dataset for Autonomous Driving on Water Surfaces description of the dataset WaterScenes, the first multi-task 4D radar-camera fusion dataset on water surfaces, which offers data from multiple sensors,…
13 papers · 2 benchmarks
Partial iLIDS is a dataset for occluded person person re-identification.
12 papers · 0 benchmarks
SemanticSTF is an adverse-weather point cloud dataset that provides dense point-level annotations and allows to study 3DSS under various adverse weather conditions.
12 papers · 1 benchmark

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.