Home › Datasets › task › Semantic Segmentation

Semantic Segmentation datasets

archive 2025-07-28

347 datasets carry the task tag "Semantic Segmentation" (the task itself: Semantic Segmentation), ordered by the archive's paper count. Page 2 of 8: 48 shown of 347. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Semantic Segmentation datasets 49–96 of 347

Structured3D is a large-scale photo-realistic dataset containing 3.5K house designs (a) created by professional designers with a variety of ground truth 3D structure annotations (b) and generate photo-realistic 2D images (c).
85 papers · 7 benchmarks
The PROMISE12 dataset was made available for the MICCAI 2012 prostate segmentation challenge.
84 papers · 2 benchmarks
LoveDA (Remote Sensing Land-Cover Dataset for Domain Adaptive Semantic Segmentation)
1.
81 papers · 1 benchmark
iSAID contains 655,451 object instances for 15 categories across 2,806 high-resolution images.
81 papers · 4 benchmarks
The ReferIt dataset contains 130,525 expressions for referring to 96,654 objects in 19,894 images of natural scenes.
80 papers · 0 benchmarks
ApolloScape is a large dataset consisting of over 140,000 video frames (73 street scene videos) from various locations in China under varying weather conditions.
74 papers · 4 benchmarks
The SemanticPOSS dataset for 3D semantic segmentation contains 2988 various and complicated LiDAR scans with large quantity of dynamic instances.
71 papers · 1 benchmark
The BraTS 2015 dataset is a dataset for brain tumor image segmentation.
69 papers · 1 benchmark
CoNSeP (Colorectal Nuclear Segmentation and Phenotypes)
The colorectal nuclear segmentation and phenotypes (CoNSeP) dataset consists of 41 H&E stained image tiles, each of size 1,000×1,000 pixels at 40× objective magnification.
68 papers · 2 benchmarks
Semantic3D is a point cloud dataset of scanned outdoor scenes with over 3 billion points.
66 papers · 1 benchmark
The Places365 dataset is a scene recognition dataset.
65 papers · 7 benchmarks
LIP (Look into Person)
The LIP (Look into Person) dataset is a large-scale dataset focusing on semantic understanding of a person.
61 papers · 1 benchmark
PanNuke is a semi automatically generated nuclei instance segmentation and classification dataset with exhaustive nuclei labels across 19 different tissue types.
61 papers · 4 benchmarks
Dark Zurich is an image dataset containing a total of 8779 images captured at nighttime, twilight, and daytime, along with the respective GPS coordinates of the camera for each image.
57 papers · 3 benchmarks
Lost and Found is a novel lost-cargo image sequence dataset comprising more than two thousand frames with pixelwise annotations of obstacle and free-space and provide a thorough comparison to several stereo-based baseline methods.
57 papers · 1 benchmark
RELLIS-3D is a multi-modal dataset for off-road robotics.
57 papers · 3 benchmarks
Fisheye cameras are commonly employed for obtaining a large field of view in surveillance, augmented reality and in particular automotive applications.
55 papers · 1 benchmark
UAVid is a high-resolution UAV semantic segmentation dataset as a complement, which brings new challenges, including large scale variation, moving object recognition and temporal consistency preservation.
54 papers · 2 benchmarks
Virtual KITTI 2 is an updated version of the well-known Virtual KITTI dataset which consists of 5 sequence clones from the KITTI tracking benchmark.
53 papers · 2 benchmarks
The xBD dataset contains over 45,000KM2 of polygon labeled pre and post disaster imagery.
51 papers · 2 benchmarks
3D-FUTURE (3D FUrniture shape with TextURE) is a 3D dataset that contains 20,240 photo-realistic synthetic images captured in 5,000 diverse scenes, and 9,992 involved unique industrial 3D CAD shapes of furniture with high-resolution…
48 papers · 0 benchmarks
WildDash is a benchmark evaluation method is presented that uses the meta-information to calculate the robustness of a given algorithm with respect to the individual hazards.
47 papers · 2 benchmarks
Stanford Background (Standford Background Dataset)
The Stanford Background dataset contains 715 RGB images and the corresponding label images.
46 papers · 0 benchmarks
Synscapes is a synthetic dataset for street scene parsing created using photorealistic rendering techniques, and show state-of-the-art results for training and validation as well as new types of analysis.
46 papers · 1 benchmark
A novel dataset and benchmark, which features 1482 RGB-D scans of 478 environments across multiple time steps.
44 papers · 4 benchmarks
DensePASS - a novel densely annotated dataset for panoramic segmentation under cross-domain conditions, specifically built to study the Pinhole-to-Panoramic transfer and accompanied with pinhole camera training examples obtained from…
43 papers · 1 benchmark
ImageNet-S (ImageNet Semantic Segmentation)
Powered by the ImageNet dataset, unsupervised learning on large-scale data has made significant advances for classification tasks.
43 papers · 6 benchmarks
DDD17 (DAVIS Driving Dataset 2017)
DDD17 has over 12 h of a 346x260 pixel DAVIS sensor recording highway and city driving in daytime, evening, night, dry and wet weather conditions, along with vehicle speed, GPS position, driver steering, throttle, and brake captured from…
41 papers · 1 benchmark
KITTI Road is road and lane estimation benchmark that consists of 289 training and 290 test images.
41 papers · 0 benchmarks
RoadTracer is a dataset for extraction of road networks from aerial images.
39 papers · 0 benchmarks
Slakh2100 (Synthesized Lakh Dataset)
The Synthesized Lakh (Slakh) Dataset is a dataset for audio source separation that is synthesized from the Lakh MIDI Dataset v0.1 using professional-grade sample-based virtual instruments.
38 papers · 3 benchmarks
WORD (Whole abdominal Organs Dataset)
WORD is a dataset for organ semantic segmentation that contains 150 abdominal CT volumes (30,495 slices) and each volume has 16 organs with fine pixel-level annotations and scribble-based sparse annotation, which may be the largest dataset…
38 papers · 0 benchmarks
Open Images V4 offers large scale across several dimensions: 30.1M image-level labels for 19.8k concepts, 15.4M bounding boxes for 600 object classes, and 375k visual relationship annotations involving 57 classes.
37 papers · 1 benchmark
Our project (STPLS3D) aims to provide a large-scale aerial photogrammetry dataset with synthetic and real annotated 3D point clouds for semantic and instance segmentation tasks.
36 papers · 3 benchmarks
PST900 is a dataset of 894 synchronized and calibrated RGB and Thermal image pairs with per pixel human annotations across four distinct classes from the DARPA Subterranean Challenge.
35 papers · 1 benchmark
OxUva is a dataset and benchmark for evaluating single-object tracking algorithms.
34 papers · 0 benchmarks
A dataset consisting of 180,662 triplets of dual-pol synthetic aperture radar (SAR) image patches, multi-spectral Sentinel-2 image patches, and MODIS land cover maps.
34 papers · 0 benchmarks
SUIM (Segmentation of Underwater IMagery)
The Segmentation of Underwater IMagery (SUIM) dataset contains over 1500 images with pixel annotations for eight object categories: fish (vertebrates), reefs (invertebrates), aquatic plants, wrecks/ruins, human divers, robots, and…
34 papers · 2 benchmarks
A large-scale dataset for transparent object segmentation, named Trans10K, consisting of 10,428 images of real scenarios with carefully manual annotations, which are 10 times larger than the existing datasets.
33 papers · 1 benchmark
OpenEDS (Open Eye Dataset) is a large scale data set of eye-images captured using a virtual-reality (VR) head mounted display mounted with two synchronized eyefacing cameras at a frame rate of 200 Hz under controlled illumination.
31 papers · 1 benchmark
PhraseCut is a dataset consisting of 77,262 images and 345,486 phrase-region pairs.
31 papers · 1 benchmark
DeepFashion2 is a versatile benchmark of four tasks including clothes detection, pose estimation, segmentation, and retrieval.
30 papers · 0 benchmarks
InteriorNet is a RGB-D for large scale interior scene understanding and mapping.
30 papers · 0 benchmarks
OCID (Object Clutter Indoor Dataset)
Developing robot perception systems for handling objects in the real-world requires computer vision algorithms to be carefully scrutinized with respect to the expected operating domain.
29 papers · 1 benchmark
PartImageNet is a large, high-quality dataset with part segmentation annotations.
29 papers · 0 benchmarks
SegTHOR (Segmentation of THoracic Organs at Risk)
SegTHOR (Segmentation of THoracic Organs at Risk) is a dataset dedicated to the segmentation of organs at risk (OARs) in the thorax, i.e.
28 papers · 0 benchmarks
The SensatUrbat dataset is an urban-scale photogrammetric point cloud dataset with nearly three billion richly annotated points, which is five times the number of labeled points than the existing largest point cloud dataset.
28 papers · 1 benchmark
GID (Gaofen Image Dataset)
Gaofen Image Dataset (GID) is a large-scale land-cover dataset constructed with Gaofen-2 (GF-2) satellite images.
27 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.