Home › Datasets › task › Semantic Segmentation
Semantic Segmentation datasets
archive 2025-07-28
347 datasets carry the task tag "Semantic Segmentation" (the task itself: Semantic Segmentation), ordered by the archive's paper count. Page 6 of 8: 48 shown of 347. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Semantic Segmentation datasets 241–288 of 347
OSAI introduces OpenTTGames - an open dataset aimed at evaluation of different computer vision tasks in Table Tennis: ball detection, semantic segmentation of humans, table and scoreboard and fast in-game events spotting.
3 papers · 0 benchmarks
PETRAW (PEg TRAnsfer Workflow recognition by different modalities)
PETRAW data set was composed of 150 sequences of peg transfer training sessions.
3 papers · 6 benchmarks
PIC (Person In Context 2021)
The Person In Context (PIC) dataset is a dataset for human-centric relation segmentation (HRS), which contains 17,122 high-resolution images and densely annotated entity segmentation and relations, including 141 object categories, 23…
3 papers · 0 benchmarks
RobotPush is a dataset for object singulation – the task of separating cluttered objects through physical interaction.
3 papers · 0 benchmarks
A dataset with high resolution (4K) images and manually-annotated dense labels every 50 frames.
3 papers · 0 benchmarks
An open source Multi-View Overhead Imagery dataset with 27 unique looks from a broad range of viewing angles (-32.5 degrees to 54.0 degrees).
3 papers · 0 benchmarks
UNDD (Urban Night Driving Dataset)
UNDD consists of 7125 unlabelled day and night images; additionally, it has 75 night images with pixel-level annotations having classes equivalent to Cityscapes dataset.
3 papers · 0 benchmarks
The Vocal Folds dataset is a dataset for automatic segmentation of laryngeal endoscopic images.
3 papers · 0 benchmarks
ATLANTIS is a benchmark for semantic segmentation of waterbody images.
2 papers · 1 benchmark
CalCROP21 is a georeferenced multi-spectral dataset of satellite Imagery and crop labels.
2 papers · 0 benchmarks
The dataset contains two subsets of synthetic, semantically segmented road-scene images, which have been created for developing and applying the methodology described in the paper "A Sim2Real Deep Learning Approach for the Transformation…
2 papers · 2 benchmarks
The training and validation data are subsets of the training split of the Cityscapes dataset.
2 papers · 1 benchmark
Colorectal Adenoma contains 177 whole slide images (156 contain adenoma) gathered and labelled by pathologists from the Department of Pathology, The Chinese PLA General Hospital.
2 papers · 0 benchmarks
A benchmark dataset for training and evaluating global cloud classification models.
2 papers · 0 benchmarks
In EMDS-6, there are 21 classes of environmental microorganisms (EMs).
2 papers · 0 benchmarks
A challenge that consists of three tasks, each targeting a different requirement for in-clinic use.
2 papers · 1 benchmark
Extended Agriculture-Vision dataset comprises two parts: 1.
2 papers · 0 benchmarks
FPDS (Fallen People Data Set)
A benchmark for detecting fallen people lying on the floor.
2 papers · 0 benchmarks
This dataset is made up of forward-looking sonar images containing ten classes of underwater debris.
2 papers · 1 benchmark
FracAtlas (A Dataset for Fracture Classification, Localization and Segmentation of Musculoskeletal Radiographs)
FractureAtlas is a musculoskeletal bone fracture dataset with annotations for deep learning tasks like classification, localization, and segmentation.
2 papers · 0 benchmarks
G-VUE (General-purpose Visual Understanding Evaluation)
General-purpose Visual Understanding Evaluation (G-VUE) is a comprehensive benchmark covering the full spectrum of visual cognitive abilities with four functional domains -- Perceive, Ground, Reason, and Act.
2 papers · 0 benchmarks
This dataset contains simulated and expert-labelled spectrograms from two radio telescopes: the Hydrogen Epoch of Reionization Array (HERA) in South Africa and the Low-Frequency Array (LOFAR) in the Netherlands.
2 papers · 1 benchmark
HuTics (Human Deictic Gestures Dataset)
HuTics contains 2040 images showing how humans use deictic gestures to interact with various daily-life objects.
2 papers · 0 benchmarks
This dataset contains simulated and expert-labelled spectrograms from two radio telescopes: the Hydrogen Epoch of Reionization Array (HERA) in South Africa and the Low-Frequency Array (LOFAR) in the Netherlands.
2 papers · 1 benchmark
Usually, the information related to the crop types available in a given territory is annual information, that is, we only know the type of main crop grown over a year and we do not know any crops that have followed one another during the…
2 papers · 1 benchmark
MUSIC (Multi-Spectral Imaging via Computed Tomography)
The Multi-Spectral Imaging via Computed Tomography (MUSIC) dataset is a two-part (2D- and 3D spectral) open access dataset for advanced image analysis of spectral radiographic (x-ray) scans, their tomographic reconstruction and the…
2 papers · 0 benchmarks
MVTec D2S (MVTec Densely Segmented Supermarket)
MVTec D2S is a benchmark for instance-aware semantic segmentation in an industrial domain.
2 papers · 0 benchmarks
MagicBathyNet is a benchmark dataset made up of image patches of Sentinel-2, SPOT-6 and aerial imagery, bathymetry in raster format and seabed classes annotations.
2 papers · 0 benchmarks
Mila Simulated Floods Dataset is a 1.5 square km virtual world using the Unity3D game engine including urban, suburban and rural areas.
2 papers · 1 benchmark
NVGaze (NVGaze: An Anatomically-Informed Dataset for Low-Latency, Near-Eye Gaze Estimation)
Quality, diversity, and size of training dataset are critical factors for learning-based gaze estimators.
2 papers · 0 benchmarks
OADAT (OADAT: Experimental and Synthetic Clinical Optoacoustic Data for Standardized Image Processing)
An experimental and synthetic (simulated) OA raw signals and reconstructed image domain datasets rendered with different experimental parameters and tomographic acquisition geometries.
2 papers · 0 benchmarks
ODMS (Object Depth via Motion and Segmentation)
ODMS is a dataset for learning Object Depth via Motion and Segmentation.
2 papers · 0 benchmarks
OmniCity is a dataset for omnipotent city understanding from multi-level and multi-view images.
2 papers · 0 benchmarks
PAX-Ray++ (Projected Anatomy in X-Ray Dataset ++)
The PAX-Ray++ dataset uses pseudo-labeled thorax CTs to enable the segmentation of anatomy in Chest X-Rays.
2 papers · 0 benchmarks
A new benchmark dataset of webcam images, Photi-LakeIce, from multiple cameras and two different winters, along with pixel-wise ground truth annotations.
2 papers · 0 benchmarks
Pothole Mix (Pothole Mix Semantic Segmentation Dataset for Road Damage Detection and Segmentation)
This dataset for the semantic segmentation of potholes and cracks on the road surface was assembled from 5 other datasets already publicly available, plus a very small addition of segmented images on our part.
2 papers · 1 benchmark
https://paperswithcode.com/sota/semantic-segmentation-on-isprs-potsdam
2 papers · 2 benchmarks
RUGD (RUGD: Robot Unstructured Ground Driving)
A Video Dataset for Visual Perception and Autonomous Navigation in Unstructured Environments.
2 papers · 1 benchmark
The Retinal Microsurgery dataset is a dataset for surgical instrument tracking.
2 papers · 0 benchmarks
A special scene-graph for intelligent vehicles.
2 papers · 0 benchmarks
SB20 (Sugar Beet 2020 University of Bonn)
Video sequences captured at a field on Campus Kleinaltendorf (CKA), University of Bonn, captured by BonBot-I, an autonomous weeding robot.
2 papers · 0 benchmarks
Pre-training is a strong strategy for enhancing visual models to efficiently train them with a limited number of labeled images.
2 papers · 0 benchmarks
Homepage | GitHub LiDARs are one of the main sensors used for autonomous driving applications, providing accurate depth estimation regardless of lighting conditions.
2 papers · 0 benchmarks
TAS-NIR is a VIS+NIR dataset of semantically annotated images in unstructured outdoor environments.
2 papers · 0 benchmarks
TOP is a synthetic dataset for topology optimization generated using Topy.
2 papers · 0 benchmarks
U-DIADS-Bib is a proprietary dataset developed through the collaboration of computer scientists and humanities at the University of Udine.
2 papers · 1 benchmark
The Vistas-NP dataset is an out-of-distribution detection dataset based on the Mapillary Vistas dataset.
2 papers · 0 benchmarks
dacl10k (dacl10k: Dataset for Semantic Bridge Damage Segmentation)
dacl10k stands for damage classification 10k images and is a multi-label semantic segmentation dataset for 19 classes (13 damages and 6 objects) present on bridges.
2 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.