Home › Datasets › task › 2D Semantic Segmentation

2D Semantic Segmentation datasets

archive 2025-07-28

81 datasets carry the task tag "2D Semantic Segmentation" (the task itself: 2D Semantic Segmentation), ordered by the archive's paper count. Page 1 of 2: 48 shown of 81. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

2D Semantic Segmentation datasets 1–48 of 81

Cityscapes is a large-scale database which focuses on semantic understanding of urban street scenes.
3,702 papers · 51 benchmarks
KITTI-360 is a large-scale dataset that contains rich sensory information and full annotations.
246 papers · 7 benchmarks
CamVid (Cambridge-driving Labeled Video Database)
CamVid (Cambridge-driving Labeled Video Database) is a road/driving scene understanding database which was originally captured as five video sequences with a 960×720 resolution camera mounted on the dashboard of a car.
227 papers · 4 benchmarks
RELLIS-3D is a multi-modal dataset for off-road robotics.
57 papers · 3 benchmarks
The xBD dataset contains over 45,000KM2 of polygon labeled pre and post disaster imagery.
51 papers · 2 benchmarks
KITTI MOTS (KITTI Multi-Object Tracking and Segmentation (MOTS) Evaluation)
The Multi-Object and Segmentation (MOTS) benchmark [2] consists of 21 training sequences and 29 test sequences.
28 papers · 1 benchmark
ISTD+ consists of shadow images, shadow-free images, and shadow masks, with 1,330 training images and 540 testing images from 135 unique background scenes.
15 papers · 1 benchmark
A Multi-Task 4D Radar-Camera Fusion Dataset for Autonomous Driving on Water Surfaces description of the dataset WaterScenes, the first multi-task 4D radar-camera fusion dataset on water surfaces, which offers data from multiple sensors,…
13 papers · 2 benchmarks
SpaceNet 7 (Multi-Temporal Urban Development SpaceNet Dataset)
Satellite imagery analytics have numerous human development and disaster response applications, particularly when time series methods are involved.
12 papers · 0 benchmarks
WildScenes is a bi-modal benchmark dataset consisting of multiple large-scale, sequential traversals in natural environments, including semantic annotations in high-resolution 2D images and dense 3D LiDAR point clouds, and accurate 6-DoF…
12 papers · 2 benchmarks
CaDIS (Cataract Dataset for Image Segmentation)
CaDIS: a Cataract Dataset for Image Segmentation is a dataset for semantic segmentation created by Digital Surgery Ltd.
9 papers · 1 benchmark
UIIS (General Underwater Image Instance Segmentation dataset)
This is the first general Underwater Image Instance Segmentation (UIIS) dataset containing 4,628 images for 7 categories with pixel-level annotations for underwater instance segmentation task
9 papers · 1 benchmark
Are current 3D object tracking methods truely robust enough for low-fidelity depth sensors like the iPhone LiDAR?
8 papers · 2 benchmarks
Open Images is a computer vision dataset covering ~9 million images with labels spanning thousands of object categories.
6 papers · 0 benchmarks
This basketball dataset was acquired under the Walloon region project DeepSport, using the Keemotion system installed in multiple arenas.
5 papers · 0 benchmarks
This brain tumor dataset contains 3064 T1-weighted contrast-enhanced images with three kinds of brain tumor.
4 papers · 0 benchmarks
Unsupervised Domain Adaptation demonstrates great potential to mitigate domain shifts by transferring models from labeled source domains to unlabeled target domains.
4 papers · 3 benchmarks
We present the CrackVision12k dataset, a collection of 12,000 crack images derived from 13 publicly available crack datasets.
4 papers · 1 benchmark
FES (Fisheye Evaluation Suite)
FES is an indoor dataset that can be used for evaluation of deep learning approaches.
4 papers · 0 benchmarks
USIS10K (Large-scale Underwater Salient Instance Segmentation Dataset)
We construct the first large-scale dataset, USIS10K, for the underwater salient instance segmentation task, which contains 10,632 images and pixel-level annotations of 7 categories.
4 papers · 0 benchmarks
A large-scale video portrait dataset that contains 291 videos from 23 conference scenes with 14K frames.
3 papers · 0 benchmarks
UIIS10K (General Underwater Image Instance Segmentation dataset 10K)
We propose a large-scale underwater instance segmentation dataset, UIIS10K, which includes 10,048 images with pixel-level annotations for 10 categories.
3 papers · 0 benchmarks
fluocells (Fluorescent Neuronal Cells)
By releasing this dataset, we aim at providing a new testbed for computer vision techniques using Deep Learning.
3 papers · 0 benchmarks
A real-world dataset, with hyper-accurate digital counterpart & comprehensive ground-truth annotation.
2 papers · 1 benchmark
CaFFe (CAlving Fronts and where to Find thEm)
The temporal variability in calving front positions of marine-terminating glaciers permits inference on the frontal ablation.
2 papers · 2 benchmarks
The dataset available for download on this webpage represents a 5x5x5µm section taken from the CA1 hippocampus region of the brain, corresponding to a 1065x2048x1536 volume.
2 papers · 1 benchmark
FaceOcc (Face Occlusion Dataset)
FaceOcc is a high-quality face occlusion dataset which contains all mislabeled occlusions in CelebAMask-HQ and complements some occlusions and textures from the internet.
2 papers · 0 benchmarks
Fetoscopic Placental Vessel Segmentation and Registration (FetReg2021) challenge was organized as part of the MICCAI2021 Endoscopic Vision (EndoVis) challenge.
2 papers · 0 benchmarks
HuTics (Human Deictic Gestures Dataset)
HuTics contains 2040 images showing how humans use deictic gestures to interact with various daily-life objects.
2 papers · 0 benchmarks
MMVR (Millimeter-wave Multi-View Radar (MMVR) Dataset)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
2 papers · 0 benchmarks
PAX-Ray++ (Projected Anatomy in X-Ray Dataset ++)
The PAX-Ray++ dataset uses pseudo-labeled thorax CTs to enable the segmentation of anatomy in Chest X-Rays.
2 papers · 0 benchmarks
RUGD (RUGD: Robot Unstructured Ground Driving)
A Video Dataset for Visual Perception and Autonomous Navigation in Unstructured Environments.
2 papers · 1 benchmark
RaidaR (RaidaR: A Rich Annotated Image Dataset of Rainy Street Scenes)
RaidaR is a rich annotated image dataset of rainy street scenes.
2 papers · 0 benchmarks
Pre-training is a strong strategy for enhancing visual models to efficiently train them with a limited number of labeled images.
2 papers · 0 benchmarks
TBBR (Thermal Bridges on Building Rooftops)
The dataset of Thermal Bridges on Building Rooftops (TBBR dataset) consists of annotated combined RGB and thermal drone images with a height map.
2 papers · 2 benchmarks
U-DIADS-Bib is a proprietary dataset developed through the collaboration of computer scientists and humanities at the University of Udine.
2 papers · 1 benchmark
ACCT Data Repository (ACCT is a fast and accessible automatic cell counting tool using machine learning for 2D image segmentation)
This dataset is a collection of fluorescent images from mice in order to test an automatic cell counting tool that we developed.
1 paper · 0 benchmarks
AUT-VI (Amirkabir campus dataset)
AUT-VI is a super-challenging visual inertial dataset with 126 diverse sequences in 17 locations.
1 paper · 0 benchmarks
BCSS (Breast Cancer Semantic Segmentation)
The BCSS dataset contains over 20,000 segmentation annotations of tissue regions from breast cancer images from The Cancer Genome Atlas (TCGA).
1 paper · 0 benchmarks
DARai (Daily Activity Recordings for AI and ML applications)
Daily Activity Recordings for Artificial Intelligence (DARai, pronounced "Dahr-ree") is a multimodal, hierarchically annotated dataset constructed to understand human activities in real-world settings.
1 paper · 0 benchmarks
Deep Indices (multi-spectral leaf/vegetation segmentation)
This dataset inclue multi-spectral acquisition of vegetation for the conception of new DeepIndices.
1 paper · 1 benchmark
DermSynth3D (3DBodyTex.DermSynth3D)
A dataset of 100K synthetic images of skin lesions, ground-truth (GT) segmentations of lesions and healthy skin, GT segmentations of seven body parts (head, torso, hips, legs, feet, arms and hands), and GT binary masks of non-skin regions…
1 paper · 0 benchmarks
A synthetic dataset including driving under adverse weather conditions | Autonomous Driving
1 paper · 0 benchmarks
The dataset X of this work is an extension of the heartSeg dataset.
1 paper · 1 benchmark
Optical images of printed circuit boards as well as detailed annotations of any text, logos, and surface-mount devices (SMDs).
1 paper · 0 benchmarks
FP4S (Floor plan image segmentation via scribble-based semi-weakly-supervised learning)
We introduce a new style- and category-agnostic floor plan image parsing benchmark developed in collaboration with professional architectural designers.
1 paper · 1 benchmark
GF-PA66 3D XCT (Glass fiber-reinforced polyamide 66 (GF-PA66) 3D X-ray Computed Tomography (XCT)))
Stack of 2D gray images of glass fiber-reinforced polyamide 66 (GF-PA66) 3D X-ray Computed Tomography (XCT) specimen.
1 paper · 1 benchmark

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.