Home › Datasets › task › Instance Segmentation

Instance Segmentation datasets

archive 2025-07-28

111 datasets carry the task tag "Instance Segmentation" (the task itself: Instance Segmentation), ordered by the archive's paper count. Page 2 of 3: 48 shown of 111. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Instance Segmentation datasets 49–96 of 111

WGISD (Embrapa Wine Grape Instance Segmentation Dataset)
Embrapa Wine Grape Instance Segmentation Dataset (WGISD) contains grape clusters properly annotated in 300 images and a novel annotation methodology for segmentation of complex objects in natural images.
5 papers · 0 benchmarks
ARMBench is a large-scale, object-centric benchmark dataset for robotic manipulation in the context of a warehouse.
4 papers · 1 benchmark
The Fraunhofer IPA Bin-Picking dataset is a large-scale dataset comprising both simulated and real-world scenes for various objects (potentially having symmetries) and is fully annotated with 6D poses.
4 papers · 0 benchmarks
Separated COCO is automatically generated subsets of COCO val dataset, collecting separated objects for a large variety of categories in real images in a scalable manner, where target object segmentation mask is separated into distinct…
4 papers · 1 benchmark
TexBiG (Text-Bild-Gefüge)
TexBiG (from the German Text-Bild-Gefüge, meaning Text-Image-Structure) is a document layout analysis dataset for historical documents in the late 19th and early 20th century.
4 papers · 4 benchmarks
The TimberSeg 1.0 dataset is composed of 220 images showing wood logs in various environments and conditions in Canada.
4 papers · 0 benchmarks
USIS10K (Large-scale Underwater Salient Instance Segmentation Dataset)
We construct the first large-scale dataset, USIS10K, for the underwater salient instance segmentation task, which contains 10,632 images and pixel-level annotations of 7 categories.
4 papers · 0 benchmarks
WISDOM (Warehouse Instance Segmentation Dataset for Object Manipulation)
Synthetic training dataset of 50,000 depth images and 320,000 object masks using simulated heaps of 3D CAD models.
4 papers · 1 benchmark
iShape is an irregular shape dataset for instance segmentation.
4 papers · 1 benchmark
The Aircraft Context Dataset, a composition of two inter-compatible large-scale and versatile image datasets focusing on manned aircraft and UAVs, is intended for training and evaluating classification, detection and segmentation models in…
3 papers · 0 benchmarks
DeepSportradar is a benchmark suite of computer vision tasks, datasets and benchmarks for automated sport understanding.
3 papers · 0 benchmarks
GraSP (Holistic and Multi-Granular Surgical Scene Understanding of Prostatectomies)
Holistic and Multi-Granular Surgical Scene Understanding of Prostatectomies (GraSP) dataset, a curated benchmark that models surgical scene understanding as a hierarchy of complementary tasks with varying levels of granularity.
3 papers · 1 benchmark
OAM-TCD is a dataset of around 5k aerial images from around the world to support robust tree detection algorithms.
3 papers · 0 benchmarks
Real-world dataset of ~400 images of cuboid-shaped parcels with full 2D and 3D annotations in the COCO format.
3 papers · 0 benchmarks
RobotPush is a dataset for object singulation – the task of separating cluttered objects through physical interaction.
3 papers · 0 benchmarks
UIIS10K (General Underwater Image Instance Segmentation dataset 10K)
We propose a large-scale underwater instance segmentation dataset, UIIS10K, which includes 10,048 images with pixel-level annotations for 10 categories.
3 papers · 0 benchmarks
This data set contains 775 video sequences, captured in the wildlife park Lindenthal (Cologne, Germany) as part of the AMMOD project, using an Intel RealSense D435 stereo camera.
2 papers · 0 benchmarks
MOTFront provides photo-realistic RGB-D images with their corresponding instance segmentation masks, class labels, 2D & 3D bounding boxes, 3D geometry, 3D poses and camera parameters.
2 papers · 0 benchmarks
MVTec D2S (MVTec Densely Segmented Supermarket)
MVTec D2S is a benchmark for instance-aware semantic segmentation in an industrial domain.
2 papers · 0 benchmarks
Occluded COCO is automatically generated subset of COCO val dataset, collecting partially occluded objects for a large variety of categories in real images in a scalable manner, where target object is partially occluded but the…
2 papers · 1 benchmark
OmniCity is a dataset for omnipotent city understanding from multi-level and multi-view images.
2 papers · 0 benchmarks
Synthetic dataset of over 13,000 images of damaged and intact parcels with full 2D and 3D annotations in the COCO format.
2 papers · 0 benchmarks
A set of 221 stereo videos captured by the SOCRATES stereo camera trap in a wildlife park in Bonn, Germany between February and July of 2022.
2 papers · 0 benchmarks
SB20 (Sugar Beet 2020 University of Bonn)
Video sequences captured at a field on Campus Kleinaltendorf (CKA), University of Bonn, captured by BonBot-I, an autonomous weeding robot.
2 papers · 0 benchmarks
TBBR (Thermal Bridges on Building Rooftops)
The dataset of Thermal Bridges on Building Rooftops (TBBR dataset) consists of annotated combined RGB and thermal drone images with a height map.
2 papers · 2 benchmarks
UW Indoor Scenes (UW-IS) Occluded dataset is curated using commodity hardware (Intel RealSense D435) to reflect real world robotics scenarios.
2 papers · 0 benchmarks
BPCIS (Bacterial Phase Contrast for Instance Segementation)
BPCIS is collection of 364 bacterial phase contrast images and corresponding label matrices for instance segmentation.
1 paper · 0 benchmarks
RGB-D instance segmentation box dataset.
1 paper · 2 benchmarks
CCSE (Chinese Character Stroke Extraction)
Chinese Character Stroke Extraction (CCSE) is a benchmark containing two large-scale datasets: Kaiti CCSE (CCSE-Kai) and Handwritten CCSE (CCSE-HW).
1 paper · 0 benchmarks
COCO-N Medium introduces a stochastic benchmark that simulates common real-world scenarios with noticeable label inaccuracies in the COCO dataset.
1 paper · 1 benchmark
COCO-WAN (Medium noise) (Benchmarking Label Noise in Instance Segmentation: Spatial Noise Matters)
The COCO-WAN benchmark is designed to assess the impact of weakly annotations (combined with auto-annotation tools) noise on instance segmentation models.
1 paper · 1 benchmark
DLO Instance Segmentation dataset (DLO Instance Segmentation dataset generated by Blender)
Contains ~60000 HD images of Deformable Linear Objects (DLOs) generated using blender.
1 paper · 0 benchmarks
ENSeg Dataset Overview This dataset represents an enhanced subset of the ENS dataset.
1 paper · 1 benchmark
The FashionFail dataset comprises 2,495 high-resolution images (2400x2400 pixels) of products found on e-commerce websites.
1 paper · 0 benchmarks
The 'Me 163' was a Second World War fighter airplane and a result of the German air force secret developments.
1 paper · 0 benchmarks
GUISS dataset (Meshes, textures, Blend files, stereo datasets, depth maps, depth estimations))
We provide all the expected data inputs to GUISS such as meshes, texture images, and blend files.
1 paper · 0 benchmarks
GraspClutter6D is a large-scale real-world dataset for robust object perception and robotic grasping in cluttered environments.
1 paper · 0 benchmarks
HT1080WT cells - 3D collagen type I matrices (HT1080WT cells embedded in 3D collagen type I matrices - manual annotations for cell instance segmentation and tracking)
Human fibrosarcoma HT1080WT (ATCC) cells at low cell densities embedded in 3D collagen type I matrices [1].
1 paper · 0 benchmarks
Heritage Pointcloud Instance Collection dataset, acquired from two large buildings and annotated at a point-wise semantic level based on existent BIM models.
1 paper · 1 benchmark
- Revision: v1.0.0-full-20210527a - DOI: 10.5281/zenodo.4817662 - Authors: J.
1 paper · 0 benchmarks
This publicly available dataset contains 1613 RGB-D images of field-grown broccoli plants.
1 paper · 0 benchmarks
LDD (LDD: A Grape Diseases Dataset Detection and Instance Segmentation)
The Instance Segmentation task, an extension of the well-known Object Detection task, is of great help in many areas, such as precision agriculture: being able to automatically identify plant organs and the possible diseases associated…
1 paper · 2 benchmarks
MIS-Check Dam (Minor Irrigation Structures- Check Dam)
Minor Irrigation Structures Check-Dam Dataset is a public dataset annotated by domain experts using images from Google static map for instance segmentation and object detection tasks.
1 paper · 0 benchmarks
MiSCS (Microscopic Shrub Cross Sections)
Microscopy images of shrub cross sections for instance segmentation of tree rings.
1 paper · 0 benchmarks
A RGB-D dataset converted from NYUDv2 into COCO-style instance segmentation format.
1 paper · 2 benchmarks
An object-centric version of Stylized COCO to benchmark texture bias and out-of-distribution robustness of vision models.
1 paper · 0 benchmarks
PWISeg (PWISeg Surgical Instruments Dataset)
Overview The Surgical Instruments Recognition Dataset is a groundbreaking collection of high-resolution images (1280x960 pixels) specifically designed for the recognition and categorization of surgical instruments.
1 paper · 0 benchmarks
A multimodal dataset of radio galaxies and their corresponding infrared hosts.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.