Home › Datasets › modality › RGB-D

RGB-D datasets

archive 2025-07-28

190 datasets carry the modality tag "RGB-D", ordered by the archive's paper count. Page 4 of 4: 46 shown of 190. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

RGB-D datasets 145–190 of 190

MetaGraspNet 2 (MetaGraspNet difficulty 2 - medium)
There has been increasing interest in smart factories powered by robotics systems to tackle repetitive, laborious tasks.
1 paper · 0 benchmarks
MetaGraspNet 3 (MetaGraspNet difficulty 3 - hard 1)
There has been increasing interest in smart factories powered by robotics systems to tackle repetitive, laborious tasks.
1 paper · 0 benchmarks
MetaGraspNet 4 (MetaGraspNet difficulty 4 - hard 2)
There has been increasing interest in smart factories powered by robotics systems to tackle repetitive, laborious tasks.
1 paper · 0 benchmarks
MetaGraspNet 5 (MetaGraspNet difficulty 5 - very hard)
There has been increasing interest in smart factories powered by robotics systems to tackle repetitive, laborious tasks.
1 paper · 0 benchmarks
MuSoHu (Toward human-like social robot navigation: A large-scale, multi-modal, social human navigation dataset)
A large-scale, egocentric, multimodal, and context-aware dataset of human demonstrations of social navigation.
1 paper · 0 benchmarks
Multiview Manipulation Data (Multiview Manipulation Expert Data and Trained Models)
Accompanying expert data and trained models for 2021 IROS paper on Multiview Manipulation.
1 paper · 0 benchmarks
NBMOD (Noisy Background Multi-Object Dataset for grasp detection)
Introduction NBMOD is a dataset created for researching the task of specific object grasp detection by robots in noisy environments.
1 paper · 1 benchmark
NPO (Negative and Positive Obstacles)
The dataset is recorded with an on-vehicle ZED stereo camera in both urban and rural environments The dataset contains various lighting conditions, such as normal lights, large-area shadows, dim lights, and sun glare.
1 paper · 1 benchmark
A RGB-D dataset converted from NYUDv2 into COCO-style instance segmentation format.
1 paper · 2 benchmarks
Omiverse Object is a large-scale synthetic dataset of 60,000 images including both transparent and opaque objects in different scenes.
1 paper · 0 benchmarks
PLAD (Point Line and Depth dataset)
PLAD is a dataset where sparse depth is provided by line-based visual SLAM to verify StructMDC.
1 paper · 1 benchmark
The RMRC 2014 indoor dataset is a dataset for indoor semantic segmentation.
1 paper · 0 benchmarks
RTB (Robot Tracking Benchmark)
The Robot Tracking Benchmark (RTB) is a synthetic dataset that facilitates the quantitative evaluation of 3D tracking algorithms for multi-body objects.
1 paper · 1 benchmark
A total of 80 real material samples were captured in a dark room.
1 paper · 0 benchmarks
The dataset contains both RGB and depth images, and the data from two accelerometers, together with ground truth calorie values from a calorimeter for calorie expenditure estimation in home environments.
1 paper · 0 benchmarks
A RGB-D dataset converted from SUN-RGBD into COCO-style instance segmentation format.
1 paper · 2 benchmarks
SceneNet-RGBD is a synthetic dataset containing large-scale photorealistic renderings of indoor scene trajectories with pixel-level annotations.
1 paper · 0 benchmarks
THEOStereo is a dataset providing synthetic stereo image pairs and their corresponding scene depth and will be published along with [1].
1 paper · 0 benchmarks
The RBO dataset of articulated objects and interactions is a collection of 358 RGB-D video sequences (67:18 minutes) of humans manipulating 14 articulated objects under varying conditions (light, perspective, background, interaction).
1 paper · 0 benchmarks
UIUC Scooping Dataset (Granular Materials Manipulation Dataset with Scooping/Digging/Excavation Action)
Overview: This dataset encompasses a compilation of 6,700 executed scoops (excavations), mapped across a vast spectrum of materials, terrain topography, and compositions.
1 paper · 0 benchmarks
In this dataset UR5 robot used 6 tools: metal-scissor, metal-whisk, plastic-knife, plastic-spoon, wooden-chopstick, and wooden-fork to perform 6 behaviors: look, stirring-slow, stirring-fast, stirring-twist, whisk, and poke.
1 paper · 0 benchmarks
VR-Folding contains garment meshes of 4 categories from CLOTH3D dataset, namely Shirt, Pants, Top and Skirt.
1 paper · 0 benchmarks
The ViCoS Towel Dataset is a state-of-the-art benchmark for grasp point localization on cloth objects, specifically towels.
1 paper · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
The YCB-Ev dataset contains synchronized RGB-D frames and event data that enables evaluating 6DoF object pose estimation algorithms using these modalities.
1 paper · 0 benchmarks
In this dataset we teleoperated UR5 arm to collect manipulation data for picking up a screwdriver in a cluttered tabletop environment.
1 paper · 0 benchmarks
rc_49 (rc_49 Grasping Dataset)
Includes several sets of synthetic stereo images labelled with grasp rectangles representing parallel-jaw grasps (Cornell-like format).
1 paper · 0 benchmarks
DAHLIA (DAily Human Life Activity)
DAHLIA dataset [1] is devoted to human activity recognition, which is a major issue for adapting smart-home services such as user assistance.
0 papers · 0 benchmarks
The Freiburg Campus 3D Scan dataset consists of 3D area maps from the Freiburg campus that were scanned with 3D lasers.
0 papers · 0 benchmarks
Freiburg Lighting Adaptable Map Tracking is a dataset for camera trajectory estimation.
0 papers · 0 benchmarks
The Freiburg RGB-D People dataset contains 3000+ RGB-D frames acquired in a university hall from three vertically mounted Kinect sensors.
0 papers · 0 benchmarks
HEADSET (HEADSET: Human Emotion Awareness under Partial Occlusions Multimodal DataSET)
The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications.
0 papers · 0 benchmarks
HouseCat6D (A Large-Scale Multi-Modal Category Level 6D Object Perception Dataset with Household Objects in Realistic Scenarios)
Estimating 6D object poses is a major challenge in 3D computer vision.
0 papers · 0 benchmarks
Overview The IITKGPFence dataset is designed for tasks related to fence-like occlusion detection, defocus blur, depth mapping, and object segmentation.
0 papers · 0 benchmarks
InfiniteRep is a synthetic, open-source dataset for fitness and physical therapy (PT) applications.
0 papers · 0 benchmarks
MHRI dataset (Multimodal Human-Robot Interaction dataset)
The dataset includes recordings from 10 different users teaching the robot different common kitchen objects, that consists of synchronized recordings from three cameras and a microphone mounted on the robot: An RGB-d camera covers the user…
0 papers · 0 benchmarks
PAVIS RGB-D is a dataset for person re-identification using depth information.
0 papers · 0 benchmarks
Pose Estimation Lunar Robot (Dataset for camera pose estimation research using computer simulated images from rovers on the lunar surface)
Overview The goal: using simulation data to train neural networks to estimate the pose of a rover's camera with respect to a known target object The mission context: A simulated lunar surface, with lunar landers and lunar rovers.
0 papers · 0 benchmarks
Dataset: RGB-D Images for Real-World and Synthetic Object Scenes This dataset consists of both real-world and synthetic RGB-D images, designed for object detection, classification, and segmentation tasks, particularly for primitive shape…
0 papers · 0 benchmarks
0 papers · 0 benchmarks
Sugar Beets 2016 is a robot dataset for plant classification as well as localization and mapping that covers the relevant stages for robotic intervention and weed control.
0 papers · 0 benchmarks
Toronto NeuroFace Dataset: A New Dataset for Facial Motion Analysis in Individuals with Neurological Disorders Toronto NeuroFace Dataset is a public dataset with videos of oro-facial gestures performed by individuals with oro-facial…
0 papers · 0 benchmarks
UNIPD-BPE (University of Padova Body Pose Estimation)
The University of Padova Body Pose Estimation dataset (UNIPD-BPE) is an extensive dataset for multi-sensor body pose estimation containing both single-person and multi-person sequences with up to 4 interacting people A network with 5…
0 papers · 0 benchmarks
The RGB-D Scenes Dataset contains 8 scenes annotated with objects that belong to the Washington RGB-D Object Dataset.
0 papers · 0 benchmarks
The RGB-D Scenes Dataset v2 consists of 14 scenes containing furniture (chair, coffee table, sofa, table) and a subset of the objects in the RGB-D Object Dataset (bowls, caps, cereal boxes, coffee mugs, and soda cans).
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.