Datasets › 2D-3D-S
2D-3D-S (2D-3D-Semantic)
The 2D-3D-S dataset provides a variety of mutually registered modalities from 2D, 2.5D and 3D domains, with instance-level semantic and geometric annotations. It covers over 6,000 m2 collected in 6 large-scale indoor areas that originate from 3 different buildings. It contains over 70,000 RGB images, along with the corresponding depths, surface normals, semantic annotations, global XYZ images (all in forms of both regular and 360° equirectangular images) as well as camera information. It also includes registered raw and semantically annotated 3D meshes and point clouds. The dataset enables development of joint and cross-modal learning models and potentially unsupervised approaches utilizing the regularities present in large-scale indoor spaces.
Source: https://github.com/alexsax/2D-3D-Semantics Image Source: https://github.com/alexsax/2D-3D-Semantics
Benchmarks archive 2025-07-28
All 6 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Semantic Segmentation | Stanford2D3D Panoramic | SFSS-MMSI (RGB+HHA) mIoU 60.6% | Single Frame Semantic Segmentation Using Multi-Modal... | sguttikon/SFSS-MMSI | 25 | Compare |
| Depth Estimation | Stanford2D3D Panoramic | HiMODE RMSE 0.2619 | HiMODE: A Hybrid Monocular Omnidirectional Depth Estimation Model | — | 18 | Compare |
| 3D Room Layouts From A Single RGB Panorama | Stanford2D3D Panoramic | DMH-Net 3DIoU 84.93 | 3D Room Layout Estimation from a Cubemap of Panorama... | starrah/dmh-net | 9 | Compare |
| Semantic Segmentation | Stanford2D3D - RGBD | CMX (SegFormer-B4) mIoU 62.1 | CMX: Cross-Modal Fusion for RGB-X Semantic Segmentation... | huaaaliu/rgbx_semantic_segmentation | 6 | Compare |
| Semantic Segmentation | Stanford2D3D Panoramic - RGBD | CBFC mAcc 70.8 | Complementary Bi-directional Feature Compression for... | — | 3 | Compare |
| Semi-Supervised Semantic Segmentation | 2D-3D-S | M3L (Linear Fusion B2) mIoU (0.1% labels) 40.05 | Missing Modality Robustness in Semi-Supervised... | harshm121/m3l | 1 | Compare |
Papers archive 2025-07-28
30 shown of 42 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 147. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
The full list of 42 is in the JSON twin.
Dataset loaders archive 2025-07-28
1 loader as listed in the archive; links are outbound and not re-checked here.
Tasks archive 2025-07-28
License archive 2025-07-28
Modalities archive 2025-07-28
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
- 2D-3D-S
- Stanford2D3D
- Stanford2D3D Panoramic
- Stanford2D3D Panoramic - RGBD
- Stanford2D3D - RGBD
5 variant names, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections