Home › Datasets › modality › 3D

3D datasets

archive 2025-07-28

380 datasets carry the modality tag "3D", ordered by the archive's paper count. Page 3 of 8: 48 shown of 380. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

3D datasets 97–144 of 380

Hi4D contains 4D textured scans of 20 subject pairs, 100 sequences, and a total of more than 11K frames.
19 papers · 0 benchmarks
SSP-3D (Sports Shape and Pose 3D)
SSP-3D is an evaluation dataset consisting of 311 images of sportspersons in tight-fitted clothes, with a variety of body shapes and poses.
19 papers · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
18 papers · 1 benchmark
The dataset collected at the University of Florence during 2012, has been captured using a Kinect camera.
18 papers · 1 benchmark
PU1K is nearly 8 times larger than the largest publicly available dataset collected by PU-GAN.
18 papers · 0 benchmarks
TUM-VIE (TUM Stereo Visual-Inertial Event Dataset)
TUM-VIE is an event camera dataset for developing 3D perception and navigation algorithms.
18 papers · 0 benchmarks
Many existing datasets for lidar place recognition are solely representative of structured urban environments, and have recently been saturated in performance by deep learning based approaches.
18 papers · 1 benchmark
ZInd (Zillow Indoor Dataset)
The Zillow Indoor Dataset (ZInD) provides extensive visual data that covers a real world distribution of unfurnished residential homes.
18 papers · 1 benchmark
ViViD++ (Vision for Visibility Dataset)
A dataset capturing diverse visual data formats that target varying luminance conditions, and was recorded from alternative vision sensors, by handheld or mounted on a car, repeatedly in the same space but in different conditions.
17 papers · 0 benchmarks
VehicleX is a large-scale synthetic dataset.
16 papers · 0 benchmarks
3D AffordanceNet is a dataset of 23k shapes for visual affordance.
15 papers · 1 benchmark
CustomHumans is recorded by a multi-view photogrammetry system equipped with 53 RGB (12 Megapixels) and 53 (4 Megapixels) IR cameras.
15 papers · 1 benchmark
A novel benchmark dataset that includes a manually annotated point cloud for over 260 million laser scanning points into 100'000 (approx.) assets from Dublin LiDAR point cloud [12] in 2015.
15 papers · 0 benchmarks
Dataset of clothing size variation which includes different subjects wearing casual clothing items in various sizes, totaling to approximately 2000 scans.
14 papers · 0 benchmarks
Purpose Medical imaging has become increasingly important in diagnosing and treating oncological patients, particularly in radiotherapy.
14 papers · 0 benchmarks
Created for MVS tasks and is a large-scale multi-view aerial dataset generated from a highly accurate 3D digital surface model produced from thousands of real aerial images with precise camera parameters.
14 papers · 0 benchmarks
Super-CLEVR is a dataset for Visual Question Answering (VQA) where different factors in VQA domain shifts can be isolated in order that their effects can be studied independently.
13 papers · 0 benchmarks
SILK (Synth It Like KITTI)
An important factor in advancing autonomous driving systems is simulation.
12 papers · 0 benchmarks
4D-OR includes a total of 6734 scenes, recorded by six calibrated RGB-D Kinect sensors 1 mounted to the ceiling of the OR, with one frame-per-second, providing synchronized RGB and depth images.
11 papers · 3 benchmarks
The Argoverse 2 Sensor Dataset is a collection of 1,000 scenarios with 3D object tracking annotations.
11 papers · 0 benchmarks
BIKED is a dataset comprised of 4500 individually designed bicycle models sourced from hundreds of designers.
11 papers · 0 benchmarks
BuildingNet is a large-scale dataset of 3D building models whose exteriors are consistently labeled.
11 papers · 0 benchmarks
A rich, extensible and efficient environment that contains 45,622 human-designed 3D scenes of visually realistic houses, ranging from single-room studios to multi-storied houses, equipped with a diverse set of fully labeled 3D objects,…
11 papers · 0 benchmarks
Housekeep a benchmark to evaluate common sense reasoning in the home for embodied AI.
11 papers · 0 benchmarks
UnrealEgo is a dataset that provides in-the-wild stereo images with a large variety of motions for 3D human pose estimation.
11 papers · 1 benchmark
AKB-48 is a large-scale Articulated object Knowledge Base which consists of 2,037 real-world 3D articulated object models of 48 categories.
3D
10 papers · 0 benchmarks
The MM-WHS 2017 dataset is a dataset for multi-modality whole heart segmentation.
10 papers · 1 benchmark
Memory Maze is a 3D domain of randomized mazes designed for evaluating the long-term memory abilities of RL agents.
10 papers · 0 benchmarks
Unite The People is a dataset for 3D body estimation.
10 papers · 0 benchmarks
xR-EgoPose is an egocentric synthetic dataset for egocentric 3D human pose estimation.
10 papers · 0 benchmarks
BLVD is a large scale 5D semantics dataset collected by the Visual Cognitive Computing and Intelligent Vehicles Lab.
9 papers · 0 benchmarks
We provide manual annotations of 14 semantic keypoints for 100,000 car instances (sedan, suv, bus, and truck) from 53,000 images captured from 18 moving cameras at Multiple intersections in Pittsburgh, PA.
9 papers · 2 benchmarks
Mindboggle is a large publicly available dataset of manually labeled brain MRI.
9 papers · 0 benchmarks
The Zenseact Open Dataset (ZOD) is a large-scale and diverse multi-modal autonomous driving (AD) dataset, created by researchers at Zenseact.
9 papers · 0 benchmarks
3DFAW contains 23k images with 66 3D face keypoint annotations.
8 papers · 2 benchmarks
ATLAS v2.0 (Anatomical Tracings of Lesions After Stroke Dataset version 2.0)
Accurate lesion segmentation is critical in stroke rehabilitation research for the quantification of lesion burden and accurate image processing.
8 papers · 1 benchmark
Are current 3D object tracking methods truely robust enough for low-fidelity depth sensors like the iPhone LiDAR?
8 papers · 2 benchmarks
ETH SfM (ETH Structure-from-Motion)
The ETH SfM (structure-from-motion) dataset is a dataset for 3D Reconstruction.
8 papers · 0 benchmarks
FeTS2022 (Federated Tumor Segmentation Challenge 2022)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
8 papers · 0 benchmarks
HANDAL (HANDAL: A Dataset of Real-World Manipulable Object Categories with Pose Annotations, Affordances, and Reconstructions)
We present the HANDAL dataset for category-level object pose estimation and affordance prediction.
8 papers · 0 benchmarks
T³Bench is the first comprehensive text-to-3D benchmark containing diverse text prompts of three increasing complexity levels that are specially designed for 3D generation (300 prompts in total).
8 papers · 1 benchmark
Thingi10K is a dataset of 3D-Printing Models.
3D
8 papers · 0 benchmarks
TruckScenes (MAN TruckScenes)
Autonomous trucking is a promising technology that can greatly impact modern logistics and the environment.
8 papers · 1 benchmark
WADS (Winter Adverse Driving dataSet)
Collected in the snow belt region of Michigan's Upper Peninsula, WADS is the first multi-modal dataset featuring dense point-wise labeled sequential LiDAR scans collected in severe winter weather.
8 papers · 0 benchmarks
X-Humans consists of 20 subjects (11 males, 9 females) with various clothing types and hair style.
8 papers · 0 benchmarks
3D-FRONT (3D Furnished Rooms with layOuts and semaNTics) is large-scale, and comprehensive repository of synthetic indoor scenes highlighted by professionally designed layouts and a large number of rooms populated by high-quality textured…
3D
7 papers · 0 benchmarks
AIOZ-GDANCE comprises 16.7 hours of whole-body motion and music audio of group dancing.
7 papers · 1 benchmark

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.