Home › Datasets › task › 3D Pose Estimation
3D Pose Estimation datasets
archive 2025-07-28
36 datasets carry the task tag "3D Pose Estimation" (the task itself: 3D Pose Estimation), ordered by the archive's paper count. Page 1 of 1: 36 shown of 36. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
3D Pose Estimation datasets 1–36 of 36
The Human3.6M dataset is one of the largest motion capture datasets, which consists of 3.6 million human poses and corresponding images captured by a high-speed motion capture system.
783 papers · 13 benchmarks
The Leeds Sports Pose (LSP) dataset is widely used as the benchmark for human pose estimation.
202 papers · 1 benchmark
For many fundamental scene understanding tasks, it is difficult or impossible to obtain per-pixel ground truth labels from real images.
108 papers · 4 benchmarks
Dataset produced for the SAPIEN simulation environment.
88 papers · 0 benchmarks
A hand-object interaction dataset with 3D pose annotations of hand and object.
36 papers · 2 benchmarks
MuCo-3DHP is a large scale training data set showing real images of sophisticated multi-person interactions and occlusions.
35 papers · 0 benchmarks
ApolloCar3DT is a dataset that contains 5,277 driving images and over 60K car instances, where each car is fitted with an industry-grade 3D CAD model with absolute model size and semantically labelled keypoints.
17 papers · 14 benchmarks
SVIRO (Synthetic Vehicle Interior Rear Seat Occupancy Dataset)
Contains bounding boxes for object detection, instance segmentation masks, keypoints for pose estimation and depth images for each synthetic scenery as well as images for each individual seat for classification.
15 papers · 0 benchmarks
MVOR (Multi-View Operating Room)
Multi-View Operating Room (MVOR) is a dataset recorded during real clinical interventions.
14 papers · 0 benchmarks
UnrealEgo is a dataset that provides in-the-wild stereo images with a large variety of motions for 3D human pose estimation.
11 papers · 1 benchmark
Unite The People is a dataset for 3D body estimation.
10 papers · 0 benchmarks
We provide manual annotations of 14 semantic keypoints for 100,000 car instances (sedan, suv, bus, and truck) from 53,000 images captured from 18 moving cameras at Multiple intersections in Pittsburgh, PA.
9 papers · 2 benchmarks
HUMAN4D is a large and multimodal 4D dataset that contains a variety of human activities simultaneously captured by a professional marker-based MoCap, a volumetric capture and an audio recording system.
9 papers · 0 benchmarks
GPA (Geometric Pose Affordance)
multi-view imagery of people interacting with a variety of rich 3D environments
8 papers · 2 benchmarks
OOD-CV (Out Of Distribution Generalization in Computer Vision)
Enhancing the robustness of vision algorithms in real-world scenarios is challenging.
6 papers · 1 benchmark
HARPER (Exploring 3D Human Pose Estimation and Forecasting from the Robot’s Perspective: The HARPER Dataset)
We introduce HARPER, a novel dataset for 3D body pose estimation and forecast in dyadic interactions between users and \spot, the quadruped robot manufactured by Boston Dynamics.
5 papers · 3 benchmarks
PedX is a large-scale multi-modal collection of pedestrians at complex urban intersections.
5 papers · 0 benchmarks
SportsPose (SportsPose - A Dynamic 3D sports pose dataset)
Accurate 3D human pose estimation is essential for sports analytics, coaching, and injury prevention.
4 papers · 0 benchmarks
The dataset is designed specifically to solve a range of computer vision problems (2D-3D tracking, posture) faced by biologists while designing behavior studies with animals.
3 papers · 0 benchmarks
Fitness-AQA (Fitness Action Quality Assessment [ECCV 2022])
Largest, first-of-its-kind, in-the-wild, fine-grained workout/exercise posture analysis dataset, covering three different exercises: BackSquat, Barbell Row, and Overhead Press.
3 papers · 0 benchmarks
HSPACE (Human-SPACE) is a large-scale photo-realistic dataset of animated humans placed in complex synthetic indoor and outdoor environments.
3 papers · 1 benchmark
Includes 100K depth images under challenging scenarios.
3 papers · 2 benchmarks
Mirrored-Human is a dataset for 3D pose estimation from a single view.
3 papers · 0 benchmarks
VBR (VBR: A Vision Benchmark in Rome)
This dataset presents a vision and perception research dataset collected in Rome, featuring RGB data, 3D point clouds, IMU, and GPS data.
3 papers · 0 benchmarks
DRACO20K dataset is used for evaluating object canonicalization on methods that estimate a canonical frame from a monocular input image.
2 papers · 0 benchmarks
Estimating camera motion in deformable scenes poses a complex and open research challenge.
2 papers · 1 benchmark
Event-Human3.6m is a challenging dataset for event-based human pose estimation by simulating events from the RGB Human3.6m dataset.
2 papers · 0 benchmarks
A Simulated Benchmark for multi-modal SLAM Systems Evaluation in Large-scale Dynamic Environments.
2 papers · 0 benchmarks
This dataset comprehends the 3D building information model (in IFC and Revit formats), manually elaborated based on the terrestrial laser scanner of the sequence 2 of ConSLAM, and the refined ground truth (GT) poses (in TUM format) of…
2 papers · 0 benchmarks
Accidental Turntables contains a challenging set of 41,212 images of cars in cluttered backgrounds, motion blur and illumination changes that serves as a benchmark for 3D pose estimation.
1 paper · 0 benchmarks
ConSLAM (Construction Dataset for SLAM)
ConSLAM is a real-world dataset collected periodically on a construction site to measure the accuracy of mobile scanners' SLAM algorithms.
1 paper · 0 benchmarks
A new large-scale dataset that consists of 409 fine-grained categories and 31,881 images with accurate 3D pose annotation.
1 paper · 0 benchmarks
Human pose estimation (HPE) in the top-view using fisheye cameras presents a promising and innovative application domain.
1 paper · 0 benchmarks
The Store Dataset is a dataset for estimating 3D poses of multiple humans in real-time.
1 paper · 0 benchmarks
This is a pose estimation dataset, consisting of symmetric 3D shapes where multiple orientations are visually indistinguishable.
1 paper · 0 benchmarks
InfiniteRep is a synthetic, open-source dataset for fitness and physical therapy (PT) applications.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.