Papers › Lifting from the Deep: Convolutional 3D Pose Estimation from a Single Image
Lifting from the Deep: Convolutional 3D Pose Estimation from a Single Image
Denis Tome, Chris Russell, Lourdes Agapito
We propose a unified formulation for the problem of 3D human pose estimation from a single raw RGB image that reasons jointly about 2D joint estimation and 3D pose reconstruction to improve both tasks. We take an integrated approach that fuses probabilistic knowledge of 3D human pose with a multi-stage CNN architecture and uses the knowledge of plausible 3D landmark locations to refine the search for better 2D locations. The entire process is trained end-to-end, is extremely efficient and obtains state- of-the-art results on Human3.6M outperforming previous approaches both on 2D and 3D errors.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Monocular 3D Human Pose Estimation | Human3.6M | Projected-pose belief maps + 2D fusion layers | Average MPJPE (mm) | 88.39 | #37 of 52 | Archive leaderboard | report |
| Monocular 3D Human Pose Estimation | Human3.6M | Projected-pose belief maps + 2D fusion layers | Frames Needed | 1 | #37 of 52 | Archive leaderboard | report |
| Monocular 3D Human Pose Estimation | Human3.6M | Projected-pose belief maps + 2D fusion layers | Need Ground Truth 2D Pose | No | #37 of 52 | Archive leaderboard | report |
| Monocular 3D Human Pose Estimation | Human3.6M | Projected-pose belief maps + 2D fusion layers | Use Video Sequence | No | #37 of 52 | Archive leaderboard | report |
| Weakly-supervised 3D Human Pose Estimation | Human3.6M | Tome et al. | 3D Annotations | No | #22 of 33 | Archive leaderboard | report |
| Weakly-supervised 3D Human Pose Estimation | Human3.6M | Tome et al. | Average MPJPE (mm) | 88.4 | #22 of 33 | Archive leaderboard | report |
| Weakly-supervised 3D Human Pose Estimation | Human3.6M | Tome et al. | Number of Frames Per View | 1 | #22 of 33 | Archive leaderboard | report |
| Weakly-supervised 3D Human Pose Estimation | Human3.6M | Tome et al. | Number of Views | 1 | #22 of 33 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections