Papers › On Boosting Single-Frame 3D Human Pose Estimation via Monocular Videos

On Boosting Single-Frame 3D Human Pose Estimation via Monocular Videos

1 Oct 2019ICCV 2019 10archive 2025-07-28

Zhi Li, Xuan Wang, Fei Wang, Peilin Jiang

The premise of training an accurate 3D human pose estimation network is the possession of huge amount of richly annotated training data. Nonetheless, manually obtaining rich and accurate annotations is, even not impossible, tedious and slow. In this paper, we propose to exploit monocular videos to complement the training dataset for the single-image 3D human pose estimation tasks. At the beginning, a baseline model is trained with a small set of annotations. By fixing some reliable estimations produced by the resulting model, our method automatically collects the annotations across the entire video as solving the 3D trajectory completion problem. Then, the baseline model is further trained with the collected annotations to learn the new poses. We evaluate our method on the broadly-adopted Human3.6M and MPI-INF-3DHP datasets. As illustrated in experiments, given only a small set of annotations, our method successfully makes the model to learn new poses from unlabelled monocular videos, promoting the accuracies of the baseline model by about 10%. By contrast with previous approaches, our method does not rely on either multi-view imagery or any explicit 2D keypoint annotations.

PaperPDF

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

3D Human Pose EstimationPose EstimationWeakly-supervised 3D Human Pose Estimation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Weakly-supervised 3D Human Pose Estimation Human3.6M Li et al. 3D Annotations S1 #23 of 33 Archive leaderboard report
Weakly-supervised 3D Human Pose Estimation Human3.6M Li et al. Average MPJPE (mm) 88.8 #23 of 33 Archive leaderboard report
Weakly-supervised 3D Human Pose Estimation Human3.6M Li et al. Number of Views 1 #23 of 33 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections