{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/hdnet-human-depth-estimation-for-multi-person","title":"HDNet: Human Depth Estimation for Multi-Person Camera-Space Localization","arxiv_id":"2007.08943","date":"2020-07-17","proceeding":"ECCV 2020 8","authors":["Jiahao Lin","Gim Hee Lee"],"abstract":"Current works on multi-person 3D pose estimation mainly focus on the estimation of the 3D joint locations relative to the root joint and ignore the absolute locations of each pose. In this paper, we propose the Human Depth Estimation Network (HDNet), an end-to-end framework for absolute root joint localization in the camera coordinate space. Our HDNet first estimates the 2D human pose with heatmaps of the joints. These estimated heatmaps serve as attention masks for pooling features from image regions corresponding to the target person. A skeleton-based Graph Neural Network (GNN) is utilized to propagate features among joints. We formulate the target depth regression as a bin index estimation problem, which can be transformed with a soft-argmax operation from the classification output of our HDNet. We evaluate our HDNet on the root joint localization and root-relative 3D pose estimation tasks with two benchmark datasets, i.e., Human3.6M and MuPoTS-3D. The experimental results show that we outperform the previous state-of-the-art consistently under multiple evaluation metrics. Our source code is available at: https://github.com/jiahaoLjh/HumanDepth.","url_abs":"https://arxiv.org/abs/2007.08943v1","url_pdf":"https://arxiv.org/pdf/2007.08943v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"hdnet-human-depth-estimation-for-multi-person","repo_url":"https://github.com/jiahaoLjh/HumanDepth","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"3d-multi-person-pose-estimation-absolute","task_name":"3D Multi-Person Pose Estimation (absolute)"},{"task_slug":"3d-multi-person-pose-estimation-root-relative","task_name":"3D Multi-Person Pose Estimation (root-relative)"},{"task_slug":"3d-pose-estimation","task_name":"3D Pose Estimation"},{"task_slug":"depth-estimation","task_name":"Depth Estimation"},{"task_slug":"graph-neural-network","task_name":"Graph Neural Network"},{"task_slug":"pose-estimation","task_name":"Pose Estimation"},{"task_slug":"root-joint-localization","task_name":"Root Joint Localization"}],"methods":[{"method_slug":"graph-neural-network","method_name":"Graph Neural Network"}],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/3d-multi-person-pose-estimation-absolute-on","task":"3D Multi-Person Pose Estimation (absolute)","dataset":"MuPoTS-3D","model":"HDNet","rank_in_archive_order":12,"of":14,"metrics":{"3DPCK":"35.2"},"uses_additional_data":false},{"leaderboard":"/sota/3d-multi-person-pose-estimation-root-relative","task":"3D Multi-Person Pose Estimation (root-relative)","dataset":"MuPoTS-3D","model":"HDNet","rank_in_archive_order":7,"of":20,"metrics":{"3DPCK":"83.7"},"uses_additional_data":false}],"syntology":{"syntology_url":null,"atlas_url":"https://app.syntology.ai/?focus=2007.08943","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}