{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/vehicle-pose-and-shape-estimation-through","title":"Vehicle Pose and Shape Estimation through Multiple Monocular Vision","arxiv_id":"1802.03515","date":"2018-02-10","proceeding":null,"authors":["Wenhao Ding","Shuaijun Li","Guilin Zhang","Xiangyu Lei","Huihuan Qian"],"abstract":"In this paper, we present an accurate approach to estimate vehicles' pose and\nshape from off-board multiview images. The images are taken by monocular\ncameras and have small overlaps. We utilize state-of-the-art convolutional\nneural networks (CNNs) to extract vehicles' semantic keypoints and introduce a\nCross Projection Optimization (CPO) method to estimate the 3D pose. During the\niterative CPO process, an adaptive shape adjustment method named Hierarchical\nWireframe Constraint (HWC) is implemented to estimate the shape. Our approach\nis evaluated under both simulated and real-world scenes for performance\nverification. It's shown that our algorithm outperforms other existing\nmonocular and stereo methods for vehicles' pose and shape estimation. This\napproach provides a new and robust solution for off-board visual vehicle\nlocalization and tracking, which can be applied to massive surveillance camera\nnetworks for intelligent transportation.","url_abs":"http://arxiv.org/abs/1802.03515v5","url_pdf":"http://arxiv.org/pdf/1802.03515v5.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"vehicle-pose-and-shape-estimation-through","repo_url":"https://github.com/GilgameshD/Multiple-View-Car-Localization","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":null}],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}