{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/roarnet-a-robust-3d-object-detection-based-on","title":"RoarNet: A Robust 3D Object Detection based on RegiOn Approximation Refinement","arxiv_id":"1811.03818","date":"2018-11-09","proceeding":null,"authors":["Kiwoo Shin","Youngwook Paul Kwon","Masayoshi Tomizuka"],"abstract":"We present RoarNet, a new approach for 3D object detection from a 2D image\nand 3D Lidar point clouds. Based on two-stage object detection framework with\nPointNet as our backbone network, we suggest several novel ideas to improve 3D\nobject detection performance. The first part of our method, RoarNet_2D,\nestimates the 3D poses of objects from a monocular image, which approximates\nwhere to examine further, and derives multiple candidates that are\ngeometrically feasible. This step significantly narrows down feasible 3D\nregions, which otherwise requires demanding processing of 3D point clouds in a\nhuge search space. Then the second part, RoarNet_3D, takes the candidate\nregions and conducts in-depth inferences to conclude final poses in a recursive\nmanner. Inspired by PointNet, RoarNet_3D processes 3D point clouds directly\nwithout any loss of data, leading to precise detection. We evaluate our method\nin KITTI, a 3D object detection benchmark. Our result shows that RoarNet has\nsuperior performance to state-of-the-art methods that are publicly available.\nRemarkably, RoarNet also outperforms state-of-the-art methods even in settings\nwhere Lidar and camera are not time synchronized, which is practically\nimportant for actual driving environments. RoarNet is implemented in Tensorflow\nand publicly available with pre-trained models.","url_abs":"http://arxiv.org/abs/1811.03818v1","url_pdf":"http://arxiv.org/pdf/1811.03818v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"3d-object-detection","task_name":"3D Object Detection"},{"task_slug":"object","task_name":"Object"},{"task_slug":"object-detection","task_name":"Object Detection"},{"task_slug":"robust-3d-object-detection","task_name":"Robust 3D Object Detection"},{"task_slug":"object-detection-1","task_name":"object-detection"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/3d-object-detection-on-kitti-cars-easy","task":"3D Object Detection","dataset":"KITTI Cars Easy","model":"RoarNet","rank_in_archive_order":20,"of":26,"metrics":{"AP":"83.71%"},"uses_additional_data":false},{"leaderboard":"/sota/3d-object-detection-on-kitti-cars-hard","task":"3D Object Detection","dataset":"KITTI Cars Hard","model":"RoarNet","rank_in_archive_order":23,"of":25,"metrics":{"AP":"59.16%"},"uses_additional_data":false},{"leaderboard":"/sota/object-detection-on-kitti-cars-easy","task":"Object Detection","dataset":"KITTI Cars Easy","model":"Roarnet","rank_in_archive_order":3,"of":5,"metrics":{"AP":"83.71"},"uses_additional_data":false}],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=1811.03818","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}