{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/probabilistic-3d-multi-modal-multi-object","title":"Probabilistic 3D Multi-Modal, Multi-Object Tracking for Autonomous Driving","arxiv_id":"2012.13755","date":"2020-12-26","proceeding":null,"authors":["Hsu-kuang Chiu","Jie Li","Rares Ambrus","Jeannette Bohg"],"abstract":"Multi-object tracking is an important ability for an autonomous vehicle to safely navigate a traffic scene. Current state-of-the-art follows the tracking-by-detection paradigm where existing tracks are associated with detected objects through some distance metric. The key challenges to increase tracking accuracy lie in data association and track life cycle management. We propose a probabilistic, multi-modal, multi-object tracking system consisting of different trainable modules to provide robust and data-driven tracking results. First, we learn how to fuse features from 2D images and 3D LiDAR point clouds to capture the appearance and geometric information of an object. Second, we propose to learn a metric that combines the Mahalanobis and feature distances when comparing a track and a new detection in data association. And third, we propose to learn when to initialize a track from an unmatched object detection. Through extensive quantitative and qualitative results, we show that when using the same object detectors our method outperforms state-of-the-art approaches on the NuScenes and KITTI datasets.","url_abs":"https://arxiv.org/abs/2012.13755v2","url_pdf":"https://arxiv.org/pdf/2012.13755v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"probabilistic-3d-multi-modal-multi-object","repo_url":"https://github.com/eddyhkchiu/mahalanobis_3d_multi_object_tracking","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok","spdx":"NOASSERTION"}}],"tasks":[{"task_slug":"3d-pedestrian-tracking","task_name":"3D Pedestrian Tracking"},{"task_slug":"autonomous-driving","task_name":"Autonomous Driving"},{"task_slug":"management","task_name":"Management"},{"task_slug":"multi-object-tracking","task_name":"Multi-Object Tracking"},{"task_slug":"navigate","task_name":"Navigate"},{"task_slug":"object","task_name":"Object"},{"task_slug":"object-detection","task_name":"Object Detection"},{"task_slug":"object-tracking","task_name":"Object Tracking"},{"task_slug":"object-detection-1","task_name":"object-detection"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/3d-pedestrian-tracking-on-nuscenes-validation","task":"3D Pedestrian Tracking","dataset":"nuScenes validation set","model":"Probabilistic 3D Tracking","rank_in_archive_order":3,"of":4,"metrics":{"AMOTA":"76.6"},"uses_additional_data":false}],"syntology":{"syntology_url":"https://syntology.ai/paper/2012.13755","atlas_url":"https://app.syntology.ai/?focus=2012.13755","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}