Papers › 3D Dual-Fusion: Dual-Domain Dual-Query Camera-LiDAR Fusion for 3D Object Detection
3D Dual-Fusion: Dual-Domain Dual-Query Camera-LiDAR Fusion for 3D Object Detection
Yecheol Kim, Konyul Park, Minwook Kim, Dongsuk Kum, Jun Won Choi
Fusing data from cameras and LiDAR sensors is an essential technique to achieve robust 3D object detection. One key challenge in camera-LiDAR fusion involves mitigating the large domain gap between the two sensors in terms of coordinates and data distribution when fusing their features. In this paper, we propose a novel camera-LiDAR fusion architecture called, 3D Dual-Fusion, which is designed to mitigate the gap between the feature representations of camera and LiDAR data. The proposed method fuses the features of the camera-view and 3D voxel-view domain and models their interactions through deformable attention. We redesign the transformer fusion encoder to aggregate the information from the two domains. Two major changes include 1) dual query-based deformable attention to fuse the dual-domain features interactively and 2) 3D local self-attention to encode the voxel-domain queries prior to dual-query decoding. The results of an experimental evaluation show that the proposed camera-LiDAR fusion architecture achieved competitive performance on the KITTI and nuScenes datasets, with state-of-the-art performances in some 3D object detection benchmarks categories.
In Syntology View this paper on Syntology: its repositories, every harvested function with whether it ran, its licence and the call to fetch it.
Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| 3D Object Detection | KITTI Cars Easy | 3D Dual-Fusion | AP | 91.01% | #5 of 26 | Archive leaderboard | report |
| 3D Object Detection | KITTI Cars Hard | 3D Dual-Fusion | AP | 79.39% | #2 of 25 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | NDS | 0.73 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mAAE | 0.13 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mAOE | 0.33 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mAP | 0.71 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mASE | 0.24 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mATE | 0.26 | #29 of 372 | Archive leaderboard | report |
| 3D Object Detection | nuScenes | 3D Dual-Fusion_T | mAVE | 0.27 | #29 of 372 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections