Papers › Robust 6DoF Pose Estimation Against Depth Noise and a Comprehensive Evaluation on a...
Robust 6DoF Pose Estimation Against Depth Noise and a Comprehensive Evaluation on a Mobile Dataset
Zixun Huang, Keling Yao, Seth Z. Zhao, Chuanyu Pan, Chenfeng Xu, Kathy Zhuang, Tianjian Xu, Weiyu Feng, Allen Y. Yang
Robust 6DoF pose estimation with mobile devices is the foundation for applications in robotics, augmented reality, and digital twin localization. In this paper, we extensively investigate the robustness of existing RGBD-based 6DoF pose estimation methods against varying levels of depth sensor noise. We highlight that existing 6DoF pose estimation methods suffer significant performance discrepancies due to depth measurement inaccuracies. In response to the robustness issue, we present a simple and effective transformer-based 6DoF pose estimation approach called DTTDNet, featuring a novel geometric feature filtering module and a Chamfer distance loss for training. Moreover, we advance the field of robust 6DoF pose estimation and introduce a new dataset -- Digital Twin Tracking Dataset Mobile (DTTD-Mobile), tailored for digital twin object tracking with noisy depth data from the mobile RGBD sensor suite of the Apple iPhone 14 Pro. Extensive experiments demonstrate that DTTDNet significantly outperforms state-of-the-art methods at least 4.32, up to 60.74 points in ADD metrics on the DTTD-Mobile. More importantly, our approach exhibits superior robustness to varying levels of measurement noise, setting a new benchmark for the robustness to noise measurements. Code and dataset are made publicly available at: https://github.com/augcog/DTTD2
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Datasets
Introduced by this paper, per the archive.
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| 3D Object Detection | DTTD-Mobile | DTTDNet | ADD AUC | 73.99 | #1 of 5 | Archive leaderboard | report |
| 3D Object Detection | DTTD-Mobile | DTTDNet | ADD-S AUC | 88.10 | #1 of 5 | Archive leaderboard | report |
| 6D Pose Estimation | DTTD-Mobile | DTTDNet | ADD AUC | 73.99 | #1 of 8 | Archive leaderboard | report |
| 6D Pose Estimation | DTTD-Mobile | DTTDNet | ADD-S AUC | 88.10 | #1 of 8 | Archive leaderboard | report |
| 6D Pose Estimation | YCB-Video | DTTD-Net w/o refiner | ADDS AUC | 94.19 | #7 of 10 | Archive leaderboard | report |
| 6D Pose Estimation using RGBD | YCB-Video | DTTDNet | ADD-S (2cm) | 96.14 | #9 of 9 | Archive leaderboard | report |
| 6D Pose Estimation using RGBD | YCB-Video | DTTDNet | ADD-S AUC | 94.19 | #9 of 9 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections