Papers › Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network
Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network
Duy-Tho Le, Hengcan Shi, Hamid Rezatofighi, Jianfei Cai
Efficiently and accurately detecting people from 3D point cloud data is of great importance in many robotic and autonomous driving applications. This fundamental perception task is still very challenging due to (i) significant deformations of human body pose and gesture over time and (ii) point cloud sparsity and scarcity for pedestrian class objects. Recent efficient 3D object detection approaches rely on pillar features to detect objects from point cloud data. However, these pillar features do not carry sufficient expressive representations to deal with all the aforementioned challenges in detecting people. To address this shortcoming, we first introduce a stackable Pillar Aware Attention (PAA) module for enhanced pillar features extraction while suppressing noises in the point clouds. By integrating multi-point-channel-pooling, point-wise, channel-wise, and task-aware attention into a simple module, the representation capabilities are boosted while requiring little additional computing resources. We also present Mini-BiFPN, a small yet effective feature network that creates bidirectional information flow and multi-level cross-scale feature fusion to better integrate multi-resolution features. Our proposed framework, namely PiFeNet, has been evaluated on three popular large-scale datasets for 3D pedestrian Detection, i.e. KITTI, JRDB, and nuScenes achieving state-of-the-art (SOTA) performance on KITTI Bird-eye-view (BEV) and JRDB and very competitive performance on nuScenes. Our approach has inference speed of 26 frame-per-second (FPS), making it a real-time detector. The code for our PiFeNet is available at https://github.com/ldtho/PiFeNet.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| 3D Object Detection | KITTI Pedestrian | PiFeNet | mAP | 0.486 | #1 of 1 | Archive leaderboard | report |
| 3D Object Detection | KITTI Pedestrian Easy | PiFeNet | Average Precision | 0.5639 | #1 of 1 | Archive leaderboard | report |
| 3D Object Detection | KITTI Pedestrian Hard | PiFeNet | Average Precision | 0.4271 | #1 of 1 | Archive leaderboard | report |
| 3D Object Detection | KITTI Pedestrian Moderate | PiFeNet | Average Precision | 0.4671 | #1 of 1 | Archive leaderboard | report |
| Birds Eye View Object Detection | KITTI Pedestrian | PiFeNet | mAP | 0.559 | #1 of 1 | Archive leaderboard | report |
| Birds Eye View Object Detection | KITTI Pedestrian Easy | PiFeNet | Average Precision | 0.6325 | #1 of 1 | Archive leaderboard | report |
| Birds Eye View Object Detection | KITTI Pedestrian Hard | PiFeNet | Average Precision | 0.5053 | #1 of 1 | Archive leaderboard | report |
| Birds Eye View Object Detection | KITTI Pedestrian Moderate | PiFeNet | Average Precision | 0.5392 | #1 of 1 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Methods
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections