Papers › Real-Time Multi-View 3D Human Pose Estimation using Semantic Feedback to Smart Edge Sensors
Real-Time Multi-View 3D Human Pose Estimation using Semantic Feedback to Smart Edge Sensors
Simon Bultmann, Sven Behnke
We present a novel method for estimation of 3D human poses from a multi-camera setup, employing distributed smart edge sensors coupled with a backend through a semantic feedback loop. 2D joint detection for each camera view is performed locally on a dedicated embedded inference processor. Only the semantic skeleton representation is transmitted over the network and raw images remain on the sensor board. 3D poses are recovered from 2D joints on a central backend, based on triangulation and a body model which incorporates prior knowledge of the human skeleton. A feedback channel from backend to individual sensors is implemented on a semantic level. The allocentric 3D pose is backprojected into the sensor views where it is fused with 2D joint detections. The local semantic model on each sensor can thus be improved by incorporating global context information. The whole pipeline is capable of real-time operation. We evaluate our method on three public datasets, where we achieve state-of-the-art results and show the benefits of our feedback architecture, as well as in our own setup for multi-person experiments. Using the feedback signal improves the 2D joint detections and in turn the estimated 3D poses.
In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| 3D Human Pose Estimation | Human3.6M | SmartEdgeSensor | Average MPJPE (mm) | 29.8 | #7 of 88 | Archive leaderboard | report |
| 3D Human Pose Estimation | Human3.6M | SmartEdgeSensor | Multi-View or Monocular | Multi-View | #7 of 88 | Archive leaderboard | report |
| 3D Human Pose Estimation | Human3.6M | SmartEdgeSensor | Using 2D ground-truth joints | No | #7 of 88 | Archive leaderboard | report |
| 3D Multi-Person Pose Estimation | Campus | SmartEdgeSensor | PCP3D | 97 | #2 of 16 | Archive leaderboard | report |
| 3D Multi-Person Pose Estimation | Shelf | SmartEdgeSensor | PCP3D | 97.4 | #11 of 27 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections