Browse State-of-the-Art › 3D Semantic Occupancy Prediction
3D Semantic Occupancy Prediction
15 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Uses sparse LiDAR semantic labels for training and testing
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
15 shown of 15 papers with code (47 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
26 Jun 2025 1 repository listedWe introduce OccOoD, a novel framework integrating OoD detection into 3D semantic occupancy prediction, with Voxel-BEV Progressive Fusion (VBPF) leveraging an RWKV-based branch to enhance OoD detection via…
-
12 Jun 2025 1 repository listedBuilding on this, we present QuadricFormer, a superquadric-based model for efficient 3D occupancy prediction, and introduce a pruning-and-splitting module to further enhance modeling efficiency by concentrating…
-
Rethinking Temporal Fusion with a Unified Gradient Descent View for 3D Semantic Occupancy Prediction17 Apr 2025 1 repository listedWe present GDFusion, a temporal fusion method for vision-based 3D semantic occupancy prediction (VisionOcc).
-
28 Jan 2025 1 repository listedTo utilize these slice features, we propose SliceOcc, an RGB camera-based model specifically tailored for indoor 3D semantic occupancy prediction.
-
17 Dec 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedIn this paper, we introduce GaussTR, a novel Gaussian Transformer that leverages alignment with foundation models to advance self-supervised 3D spatial understanding.
-
5 Dec 2024 1 repository listed Syntology ran 1 of 4 samples · 3 unverified · 4 pointer-only (licence)To address this, we propose a probabilistic Gaussian superposition model which interprets each Gaussian as a probability distribution of its neighborhood being occupied and conforms to probabilistic multiplication to…
-
19 Nov 2024 1 repository listedRecent methods are mainly built on the 2D-to-3D transformation that relies on sensor calibration to project the 2D image information into the 3D space.
-
12 Nov 2024 1 repository listedIn this work, we strive to improve performance by introducing a series of targeted improvements for 3D semantic occupancy prediction and flow estimation.
-
30 Sep 2024 1 repository listedMulti-sensor fusion significantly enhances the accuracy and robustness of 3D semantic occupancy prediction, which is crucial for autonomous driving and robotics.
-
27 May 2024 1 repository listedTo address this, we propose an object-centric representation to describe 3D scenes with sparse 3D semantic Gaussians where each Gaussian represents a flexible region of interest and its semantic features.
-
12 Mar 2024 1 repository listedHyDRa achieves a new state-of-the-art for camera-radar fusion of 64.
-
3 Mar 2024 1 repository listedA comprehensive understanding of 3D scenes is crucial in autonomous vehicles (AVs), and recent models for 3D semantic occupancy prediction have successfully addressed the challenge of describing real-world objects with…
-
23 Jan 2024 1 repository listedIn contrast, our approach leverages two projection matrices to store the static mapping relationships and matrix multiplications to efficiently generate global Bird's Eye View (BEV) features and local 3D feature volumes.
-
31 Aug 2023 1 repository listedTo address this, we propose a cylindrical tri-perspective view to represent point clouds effectively and comprehensively and a PointOcc model to process them efficiently.
-
11 Apr 2023 1 repository listed Syntology ran 1 of 3 samples · 2 unverifiedThe vision-based perception for autonomous driving has undergone a transformation from the bird-eye-view (BEV) representations to the 3D semantic occupancy.
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections