Browse State-of-the-Art › Scene Understanding › Papers, page 11
Scene Understanding
Papers archive 2025-07-28
archive papers tagged: 1,723 · with a code link: 720 · where Syntology ran a sample: 208 (182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (208 of 1,723 tagged: 182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 11 of 18: papers 1,001 to 1,100 of 1,723, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
BOX3D: Lightweight Camera-LiDAR Fusion for 3D Object Detection and Localization27 Aug 2024 0 repositories listed
-
Interactive Occlusion Boundary Estimation through Exploitation of Synthetic Data27 Aug 2024 0 repositories listed
-
FusionSAM: Latent Space driven Segment Anything Model for Multimodal Fusion and Segmentation26 Aug 2024 0 repositories listed
-
3D-VirtFusion: Synthetic 3D Data Augmentation through Generative Diffusion Models and Controllable Editing25 Aug 2024 0 repositories listed
-
Making Large Language Models Better Planners with Reasoning-Decision Alignment25 Aug 2024 0 repositories listed
-
Near, far: Patch-ordering enhances vision foundation models' scene understanding20 Aug 2024 0 repositories listed
-
3D-Aware Instance Segmentation and Tracking in Egocentric Videos19 Aug 2024 0 repositories listed
-
SceneGPT: A Language Model for 3D Scene Understanding13 Aug 2024 0 repositories listed
-
SpectralGaussians: Semantic, spectral 3D Gaussian splatting for multi-spectral scene representation, visualization and analysis13 Aug 2024 0 repositories listed
-
HeLiMOS: A Dataset for Moving Object Segmentation in 3D Point Clouds From Heterogeneous LiDAR Sensors12 Aug 2024 0 repositories listed
-
Spherical World-Locking for Audio-Visual Localization in Egocentric Videos9 Aug 2024 0 repositories listed
-
1 Aug 2024 0 repositories listed
-
DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations31 Jul 2024 0 repositories listed
-
NIS-SLAM: Neural Implicit Semantic RGB-D SLAM for 3D Consistent Scene Understanding30 Jul 2024 0 repositories listed
-
Rethinking RGB-D Fusion for Semantic Segmentation in Surgical Datasets29 Jul 2024 0 repositories listed
-
GP-VLS: A general-purpose vision language model for surgery27 Jul 2024 0 repositories listed
-
Answerability Fields: Answerable Location Estimation via Diffusion Models26 Jul 2024 0 repositories listed
-
3D Question Answering for City Scene Understanding24 Jul 2024 0 repositories listed
-
Augmented Efficiency: Reducing Memory Footprint and Accelerating Inference for 3D Semantic Segmentation through Hybrid Vision23 Jul 2024 0 repositories listed
-
InLUT3D: Challenging real indoor dataset for point cloud analysis22 Jul 2024 0 repositories listed
-
VideoGameBunny: Towards vision assistants for video games21 Jul 2024 0 repositories listed
-
GaussianBeV: 3D Gaussian Representation meets Perception Models for BeV Segmentation19 Jul 2024 0 repositories listed
-
OpenSU3D: Open World 3D Scene Understanding using Foundation Models19 Jul 2024 0 repositories listed
-
18 Jul 2024 0 repositories listed Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Training-Free Model Merging for Multi-target Domain Adaptation18 Jul 2024 0 repositories listed
-
Benchmarking Vision Language Models for Cultural Understanding15 Jul 2024 0 repositories listed
-
Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding13 Jul 2024 0 repositories listed
-
BLOS-BEV: Navigation Map Enhanced Lane Segmentation Network, Beyond Line of Sight11 Jul 2024 0 repositories listed
-
10 Jul 2024 0 repositories listed Syntology 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 11 harvested samples)
-
Joint prototype and coefficient prediction for 3D instance segmentation9 Jul 2024 0 repositories listed
-
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition9 Jul 2024 0 repositories listed
-
Self-supervised Learning via Cluster Distance Prediction for Operating Room Context Awareness7 Jul 2024 0 repositories listed
-
Hybrid Primal Sketch: Combining Analogy, Qualitative Representations, and Computer Vision for Scene Understanding5 Jul 2024 0 repositories listed
-
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction1 Jul 2024 0 repositories listed
-
ESGNN: Towards Equivariant Scene Graph Neural Network for 3D Scene Understanding30 Jun 2024 0 repositories listed
-
EgoGaussian: Dynamic Scene Understanding from Egocentric Video with 3D Gaussian Splatting28 Jun 2024 0 repositories listed
-
PPTFormer: Pseudo Multi-Perspective Transformer for UAV Segmentation28 Jun 2024 0 repositories listed
-
3D-MVP: 3D Multiview Pretraining for Robotic Manipulation26 Jun 2024 0 repositories listed
-
GPT-4V Explorations: Mining Autonomous Driving24 Jun 2024 0 repositories listed
-
EvSegSNN: Neuromorphic Semantic Segmentation for Event Data20 Jun 2024 0 repositories listed
-
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features17 Jun 2024 0 repositories listed
-
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding17 Jun 2024 0 repositories listed
-
MapVision: CVPR 2024 Autonomous Grand Challenge Mapless Driving Tech Report14 Jun 2024 0 repositories listed
-
FastLGS: Speeding up Language Embedded Gaussians with Feature Grid Mapping4 Jun 2024 0 repositories listed
-
CYCLO: Cyclic Graph Transformer Approach to Multi-Object Relationship Modeling in Aerial Videos3 Jun 2024 0 repositories listed
-
EAGLE: Efficient Adaptive Geometry-based Learning in Cross-view Understanding3 Jun 2024 0 repositories listed
-
Object Aware Egocentric Online Action Detection3 Jun 2024 0 repositories listed
-
Semi-supervised Video Semantic Segmentation Using Unreliable Pseudo Labels for PVUW20242 Jun 2024 0 repositories listed
-
Learning 3D Robotics Perception using Inductive Priors30 May 2024 0 repositories listed
-
30 May 2024 0 repositories listed
-
Kestrel: Point Grounding Multimodal LLM for Part-Aware 3D Vision-Language Understanding29 May 2024 0 repositories listed
-
GOI: Find 3D Gaussians of Interest with an Optimizable Open-vocabulary Semantic-space Hyperplane27 May 2024 0 repositories listed
-
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding24 May 2024 0 repositories listed
-
Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis23 May 2024 0 repositories listed
-
Transformers for Image-Goal Navigation23 May 2024 0 repositories listed
-
GameVLM: A Decision-making Framework for Robotic Task Planning Based on Visual Language Models and Zero-sum Games22 May 2024 0 repositories listed
-
TS40K: a 3D Point Cloud Dataset of Rural Terrain and Electrical Transmission System22 May 2024 0 repositories listed
-
Anticipating Object State Changes in Long Procedural Videos21 May 2024 0 repositories listed
-
A Preprocessing and Postprocessing Voxel-based Method for LiDAR Semantic Segmentation Improvement in Long Distance16 May 2024 0 repositories listed
-
3D Shape Augmentation with Content-Aware Shape Resizing15 May 2024 0 repositories listed
-
BEHAVIOR Vision Suite: Customizable Dataset Generation via Simulation15 May 2024 0 repositories listed
-
DriveWorld: 4D Pre-trained Scene Understanding via World Models for Autonomous Driving7 May 2024 0 repositories listed
-
Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM29 Apr 2024 0 repositories listed
-
Seeing Beyond Classes: Zero-Shot Grounded Situation Recognition via Language Explainer24 Apr 2024 0 repositories listed
-
CloudFort: Enhancing Robustness of 3D Point Cloud Classification Against Backdoor Attacks via Spatial Partitioning and Ensemble Prediction22 Apr 2024 0 repositories listed
-
On Support Relations Inference and Scene Hierarchy Graph Construction from Point Cloud in Clustered Environments22 Apr 2024 0 repositories listed
-
Unified Scene Representation and Reconstruction for 3D Large Language Models19 Apr 2024 0 repositories listed
-
AccidentBlip: Agent of Accident Warning based on MA-former18 Apr 2024 0 repositories listed
-
Multimodal 3D Object Detection on Unseen Domains17 Apr 2024 0 repositories listed
-
PreGSU-A Generalized Traffic Scene Understanding Model for Autonomous Driving based on Pre-trained Graph Attention Network16 Apr 2024 0 repositories listed
-
Depth Estimation using Weighted-loss and Transfer Learning11 Apr 2024 0 repositories listed
-
Gaga: Group Any Gaussians via 3D-aware Memory Bank11 Apr 2024 0 repositories listed
-
Incorporating Explanations into Human-Machine Interfaces for Trust and Situation Awareness in Autonomous Vehicles10 Apr 2024 0 repositories listed
-
O2V-Mapping: Online Open-Vocabulary Mapping with Neural Implicit Representation10 Apr 2024 0 repositories listed
-
DaF-BEVSeg: Distortion-aware Fisheye Camera based Bird's Eye View Segmentation with Occlusion Reasoning9 Apr 2024 0 repositories listed
-
QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding9 Apr 2024 0 repositories listed
-
Panoptic Perception: A Novel Task and Fine-grained Dataset for Universal Remote Sensing Image Interpretation6 Apr 2024 0 repositories listed
-
You Only Scan Once: A Dynamic Scene Reconstruction Pipeline for 6-DoF Robotic Grasping of Novel Objects4 Apr 2024 0 repositories listed
-
1 Apr 2024 0 repositories listed
-
MM3DGS SLAM: Multi-modal 3D Gaussian Splatting for SLAM Using Vision, Depth, and Inertial Measurements1 Apr 2024 0 repositories listed
-
Adapting to Length Shift: FlexiLength Network for Trajectory Prediction31 Mar 2024 0 repositories listed
-
Neural Radiance Field-based Visual Rendering: A Comprehensive Review31 Mar 2024 0 repositories listed
-
HGS-Mapping: Online Dense Mapping Using Hybrid Gaussian Representation in Urban Scenes29 Mar 2024 0 repositories listed
-
Efficient 3D Instance Mapping and Localization with Neural Fields28 Mar 2024 0 repositories listed
-
Towards Trustworthy Automated Driving through Qualitative Scene Understanding and Explanations25 Mar 2024 0 repositories listed
-
Multi-Task Learning with Multi-Task Optimization24 Mar 2024 0 repositories listed
-
Semantic Is Enough: Only Semantic Information For NeRF Reconstruction24 Mar 2024 0 repositories listed
-
DiffusionMTL: Learning Multi-Task Denoising Diffusion Model from Partially Annotated Data22 Mar 2024 0 repositories listed
-
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting22 Mar 2024 0 repositories listed
-
3D Object Detection from Point Cloud via Voting Step Diffusion21 Mar 2024 0 repositories listed
-
Exosense: A Vision-Based Scene Understanding System For Exoskeletons21 Mar 2024 0 repositories listed
-
SurroundSDF: Implicit 3D Scene Understanding Based on Signed Distance Field21 Mar 2024 0 repositories listed
-
Geometric Constraints in Deep Learning Frameworks: A Survey19 Mar 2024 0 repositories listed
-
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting19 Mar 2024 0 repositories listed
-
M2DA: Multi-Modal Fusion Transformer Incorporating Driver Attention for Autonomous Driving19 Mar 2024 0 repositories listed
-
Agent3D-Zero: An Agent for Zero-shot 3D Understanding18 Mar 2024 0 repositories listed
-
R3DS: Reality-linked 3D Scenes for Panoramic Scene Understanding18 Mar 2024 0 repositories listed
-
Urban Scene Diffusion through Semantic Occupancy Map18 Mar 2024 0 repositories listed
-
N2F2: Hierarchical Scene Understanding with Nested Neural Feature Fields16 Mar 2024 0 repositories listed
-
Segment Any Object Model (SAOM): Real-to-Simulation Fine-Tuning Strategy for Multi-Class Multi-Instance Segmentation16 Mar 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.