Browse State-of-the-Art › Scene Understanding › Papers, page 12
Scene Understanding
Papers archive 2025-07-28
archive papers tagged: 1,723 · with a code link: 720 · where Syntology ran a sample: 208 (182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (208 of 1,723 tagged: 182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 12 of 18: papers 1,101 to 1,200 of 1,723, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning15 Mar 2024 0 repositories listed
-
Mapping High-level Semantic Regions in Indoor Environments without Object Recognition11 Mar 2024 0 repositories listed
-
Out of the Room: Generalizing Event-Based Dynamic Motion Segmentation for Complex Scenes7 Mar 2024 0 repositories listed
-
GSNeRF: Generalizable Semantic Neural Radiance Fields with Enhanced 3D Scene Understanding6 Mar 2024 0 repositories listed
-
HUNTER: Unsupervised Human-centric 3D Detection via Transferring Knowledge from Synthetic Instances to Real Scenes5 Mar 2024 0 repositories listed
-
PCDepth: Pattern-based Complementary Learning for Monocular Depth Estimation by Best of Both Worlds29 Feb 2024 0 repositories listed
-
LiveHPS: LiDAR-based Scene-level Human Pose and Shape Estimation in Free Environment27 Feb 2024 0 repositories listed
-
OpenSUN3D: 1st Workshop Challenge on Open-Vocabulary 3D Scene Understanding23 Feb 2024 0 repositories listed
-
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models19 Feb 2024 0 repositories listed
-
Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation14 Feb 2024 0 repositories listed
-
InCoRo: In-Context Learning for Robotics Control with Feedback Loops7 Feb 2024 0 repositories listed
-
Neural Language of Thought Models2 Feb 2024 0 repositories listed
-
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents23 Jan 2024 0 repositories listed
-
Digital Divides in Scene Recognition: Uncovering Socioeconomic Biases in Deep Learning Systems23 Jan 2024 0 repositories listed
-
S³M-Net: Joint Learning of Semantic Segmentation and Stereo Matching for Autonomous Driving21 Jan 2024 0 repositories listed
-
BPDO:Boundary Points Dynamic Optimization for Arbitrary Shape Scene Text Detection18 Jan 2024 0 repositories listed
-
SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding17 Jan 2024 0 repositories listed
-
Class-Imbalanced Semi-Supervised Learning for Large-Scale Point Cloud Semantic Segmentation via Decoupling Optimization13 Jan 2024 0 repositories listed
-
Learning Segmented 3D Gaussians via Efficient Feature Unprojection for Zero-shot Neural Scene Segmentation11 Jan 2024 0 repositories listed
-
Exploring Self- and Cross-Triplet Correlations for Human-Object Interaction Detection11 Jan 2024 0 repositories listed
-
VLP: Vision Language Planning for Autonomous Driving10 Jan 2024 0 repositories listed
-
FMGS: Foundation Model Embedded 3D Gaussian Splatting for Holistic 3D Scene Understanding3 Jan 2024 0 repositories listed
-
Bilateral Adaptation for Human-Object Interaction Detection with Occlusion-Robustness1 Jan 2024 0 repositories listed
-
Going Beyond Multi-Task Dense Prediction with Synergy Embedding Models1 Jan 2024 0 repositories listed
-
Omni-Q: Omni-Directional Scene Understanding for Unsupervised Visual Grounding1 Jan 2024 0 repositories listed
-
PanoRecon: Real-Time Panoptic 3D Reconstruction from Monocular Video1 Jan 2024 0 repositories listed
-
SceneFun3D: Fine-Grained Functionality and Affordance Understanding in 3D Scenes1 Jan 2024 0 repositories listed
-
Towards CLIP-driven Language-free 3D Visual Grounding via 2D-3D Relational Enhancement and Consistency1 Jan 2024 0 repositories listed
-
Unsupervised 3D Structure Inference from Category-Specific Image Collections1 Jan 2024 0 repositories listed
-
When Visual Grounding Meets Gigapixel-level Large-scale Scenes: Benchmark and Approach1 Jan 2024 0 repositories listed
-
Robust Multi-Modal Image Stitching for Improved Scene Understanding28 Dec 2023 0 repositories listed
-
Cloud-Device Collaborative Learning for Multimodal Large Language Models26 Dec 2023 0 repositories listed
-
BridgeNet: Comprehensive and Effective Feature Interactions via Bridge Feature for Multi-task Dense Predictions21 Dec 2023 0 repositories listed
-
AccidentGPT: Accident Analysis and Prevention from V2X Environmental Perception with Multi-modal Large Model20 Dec 2023 0 repositories listed
-
Language-Assisted 3D Scene Understanding18 Dec 2023 0 repositories listed
-
Weakly-Supervised 3D Visual Grounding based on Visual Linguistic Alignment15 Dec 2023 0 repositories listed
-
Dietary Assessment with Multimodal ChatGPT: A Systematic Analysis14 Dec 2023 0 repositories listed
-
VMT-Adapter: Parameter-Efficient Transfer Learning for Multi-Task Dense Scene Understanding14 Dec 2023 0 repositories listed
-
Cataract-1K: Cataract Surgery Dataset for Scene Segmentation, Phase Recognition, and Irregularity Detection11 Dec 2023 0 repositories listed
-
SkyScenes: A Synthetic Dataset for Aerial Scene Understanding11 Dec 2023 0 repositories listed
-
Spatiotemporal Event Graphs for Dynamic Scene Understanding11 Dec 2023 0 repositories listed
-
Prospective Role of Foundation Models in Advancing Autonomous Vehicles8 Dec 2023 0 repositories listed
-
A Review and A Robust Framework of Data-Efficient 3D Scene Parsing with Traditional/Learned 3D Descriptors3 Dec 2023 0 repositories listed
-
Segment Any 3D Gaussians1 Dec 2023 0 repositories listed
-
HAtt-Flow: Hierarchical Attention-Flow Mechanism for Group Activity Scene Graph Generation in Videos28 Nov 2023 0 repositories listed
-
Scene Summarization: Clustering Scene Videos into Spatially Diverse Frames28 Nov 2023 0 repositories listed
-
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding27 Nov 2023 0 repositories listed
-
REACT: Recognize Every Action Everywhere All At Once27 Nov 2023 0 repositories listed
-
GPT-4V Takes the Wheel: Promises and Challenges for Pedestrian Behavior Prediction24 Nov 2023 0 repositories listed
-
GP-NeRF: Generalized Perception NeRF for Context-Aware 3D Scene Understanding20 Nov 2023 0 repositories listed
-
SeaDSC: A video-based unsupervised method for dynamic scene change detection in unmanned surface vehicles20 Nov 2023 0 repositories listed
-
Two Stream Scene Understanding on Graph Embedding12 Nov 2023 0 repositories listed
-
Leveraging Large-Scale Pretrained Vision Foundation Models for Label-Efficient 3D Point Cloud Segmentation3 Nov 2023 0 repositories listed
-
Single-view 3D Scene Reconstruction with High-fidelity Shape and Texture1 Nov 2023 0 repositories listed
-
Recent Advances in Multi-modal 3D Scene Understanding: A Comprehensive Survey and Evaluation24 Oct 2023 0 repositories listed
-
Panoptic Out-of-Distribution Segmentation18 Oct 2023 0 repositories listed
-
S4C: Self-Supervised Semantic Scene Completion with Neural Fields11 Oct 2023 0 repositories listed
-
TextPSG: Panoptic Scene Graph Generation from Textual Descriptions10 Oct 2023 0 repositories listed
-
Zero-Shot Open-Vocabulary Tracking with Large Pre-Trained Models10 Oct 2023 0 repositories listed
-
Elastic Interaction Energy-Informed Real-Time Traffic Scene Perception2 Oct 2023 0 repositories listed
-
Logical Bias Learning for Object Relation Prediction1 Oct 2023 0 repositories listed
-
SGRec3D: Self-Supervised 3D Scene Graph Learning via Object-Level Scene Reconstruction27 Sep 2023 0 repositories listed
-
Language-EXtended Indoor SLAM (LEXIS): A Versatile System for Real-time Visual Scene Understanding26 Sep 2023 0 repositories listed
-
LLMR: Real-time Prompting of Interactive Worlds using Large Language Models21 Sep 2023 0 repositories listed
-
SANPO: A Scene Understanding, Accessibility and Human Navigation Dataset21 Sep 2023 0 repositories listed
-
Survey of Action Recognition, Spotting and Spatio-Temporal Localization in Soccer -- Current Trends and Research Perspectives21 Sep 2023 0 repositories listed
-
PanoMixSwap Panorama Mixing via Structural Swapping for Indoor Scene Understanding18 Sep 2023 0 repositories listed
-
So you think you can track?13 Sep 2023 0 repositories listed
-
AmodalSynthDrive: A Synthetic Amodal Perception Dataset for Autonomous Driving12 Sep 2023 0 repositories listed
-
Rank2Tell: A Multimodal Driving Dataset for Joint Importance Ranking and Reasoning12 Sep 2023 0 repositories listed
-
Can you text what is happening? Integrating pre-trained language encoders into trajectory prediction models for autonomous driving11 Sep 2023 0 repositories listed
-
PAg-NeRF: Towards fast and efficient end-to-end panoptic 3D representations for agricultural robotics11 Sep 2023 0 repositories listed
-
Weakly Supervised Point Clouds Transformer for 3D Object Detection8 Sep 2023 0 repositories listed
-
Structural Concept Learning via Graph Attention for Multi-Level Rearrangement Planning5 Sep 2023 0 repositories listed
-
Expanding Frozen Vision-Language Models without Retraining: Towards Improved Robot Perception31 Aug 2023 0 repositories listed
-
Semi-Supervised Semantic Depth Estimation using Symbiotic Transformer and NearFarMix Augmentation28 Aug 2023 0 repositories listed
-
End-to-end Autonomous Driving using Deep Learning: A Systematic Review27 Aug 2023 0 repositories listed
-
Synergizing Contrastive Learning and Optimal Transport for 3D Point Cloud Domain Adaptation27 Aug 2023 0 repositories listed
-
SurGNN: Explainable visual scene understanding and assessment of surgical skill using graph neural networks24 Aug 2023 0 repositories listed
-
Novel-view Synthesis and Pose Estimation for Hand-Object Interaction from Sparse Views22 Aug 2023 0 repositories listed
-
Explore and Tell: Embodied Visual Captioning in 3D Environments21 Aug 2023 0 repositories listed
-
15 Aug 2023 0 repositories listed
-
Temporal DINO: A Self-supervised Video Strategy to Enhance Action Prediction8 Aug 2023 0 repositories listed
-
Syn-Mediverse: A Multimodal Synthetic Dataset for Intelligent Scene Understanding of Healthcare Facilities6 Aug 2023 0 repositories listed
-
Scene-aware Human Pose Generation using Transformer4 Aug 2023 0 repositories listed
-
Weakly Supervised 3D Instance Segmentation without Instance-level Annotations3 Aug 2023 0 repositories listed
-
Interpretable End-to-End Driving Model for Implicit Scene Understanding2 Aug 2023 0 repositories listed
-
1 Aug 2023 0 repositories listed
-
Enhancing image captioning with depth information using a Transformer-based framework24 Jul 2023 0 repositories listed
-
Challenges for Monocular 6D Object Pose Estimation in Robotics22 Jul 2023 0 repositories listed
-
Improving Online Lane Graph Extraction by Object-Lane Clustering20 Jul 2023 0 repositories listed
-
Mining Conditional Part Semantics with Occluded Extrapolation for Human-Object Interaction Detection19 Jul 2023 0 repositories listed
-
Human Action Recognition in Still Images Using ConViT18 Jul 2023 0 repositories listed
-
Towards A Unified Agent with Foundation Models18 Jul 2023 0 repositories listed
-
Smart Infrastructure: A Research Junction12 Jul 2023 0 repositories listed
-
Test-Time Adaptation for Nighttime Color-Thermal Semantic Segmentation10 Jul 2023 0 repositories listed
-
PSDR-Room: Single Photo to Scene using Differentiable Rendering6 Jul 2023 0 repositories listed
-
Object Recognition System on a Tactile Device for Visually Impaired5 Jul 2023 0 repositories listed
-
Artifacts Mapping: Multi-Modal Semantic Mapping for Object Detection and 3D Localization3 Jul 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.