Browse State-of-the-Art › Scene Understanding › Papers, page 8
Scene Understanding
Papers archive 2025-07-28
archive papers tagged: 1,723 · with a code link: 720 · where Syntology ran a sample: 208 (182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (208 of 1,723 tagged: 182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 8 of 18: papers 701 to 800 of 1,723, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
16 Apr 2018 1 repository listed Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
12 Apr 2018 1 repository listed
-
12 Mar 2018 1 repository listed
-
18 Feb 2018 1 repository listed
-
1 Oct 2017 1 repository listed
-
18 Sep 2017 1 repository listed
-
8 Aug 2017 1 repository listed
-
2 Aug 2017 1 repository listed
-
31 Jul 2017 1 repository listed
-
1 Jul 2017 1 repository listed
-
1 Jul 2017 1 repository listed
-
15 May 2017 1 repository listed
-
29 Mar 2017 1 repository listed
-
9 Mar 2017 1 repository listed
-
14 Feb 2017 1 repository listed
-
23 Jan 2017 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
15 Dec 2016 1 repository listed
-
25 Nov 2016 1 repository listed
-
22 Nov 2015 1 repository listed
-
1 Jun 2011 1 repository listed
-
Advancing Complex Wide-Area Scene Understanding with Hierarchical Coresets Selection17 Jul 2025 0 repositories listed
-
Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models17 Jul 2025 0 repositories listed
-
City-VLM: Towards Multidomain Perception Scene Understanding via Multimodal Incomplete Learning17 Jul 2025 0 repositories listed
-
Seeing the Signs: A Survey of Edge-Deployable OCR Models for Billboard Visibility Analysis15 Jul 2025 0 repositories listed
-
Tactical Decision for Multi-UGV Confrontation with a Vision-Language Model-Based Commander15 Jul 2025 0 repositories listed
-
EmbRACE-3K: Embodied Reasoning and Action in Complex Environments14 Jul 2025 0 repositories listed
-
MUVOD: A Novel Multi-view Video Object Segmentation Dataset and A Benchmark for 3D Segmentation10 Jul 2025 0 repositories listed
-
What Demands Attention in Urban Street Scenes? From Scene Understanding towards Road Safety: A Survey of Vision-driven Datasets and Studies9 Jul 2025 0 repositories listed
-
VoteSplat: Hough Voting Gaussian Splatting for 3D Scene Understanding28 Jun 2025 0 repositories listed
-
CoPa-SG: Dense Scene Graphs with Parametric and Proto-Relations26 Jun 2025 0 repositories listed
-
Case-based Reasoning Augmented Large Language Model Framework for Decision Making in Realistic Safety-Critical Driving Scenarios25 Jun 2025 0 repositories listed
-
DreamAnywhere: Object-Centric Panoramic 3D Scene Generation25 Jun 2025 0 repositories listed
-
IPFormer: Visual 3D Panoptic Scene Completion with Context-Adaptive Instance Proposals25 Jun 2025 0 repositories listed
-
HOIverse: A Synthetic Scene Graph Dataset With Human Object Interactions24 Jun 2025 0 repositories listed
-
Scene-R1: Video-Grounded Large Language Models for 3D Scene Reasoning without 3D Annotations21 Jun 2025 0 repositories listed
-
Image Segmentation with Large Language Models: A Survey with Perspectives for Intelligent Transportation Systems17 Jun 2025 0 repositories listed
-
Leader360V: The Large-scale, Real-world 360 Video Dataset for Multi-task Learning in Diverse Environment17 Jun 2025 0 repositories listed
-
SceneAware: Scene-Constrained Pedestrian Trajectory Prediction with LLM-Guided Walkability17 Jun 2025 0 repositories listed
-
Unified Representation Space for 3D Visual Grounding17 Jun 2025 0 repositories listed
-
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding16 Jun 2025 0 repositories listed
-
SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis12 Jun 2025 0 repositories listed
-
SemanticSplat: Feed-Forward 3D Scene Understanding with Language-Aware Gaussian Fields11 Jun 2025 0 repositories listed
-
PhyBlock: A Progressive Benchmark for Physical Understanding and Planning via 3D Block Assembly10 Jun 2025 0 repositories listed
-
Robust Visual Localization via Semantic-Guided Multi-Scale Transformer10 Jun 2025 0 repositories listed
-
SceneSplat++: A Large Dataset and Comprehensive Benchmark for Language Gaussian Splatting10 Jun 2025 0 repositories listed
-
Design and Evaluation of Deep Learning-Based Dual-Spectrum Image Fusion Methods9 Jun 2025 0 repositories listed
-
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting9 Jun 2025 0 repositories listed
-
SpatialLM: Training Large Language Models for Structured Indoor Modeling9 Jun 2025 0 repositories listed
-
Does Your 3D Encoder Really Work? When Pretrain-SFT from 2D VLMs Meets 3D VLMs5 Jun 2025 0 repositories listed
-
ProJo4D: Progressive Joint Optimization for Sparse-View Inverse Physics Estimation5 Jun 2025 0 repositories listed
-
Attention-based transformer models for image captioning across languages: An in-depth survey and evaluation3 Jun 2025 0 repositories listed
-
Tactile MNIST: Benchmarking Active Tactile Perception3 Jun 2025 0 repositories listed
-
Learning Sparsity for Effective and Efficient Music Performance Question Answering2 Jun 2025 0 repositories listed
-
SAM2-LOVE: Segment Anything Model 2 in Language-aided Audio-Visual Scenes2 Jun 2025 0 repositories listed
-
SeG-SR: Integrating Semantic Knowledge into Remote Sensing Image Super-Resolution via Vision-Language Model29 May 2025 0 repositories listed
-
DORAEMON: Decentralized Ontology-aware Reliable Agent with Enhanced Memory Oriented Navigation28 May 2025 0 repositories listed
-
LiDAR Based Semantic Perception for Forklifts in Outdoor Environments28 May 2025 0 repositories listed
-
A Graph Completion Method that Jointly Predicts Geometry and Topology Enables Effective Molecule Assembly27 May 2025 0 repositories listed
-
Compositional Scene Understanding through Inverse Generative Modeling27 May 2025 0 repositories listed
-
OccLE: Label-Efficient 3D Semantic Occupancy Prediction27 May 2025 0 repositories listed
-
OmniIndoor3D: Comprehensive Indoor 3D Reconstruction27 May 2025 0 repositories listed
-
Right Side Up? Disentangling Orientation Understanding in MLLMs with Fine-grained Multi-axis Perception Tasks27 May 2025 0 repositories listed
-
Underwater Diffusion Attention Network with Contrastive Language-Image Joint Learning for Underwater Image Enhancement26 May 2025 0 repositories listed
-
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection25 May 2025 0 repositories listed
-
FHGS: Feature-Homogenized Gaussian Splatting25 May 2025 0 repositories listed
-
Can MLLMs Guide Me Home? A Benchmark Study on Fine-Grained Visual Reasoning from Transit Maps24 May 2025 0 repositories listed
-
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding24 May 2025 0 repositories listed
-
From Flight to Insight: Semantic 3D Reconstruction for Aerial Inspection via Gaussian Splatting and Language-Guided Segmentation23 May 2025 0 repositories listed
-
Assessing the generalization performance of SAM for ureteroscopy scene understanding22 May 2025 0 repositories listed
-
HAMF: A Hybrid Attention-Mamba Framework for Joint Scene Context Understanding and Future Motion Representation Learning21 May 2025 0 repositories listed
-
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation21 May 2025 0 repositories listed
-
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets21 May 2025 0 repositories listed
-
AdaToken-3D: Dynamic Spatial Gating for Efficient 3D Large Multimodal-Models Reasoning19 May 2025 0 repositories listed
-
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps19 May 2025 0 repositories listed
-
Can Large Multimodal Models Understand Agricultural Scenes? Benchmarking with AgroMind18 May 2025 0 repositories listed
-
LLaVA-4D: Embedding SpatioTemporal Prompt into LMMs for 4D Scene Understanding18 May 2025 0 repositories listed
-
SEPT: Standard-Definition Map Enhanced Scene Perception and Topology Reasoning for Autonomous Driving18 May 2025 0 repositories listed
-
TinyRS-R1: Compact Multimodal Language Model for Remote Sensing17 May 2025 0 repositories listed
-
Seeing Beyond the Scene: Enhancing Vision-Language Models with Interactional Reasoning14 May 2025 0 repositories listed
-
Deep Learning Advances in Vision-Based Traffic Accident Anticipation: A Comprehensive Review of Methods,Datasets,and Future Directions12 May 2025 0 repositories listed
-
Boosting Cross-spectral Unsupervised Domain Adaptation for Thermal Semantic Segmentation11 May 2025 0 repositories listed
-
Technical Report for ICRA 2025 GOOSE 2D Semantic Segmentation Challenge: Leveraging Color Shift Correction, RoPE-Swin Backbone, and Quantile-based Label Denoising Strategy for Robust Outdoor Scene Understanding11 May 2025 0 repositories listed
-
Camera Control at the Edge with Language Models for Scene Understanding9 May 2025 0 repositories listed
-
Camera-Only Bird's Eye View Perception: A Neural Approach to LiDAR-Free Environmental Mapping for Autonomous Vehicles9 May 2025 0 repositories listed
-
Does CLIP perceive art the same way we do?8 May 2025 0 repositories listed
-
PADriver: Towards Personalized Autonomous Driving8 May 2025 0 repositories listed
-
RAFT: Robust Augmentation of FeaTures for Image Segmentation7 May 2025 0 repositories listed
-
Segment Any RGB-Thermal Model with Language-aided Distillation4 May 2025 0 repositories listed
-
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models3 May 2025 0 repositories listed
-
Embracing Diffraction: A Paradigm Shift in Wireless Sensing and Communication2 May 2025 0 repositories listed
-
V3LMA: Visual 3D-enhanced Language Model for Autonomous Driving30 Apr 2025 0 repositories listed
-
Category-Level and Open-Set Object Pose Estimation for Robotics28 Apr 2025 0 repositories listed
-
Masked Point-Entity Contrast for Open-Vocabulary 3D Scene Understanding28 Apr 2025 0 repositories listed
-
TraveLLaMA: Facilitating Multi-modal Large Language Models to Understand Urban Scenes and Provide Travel Assistance23 Apr 2025 0 repositories listed
-
Multimodal Large Language Models for Enhanced Traffic Safety: A Comprehensive Review and Future Trends21 Apr 2025 0 repositories listed
-
Vision-Centric Representation-Efficient Fine-Tuning for Robust Universal Foreground Segmentation20 Apr 2025 0 repositories listed
-
HAECcity: Open-Vocabulary Scene Understanding of City-Scale Point Clouds with Superpoint Graph Clustering18 Apr 2025 0 repositories listed
-
Temporal Propagation of Asymmetric Feature Pyramid for Surgical Scene Segmentation18 Apr 2025 0 repositories listed
-
Explainable Scene Understanding with Qualitative Representations and Graph Neural Networks17 Apr 2025 0 repositories listed
-
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting16 Apr 2025 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.