Browse State-of-the-Art › Scene Understanding › Papers, page 10
Scene Understanding
Papers archive 2025-07-28
archive papers tagged: 1,723 · with a code link: 720 · where Syntology ran a sample: 208 (182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (208 of 1,723 tagged: 182 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 10 of 18: papers 901 to 1,000 of 1,723, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning8 Jan 2025 0 repositories listed
-
Advancing the Understanding of Fine-Grained 3D Forest Structures using Digital Cousins and Simulation-to-Reality: Methods and Datasets7 Jan 2025 0 repositories listed
-
CL3DOR: Contrastive Learning for 3D Large Multimodal Models via Odds Ratio on High-Resolution Point Clouds7 Jan 2025 0 repositories listed
-
LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving7 Jan 2025 0 repositories listed
-
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer2 Jan 2025 0 repositories listed
-
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction2 Jan 2025 0 repositories listed
-
3D-MVP: 3D Multiview Pretraining for Manipulation1 Jan 2025 0 repositories listed
-
Beyond Human Perception: Understanding Multi-Object World from Monocular View1 Jan 2025 0 repositories listed
-
GroundingFace: Fine-grained Face Understanding via Pixel Grounding Multimodal Large Language Model1 Jan 2025 0 repositories listed
-
HUSH: Holistic Panoramic 3D Scene Understanding using Spherical Harmonics1 Jan 2025 0 repositories listed
-
Scene Map-based Prompt Tuning for Navigation Instruction Generation1 Jan 2025 0 repositories listed
-
TADFormer: Task-Adaptive Dynamic TransFormer for Efficient Multi-Task Learning1 Jan 2025 0 repositories listed
-
Vision-Language Embodiment for Monocular Depth Estimation1 Jan 2025 0 repositories listed
-
Embodied VideoAgent: Persistent Memory from Egocentric Videos and Embodied Sensors Enables Dynamic Scene Understanding31 Dec 2024 0 repositories listed
-
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies31 Dec 2024 0 repositories listed
-
4D Gaussian Splatting: Modeling Dynamic Scenes with Native 4D Primitives30 Dec 2024 0 repositories listed
-
Text-to-Image GAN with Pretrained Representations30 Dec 2024 0 repositories listed
-
UniPLV: Towards Label-Efficient Open-World 3D Scene Understanding by Regional Visual Language Supervision24 Dec 2024 0 repositories listed
-
LangSurf: Language-Embedded Surface Gaussians for 3D Scene Understanding23 Dec 2024 0 repositories listed
-
Application of Multimodal Large Language Models in Autonomous Driving21 Dec 2024 0 repositories listed
-
ObjVariantEnsemble: Advancing Point Cloud LLM Evaluation in Challenging Scenes with Subtly Distinguished Objects19 Dec 2024 0 repositories listed
-
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting18 Dec 2024 0 repositories listed
-
Multi-View Pedestrian Occupancy Prediction with a Novel Synthetic Dataset18 Dec 2024 0 repositories listed
-
An Enhanced Classification Method Based on Adaptive Multi-Scale Fusion for Long-tailed Multispectral Point Clouds16 Dec 2024 0 repositories listed
-
SuperGSeg: Open-Vocabulary 3D Segmentation with Structured Super-Gaussians13 Dec 2024 0 repositories listed
-
MAGIC: Mastering Physical Adversarial Generation in Context through Collaborative LLM Agents11 Dec 2024 0 repositories listed
-
SLGaussian: Fast Language Gaussian Splatting in Sparse Views11 Dec 2024 0 repositories listed
-
TGOSPA Metric Parameters Selection and Evaluation for Visual Multi-object Tracking11 Dec 2024 0 repositories listed
-
Event fields: Capturing light fields at high speed, resolution, and dynamic range9 Dec 2024 0 repositories listed
-
Visual Lexicon: Rich Image Features in Language Space9 Dec 2024 0 repositories listed
-
TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances7 Dec 2024 0 repositories listed
-
Assessing the performance of CT image denoisers using Laguerre-Gauss Channelized Hotelling Observer for lesion detection4 Dec 2024 0 repositories listed
-
Designing DNNs for a trade-off between robustness and processing performance in embedded devices4 Dec 2024 0 repositories listed
-
BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding3 Dec 2024 0 repositories listed
-
SparseLGS: Sparse View Language Embedded Gaussian Splatting3 Dec 2024 0 repositories listed
-
A Semantic Communication System for Real-time 3D Reconstruction Tasks2 Dec 2024 0 repositories listed
-
Holistic Understanding of 3D Scenes as Universal Scene Description2 Dec 2024 0 repositories listed
-
Occam's LGS: A Simple Approach for Language Gaussian Splatting2 Dec 2024 0 repositories listed
-
ChatSplat: 3D Conversational Gaussian Splatting1 Dec 2024 0 repositories listed
-
Quantifying the synthetic and real domain gap in aerial scene understanding29 Nov 2024 0 repositories listed
-
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation29 Nov 2024 0 repositories listed
-
InstanceGaussian: Appearance-Semantic Joint Gaussian Representation for 3D Instance-Level Perception28 Nov 2024 0 repositories listed
-
On-chip Hyperspectral Image Segmentation with Fully Convolutional Networks for Scene Understanding in Autonomous Driving28 Nov 2024 0 repositories listed
-
SceneTAP: Scene-Coherent Typographic Adversarial Planner against Vision-Language Models in Real-World Environments28 Nov 2024 0 repositories listed
-
Reconstructing Animals and the Wild27 Nov 2024 0 repositories listed
-
26 Nov 2024 0 repositories listed
-
Open-Vocabulary Octree-Graph for 3D Scene Understanding25 Nov 2024 0 repositories listed
-
RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Robotics25 Nov 2024 0 repositories listed
-
UniGaussian: Driving Scene Reconstruction from Multiple Camera Models via Unified Gaussian Representations22 Nov 2024 0 repositories listed
-
Multimodal 3D Reasoning Segmentation with Complex Scenes21 Nov 2024 0 repositories listed
-
Classification of Geographical Land Structure Using Convolution Neural Network and Transfer Learning19 Nov 2024 0 repositories listed
-
Calibrated and Efficient Sampling-Free Confidence Estimation for LiDAR Scene Semantic Segmentation18 Nov 2024 0 repositories listed
-
MGNiceNet: Unified Monocular Geometric Scene Understanding18 Nov 2024 0 repositories listed
-
Reducing Label Dependency for Underwater Scene Understanding: A Survey of Datasets, Techniques and Applications18 Nov 2024 0 repositories listed
-
The ADUULM-360 Dataset -- A Multi-Modal Dataset for Depth Estimation in Adverse Weather18 Nov 2024 0 repositories listed
-
Memory-Augmented Multimodal LLMs for Surgical VQA via Self-Contained Inquiry17 Nov 2024 0 repositories listed
-
Large Language Models (LLMs) as Traffic Control Systems at Urban Intersections: A New Paradigm16 Nov 2024 0 repositories listed
-
Content-Aware Preserving Image Generation15 Nov 2024 0 repositories listed
-
SE(3) Equivariant Ray Embeddings for Implicit Multi-View Depth Estimation11 Nov 2024 0 repositories listed
-
Graph-Based Multi-Modal Sensor Fusion for Autonomous Driving6 Nov 2024 0 repositories listed
-
Modeling Uncertainty in 3D Gaussian Splatting through Continuous Semantic Splatting4 Nov 2024 0 repositories listed
-
Symbolic Graph Inference for Compound Scene Understanding30 Oct 2024 0 repositories listed
-
UniRiT: Towards Few-Shot Non-Rigid Point Cloud Registration30 Oct 2024 0 repositories listed
-
Towards Robust Algorithms for Surgical Phase Recognition via Digital Twin-based Scene Representation26 Oct 2024 0 repositories listed
-
PerspectiveNet: Multi-View Perception for Dynamic Scene Understanding22 Oct 2024 0 repositories listed
-
Large Language Models for Autonomous Driving (LLM4AD): Concept, Benchmark, Experiments, and Challenges20 Oct 2024 0 repositories listed
-
SAM-Guided Masked Token Prediction for 3D Scene Understanding16 Oct 2024 0 repositories listed
-
3DArticCyclists: Generating Synthetic Articulated 8D Pose-Controllable Cyclist Data for Computer Vision Applications14 Oct 2024 0 repositories listed
-
Enhancing Single Image to 3D Generation using Gaussian Splatting and Hybrid Diffusion Priors12 Oct 2024 0 repositories listed
-
3D Vision-Language Gaussian Splatting10 Oct 2024 0 repositories listed
-
A transition towards virtual representations of visual scenes10 Oct 2024 0 repositories listed
-
Test-Time Intensity Consistency Adaptation for Shadow Detection10 Oct 2024 0 repositories listed
-
Evaluating the Impact of Point Cloud Colorization on Semantic Segmentation Accuracy9 Oct 2024 0 repositories listed
-
Open-RGBT: Open-vocabulary RGB-T Zero-shot Semantic Segmentation in Open-world Environments9 Oct 2024 0 repositories listed
-
Diffusion Models in 3D Vision: A Survey7 Oct 2024 0 repositories listed
-
Resource-Efficient Multiview Perception: Integrating Semantic Masking with Masked Autoencoders7 Oct 2024 0 repositories listed
-
In-Place Panoptic Radiance Field Segmentation with Perceptual Prior for 3D Scene Understanding6 Oct 2024 0 repositories listed
-
Fast Object Detection with a Machine Learning Edge Device5 Oct 2024 0 repositories listed
-
SPARTUN3D: Situated Spatial Understanding of 3D World in Large Language Models4 Oct 2024 0 repositories listed
-
Class-Agnostic Visio-Temporal Scene Sketch Semantic Segmentation30 Sep 2024 0 repositories listed
-
You Only Speak Once to See27 Sep 2024 0 repositories listed
-
26 Sep 2024 0 repositories listed
-
Scene Understanding in Pick-and-Place Tasks: Analyzing Transformations Between Initial and Final Scenes26 Sep 2024 0 repositories listed
-
OW-Rep: Open World Object Detection with Instance Representation Learning24 Sep 2024 0 repositories listed
-
23 Sep 2024 0 repositories listed
-
Multilateral Cascading Network for Semantic Segmentation of Large-Scale Outdoor Point Clouds21 Sep 2024 0 repositories listed
-
MOSE: Monocular Semantic Reconstruction Using NeRF-Lifted Noisy Priors21 Sep 2024 0 repositories listed
-
Relevance-driven Decision Making for Safer and More Efficient Human Robot Collaboration21 Sep 2024 0 repositories listed
-
DAE-Fuse: An Adaptive Discriminative Autoencoder for Multi-Modality Image Fusion16 Sep 2024 0 repositories listed
-
Video Token Sparsification for Efficient Multimodal LLMs in Autonomous Driving16 Sep 2024 0 repositories listed
-
Relevance for Human Robot Collaboration12 Sep 2024 0 repositories listed
-
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data10 Sep 2024 0 repositories listed
-
TanDepth: Leveraging Global DEMs for Metric Monocular Depth Estimation in UAVs8 Sep 2024 0 repositories listed
-
Future Does Matter: Boosting 3D Object Detection with Temporal Motion Estimation in Point Cloud Sequences6 Sep 2024 0 repositories listed
-
Optimizing 3D Gaussian Splatting for Sparse Viewpoint Scene Reconstruction5 Sep 2024 0 repositories listed
-
Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving4 Sep 2024 0 repositories listed
-
GaussianPU: A Hybrid 2D-3D Upsampling Framework for Enhancing Color Point Clouds via 3D Gaussian Splatting3 Sep 2024 0 repositories listed
-
Leaky Wave Antenna-Equipped RF Chipless Tags for Orientation Estimation31 Aug 2024 0 repositories listed
-
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving29 Aug 2024 0 repositories listed
-
Str-L Pose: Integrating Point and Structured Line for Relative Pose Estimation in Dual-Graph28 Aug 2024 0 repositories listed