Browse State-of-the-Art › Visual Grounding › Papers, page 6
Visual Grounding
Papers archive 2025-07-28
archive papers tagged: 571 · with a code link: 299 · where Syntology ran a sample: 111 (95 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (111 of 571 tagged: 95 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 6 of 6: papers 501 to 571 of 571, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Suspected Object Matters: Rethinking Model's Prediction for One-stage Visual Grounding10 Mar 2022 0 repositories listed
-
1 Jan 2022 0 repositories listed
-
Deconfounded Visual Grounding31 Dec 2021 0 repositories listed
-
RoViST: Learning Robust Metrics for Visual Storytelling17 Dec 2021 0 repositories listed
-
2 Dec 2021 0 repositories listed
-
Less is More: Generating Grounded Navigation Instructions from Landmarks25 Nov 2021 0 repositories listed
-
Zero-Shot Visual Grounding of Referring Utterances in Dialogue16 Nov 2021 0 repositories listed
-
Attention as Grounding: Exploring Textual and Cross-Modal Attention on Entities and Relations in Language-and-Vision Transformer16 Oct 2021 0 repositories listed
-
Efficient Multi-Modal Embeddings from Structured Data6 Oct 2021 0 repositories listed
-
Retrieve, Caption, Generate: Visual Grounding for Enhancing Commonsense in Text Generation Models8 Sep 2021 0 repositories listed
-
INVIGORATE: Interactive Visual Grounding and Grasping in Clutter25 Aug 2021 0 repositories listed
-
TransRefer3D: Entity-and-Relation Aware Transformer for Fine-Grained 3D Visual Grounding5 Aug 2021 0 repositories listed
-
Attending Self-Attention: A Case Study of Visually Grounded Supervision in Vision-and-Language Transformers1 Aug 2021 0 repositories listed
-
Word2Pix: Word to Pixel Cross Attention Transformer in Visual Grounding31 Jul 2021 0 repositories listed
-
LanguageRefer: Spatial-Language Model for 3D Visual Grounding7 Jul 2021 0 repositories listed
-
Adventurer's Treasure Hunt: A Transparent System for Visually Grounded Compositional Visual Question Answering based on Scene Graphs28 Jun 2021 0 repositories listed
-
AIFit: Automatic 3D Human-Interpretable Feedback Models for Fitness Training19 Jun 2021 0 repositories listed
-
Attention-Based Keyword Localisation in Speech using Visual Grounding16 Jun 2021 0 repositories listed
-
Visual Grounding Strategies for Text-Only Natural Language Processing25 Mar 2021 0 repositories listed
-
Scene-Intuitive Agent for Remote Embodied Visual Grounding24 Mar 2021 0 repositories listed
-
Decoupled Spatial Temporal Graphs for Generic Visual Grounding18 Mar 2021 0 repositories listed
-
Few-Shot Visual Grounding for Natural Human-Robot Interaction17 Mar 2021 0 repositories listed
-
Transformers in Vision: A Survey4 Jan 2021 0 repositories listed
-
3DVG-Transformer: Relation Modeling for Visual Grounding on Point Clouds1 Jan 2021 0 repositories listed
-
Explainable Video Entailment With Grounded Visual Evidence1 Jan 2021 0 repositories listed
-
CASTing Your Model: Learning to Localize Improves Self-Supervised Representations8 Dec 2020 0 repositories listed
-
28 Nov 2020 0 repositories listed
-
Commands 4 Autonomous Vehicles (C4AV) Workshop Summary18 Sep 2020 0 repositories listed
-
Propagating Over Phrase Relations for One-Stage Visual Grounding1 Aug 2020 0 repositories listed
-
Reducing Language Biases in Visual Question Answering with Visually-Grounded Question Encoder13 Jul 2020 0 repositories listed
-
Multi-Granularity Modularized Network for Abstract Visual Reasoning9 Jul 2020 0 repositories listed
-
Knowledge Supports Visual Language Grounding: A Case Study on Colour Terms1 Jul 2020 0 repositories listed
-
Fast visual grounding in interaction: bringing few-shot learning with neural networks to an interactive robot1 Jun 2020 0 repositories listed
-
Visual Grounding Annotation of Recipe Flow Graph1 May 2020 0 repositories listed
-
Spatio-Temporal Graph for Video Captioning with Knowledge Distillation31 Mar 2020 0 repositories listed
-
Giving Commands to a Self-driving Car: A Multimodal Reasoner for Visual Grounding19 Mar 2020 0 repositories listed
-
Emergent Communication with World Models22 Feb 2020 0 repositories listed
-
Exploring Context, Attention and Audio Features for Audio Visual Scene-Aware Dialog20 Dec 2019 0 repositories listed
-
Compositional Temporal Visual Grounding of Natural Language Event Descriptions4 Dec 2019 0 repositories listed
-
OptiBox: Breaking the Limits of Proposals for Visual Grounding29 Nov 2019 0 repositories listed
-
Leveraging Past References for Robust Language Grounding1 Nov 2019 0 repositories listed
-
Countering Language Drift via Visual Grounding10 Sep 2019 0 repositories listed
-
Differentiable Disentanglement Filter: an Application Agnostic Core Concept Discovery Probe4 Sep 2019 0 repositories listed
-
Multimodal Unified Attention Networks for Vision-and-Language Interactions12 Aug 2019 0 repositories listed
-
Differentiable Disentanglement Filter: an Application Agnostic Core Concept Discovery Probe17 Jul 2019 0 repositories listed
-
Transfer Learning from Audio-Visual Grounding to Speech Recognition9 Jul 2019 0 repositories listed
-
Joint Visual Grounding with Language Scene Graphs9 Jun 2019 0 repositories listed
-
Visually Grounded Neural Syntax Acquisition7 Jun 2019 0 repositories listed
-
Learning to Compose and Reason with Language Tree Structures for Visual Grounding5 Jun 2019 0 repositories listed
-
On the Contributions of Visual and Textual Supervision in Low-Resource Semantic Speech Retrieval24 Apr 2019 0 repositories listed
-
4 Apr 2019 0 repositories listed
-
Revisiting Visual Grounding3 Apr 2019 0 repositories listed
-
Align2Ground: Weakly Supervised Phrase Grounding Guided by Image-Caption Alignment27 Mar 2019 0 repositories listed
-
You Only Look & Listen Once: Towards Fast and Accurate Visual Grounding12 Feb 2019 0 repositories listed
-
Taking a HINT: Leveraging Explanations to Make Vision and Language Models More Grounded11 Feb 2019 0 repositories listed
-
Learning to Assemble Neural Module Tree Networks for Visual Grounding8 Dec 2018 0 repositories listed
-
Multi-task Learning of Hierarchical Vision-Language Representation3 Dec 2018 0 repositories listed
-
Being data-driven is not enough: Revisiting interactive instruction giving as a challenge for NLG1 Nov 2018 0 repositories listed
-
Overcoming Language Priors in Visual Question Answering with Adversarial Regularization8 Oct 2018 0 repositories listed
-
Interpretable Visual Question Answering by Visual Grounding from Attention Supervision Mining1 Aug 2018 0 repositories listed
-
Illustrative Language Understanding: Large-Scale Visual Grounding with Image Search1 Jul 2018 0 repositories listed
-
Visually grounded cross-lingual keyword spotting in speech13 Jun 2018 0 repositories listed
-
Interactive Visual Grounding of Referring Expressions for Human-Robot Interaction11 Jun 2018 0 repositories listed
-
Finding "It": Weakly-Supervised Reference-Aware Visual Grounding in Instructional Videos1 Jun 2018 0 repositories listed
-
Visual Grounding via Accumulated Attention1 Jun 2018 0 repositories listed
-
Learning Unsupervised Visual Grounding Through Semantic Self-Supervision17 Mar 2018 0 repositories listed
-
Improving Visually Grounded Sentence Representations with Self-Attention2 Dec 2017 0 repositories listed
-
Interactive Reinforcement Learning for Object Grounding via Self-Talking2 Dec 2017 0 repositories listed
-
23 Sep 2017 0 repositories listed
-
Weakly-supervised Visual Grounding of Phrases with Linguistic Structures3 May 2017 0 repositories listed
-
Image-Grounded Conversations: Multimodal Context for Natural Question and Response Generation28 Jan 2017 0 repositories listed