Browse State-of-the-Art › Referring Expression Comprehension › Papers, page 2
Referring Expression Comprehension
Papers archive 2025-07-28
archive papers tagged: 167 · with a code link: 98 · where Syntology ran a sample: 50 (47 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (50 of 167 tagged: 47 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 167 of 167, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
24 May 2025 0 repositories listed Syntology 14 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding25 Mar 2025 0 repositories listed
-
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing16 Mar 2025 0 repositories listed
-
Exploring Spatial Language Grounding Through Referring Expressions4 Feb 2025 0 repositories listed
-
FLORA: Formal Language Model Enables Robust Training-free Zero-shot Object Referring Analysis17 Jan 2025 0 repositories listed
-
Omni-RGPT: Unifying Image and Video Region-level Understanding via Token Marks14 Jan 2025 0 repositories listed
-
Hierarchical Alignment-enhanced Adaptive Grounding Network for Generalized Referring Expression Comprehension2 Jan 2025 0 repositories listed
-
DViN: Dynamic Visual Routing Network for Weakly Supervised Referring Expression Comprehension1 Jan 2025 0 repositories listed
-
Task-aware Cross-modal Feature Refinement Transformer with Large Language Models for Visual Grounding1 Jan 2025 0 repositories listed
-
Harlequin: Color-driven Generation of Synthetic Data for Referring Expression Comprehension22 Nov 2024 0 repositories listed
-
Make Graph-based Referring Expression Comprehension Great Again through Expression-guided Dynamic Gating and Regression5 Sep 2024 0 repositories listed
-
Revisiting Multi-Modal LLM Evaluation9 Aug 2024 0 repositories listed
-
MaskInversion: Localized Embeddings via Optimization of Explainability Maps29 Jul 2024 0 repositories listed
-
Learning Visual Grounding from Generative Vision and Language Model18 Jul 2024 0 repositories listed
-
The Solution for the 5th GCAIAC Zero-shot Referring Expression Comprehension Challenge6 Jul 2024 0 repositories listed
-
M²IST: Multi-Modal Interactive Side-Tuning for Efficient Referring Expression Comprehension1 Jul 2024 0 repositories listed
-
Segment Anything Model for automated image data annotation: empirical studies using text prompts from Grounding DINO27 Jun 2024 0 repositories listed
-
ScanFormer: Referring Expression Comprehension by Iteratively Scanning26 Jun 2024 0 repositories listed
-
Text-driven Affordance Learning from Egocentric Vision3 Apr 2024 0 repositories listed
-
PropTest: Automatic Property Testing for Improved Visual Programming25 Mar 2024 0 repositories listed
-
WaterVG: Waterway Visual Grounding based on Text-Guided Vision and mmWave Radar19 Mar 2024 0 repositories listed
-
Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training4 Mar 2024 0 repositories listed
-
Compositional Zero-Shot Learning for Attribute-Based Object Reference in Human-Robot Interaction21 Dec 2023 0 repositories listed
-
8 Dec 2023 0 repositories listed
-
Learning Pseudo-Labeler beyond Noun Concepts for Open-Vocabulary Object Detection4 Dec 2023 0 repositories listed
-
Enhancing Visual Grounding and Generalization: A Multi-Task Cycle Training Approach for Vision-Language Models21 Nov 2023 0 repositories listed
-
CoVLM: Composing Visual Entities and Relationships in Large Language Models Via Communicative Decoding6 Nov 2023 0 repositories listed
-
Video Referring Expression Comprehension via Transformer with Content-conditioned Query25 Oct 2023 0 repositories listed
-
Switching Head-Tail Funnel UNITER for Dual Referring Expression Comprehension with Fetch-and-Carry Tasks14 Jul 2023 0 repositories listed
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input25 Jun 2023 0 repositories listed
-
Language-Guided 3D Object Detection in Point Cloud for Autonomous Driving25 May 2023 0 repositories listed
-
Dynamic Inference With Grounding Based Vision and Language Models1 Jan 2023 0 repositories listed
-
RefCLIP: A Universal Teacher for Weakly Supervised Referring Expression Comprehension1 Jan 2023 0 repositories listed
-
RefTeacher: A Strong Baseline for Semi-Supervised Referring Expression Comprehension1 Jan 2023 0 repositories listed
-
Video Referring Expression Comprehension via Transformer with Content-aware Query6 Oct 2022 0 repositories listed
-
One for All: One-stage Referring Expression Comprehension with Dynamic Reasoning31 Jul 2022 0 repositories listed
-
17 Jun 2022 0 repositories listed
-
RefCrowd: Grounding the Target in Crowd with Referring Expressions16 Jun 2022 0 repositories listed
-
Self-paced Multi-grained Cross-modal Interaction Modeling for Referring Expression Comprehension21 Apr 2022 0 repositories listed
-
FindIt: Generalized Localization with Natural Language Queries31 Mar 2022 0 repositories listed
-
Differentiated Relevances Embedding for Group-based Referring Expression Comprehension12 Mar 2022 0 repositories listed
-
4 Feb 2022 0 repositories listed
-
Lite-MDETR: A Lightweight Multi-Modal Detector1 Jan 2022 0 repositories listed
-
ReCLIP: A Strong Zero-Shot Baseline for Referring Expression Comprehension16 Nov 2021 0 repositories listed
-
Evaluating and Improving Interactions with Hazy Oracles19 Oct 2021 0 repositories listed
-
Proposal-free One-stage Referring Expression via Grid-Word Cross-Attention5 May 2021 0 repositories listed
-
Playing Lottery Tickets with Vision and Language23 Apr 2021 0 repositories listed
-
Co-Grounding Networks with Semantic Attention for Referring Expression Comprehension in Videos23 Mar 2021 0 repositories listed
-
Language-Mediated, Object-Centric Representation Learning31 Dec 2020 0 repositories listed
-
PPGN: Phrase-Guided Proposal Generation Network For Referring Expression Comprehension20 Dec 2020 0 repositories listed
-
Modular Graph Attention Network for Complex Visual Relational Reasoning22 Nov 2020 0 repositories listed
-
15 Nov 2020 0 repositories listed
-
Commands 4 Autonomous Vehicles (C4AV) Workshop Summary18 Sep 2020 0 repositories listed
-
Referring Expression Comprehension: A Survey of Methods and Datasets19 Jul 2020 0 repositories listed
-
30 Jun 2020 0 repositories listed
-
Leveraging Non-Specialists for Accurate and Time Efficient AMR Annotation1 May 2020 0 repositories listed
-
Giving Commands to a Self-driving Car: A Multimodal Reasoner for Visual Grounding19 Mar 2020 0 repositories listed
-
MUTATT: Visual-Textual Mutual Guidance for Referring Expression Comprehension18 Mar 2020 0 repositories listed
-
1 Mar 2020 0 repositories listed
-
UNITER: Learning UNiversal Image-TExt Representations25 Sep 2019 0 repositories listed
-
Dynamic Graph Attention for Referring Expression Comprehension18 Sep 2019 0 repositories listed
-
A Real-Time Cross-modality Correlation Filtering Method for Referring Expression Comprehension16 Sep 2019 0 repositories listed
-
4 Apr 2019 0 repositories listed
-
Neighbourhood Watch: Referring Expression Comprehension via Language-guided Graph Attention Networks12 Dec 2018 0 repositories listed
-
Real-Time Referring Expression Comprehension by Single-Stage Grounding Network9 Dec 2018 0 repositories listed
-
Parallel Attention: A Unified Framework for Visual Object Discovery through Dialogs and Queries17 Nov 2017 0 repositories listed
-
Deep Fragment Embeddings for Bidirectional Image Sentence Mapping22 Jun 2014 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.