Browse State-of-the-Art › Visual Reasoning › Papers, page 6
Visual Reasoning
Papers archive 2025-07-28
archive papers tagged: 698 · with a code link: 356 · where Syntology ran a sample: 165 (130 with a run with no instrument failure, 35 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (165 of 698 tagged: 130 with a run with no instrument failure, 35 where every run was a failure of Syntology's instrument)
Page 6 of 7: papers 501 to 600 of 698, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Beyond Visual Appearances: Privacy-sensitive Objects Identification via Hybrid Graph Reasoning18 Jun 2024 0 repositories listed
-
A Unified View of Abstract Visual Reasoning Problems16 Jun 2024 0 repositories listed
-
A-I-RAVEN and I-RAVEN-Mesh: Two New Benchmarks for Abstract Visual Reasoning16 Jun 2024 0 repositories listed
-
Comparison Visual Instruction Tuning13 Jun 2024 0 repositories listed
-
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems11 Jun 2024 0 repositories listed
-
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model1 Jun 2024 0 repositories listed
-
28 May 2024 0 repositories listed
-
Do Vision-Language Transformers Exhibit Visual Commonsense? An Empirical Study of VCR27 May 2024 0 repositories listed
-
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff26 May 2024 0 repositories listed
-
22 May 2024 0 repositories listed
-
Analogist: Out-of-the-box Visual In-Context Learning with Image Diffusion Model16 May 2024 0 repositories listed
-
CLIP-Powered TASS: Target-Aware Single-Stream Network for Audio-Visual Question Answering13 May 2024 0 repositories listed
-
Naturally Supervised 3D Visual Grounding with Language-Regularized Concept Learners30 Apr 2024 0 repositories listed
-
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models26 Apr 2024 0 repositories listed
-
Cantor: Inspiring Multimodal Chain-of-Thought of MLLM24 Apr 2024 0 repositories listed
-
Think-Program-reCtify: 3D Situated Reasoning with Large Language Models23 Apr 2024 0 repositories listed
-
Automated Evaluation of Large Vision-Language Models on Self-driving Corner Cases16 Apr 2024 0 repositories listed
-
Wu's Method can Boost Symbolic AI to Rival Silver Medalists and AlphaGeometry to Outperform Gold Medalists at IMO Geometry9 Apr 2024 0 repositories listed
-
28 Mar 2024 0 repositories listed
-
PropTest: Automatic Property Testing for Improved Visual Programming25 Mar 2024 0 repositories listed
-
Just Say the Name: Online Continual Learning with Category Names Only via Data Generation16 Mar 2024 0 repositories listed
-
Test-time Distribution Learning Adapter for Cross-modal Visual Reasoning10 Mar 2024 0 repositories listed
-
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection5 Mar 2024 0 repositories listed
-
VISREAS: Complex Visual Reasoning with Unanswerable Questions23 Feb 2024 0 repositories listed
-
Visual In-Context Learning for Large Vision-Language Models18 Feb 2024 0 repositories listed
-
Muffin or Chihuahua? Challenging Multimodal Large Language Models with Multipanel VQA29 Jan 2024 0 repositories listed
-
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking8 Jan 2024 0 repositories listed
-
Towards Truly Zero-shot Compositional Visual Reasoning with LLMs as Programmers3 Jan 2024 0 repositories listed
-
Generate Subgoal Images before Act: Unlocking the Chain-of-Thought Reasoning in Diffusion Model for Robot Manipulation with Multimodal Prompts1 Jan 2024 0 repositories listed
-
ChartBench: A Benchmark for Complex Visual Reasoning in Charts26 Dec 2023 0 repositories listed
-
Leveraging VLM-Based Pipelines to Annotate 3D Objects29 Nov 2023 0 repositories listed
-
From Wrong To Right: A Recursive Approach Towards Vision-Language Explanation21 Nov 2023 0 repositories listed
-
17 Nov 2023 0 repositories listed
-
15 Nov 2023 0 repositories listed
-
Adaptive recurrent vision performs zero-shot computation scaling to unseen difficulty levels12 Nov 2023 0 repositories listed
-
Visual Commonsense based Heterogeneous Graph Contrastive Learning11 Nov 2023 0 repositories listed
-
Towards A Unified Neural Architecture for Visual Recognition and Reasoning10 Nov 2023 0 repositories listed
-
OC-NMN: Object-centric Compositional Neural Module Network for Generative Visual Analogical Reasoning28 Oct 2023 0 repositories listed
-
Open Visual Knowledge Extraction via Relation-Oriented Multimodality Model Prompting28 Oct 2023 0 repositories listed
-
Multimodal Representations for Teacher-Guided Compositional Visual Reasoning24 Oct 2023 0 repositories listed
-
Superpixel Semantics Representation and Pre-training for Vision-Language Task20 Oct 2023 0 repositories listed
-
Visual Question Answering in the Medical Domain20 Sep 2023 0 repositories listed
-
A Continual Learning Paradigm for Non-differentiable Visual Programming Frameworks on Visual Reasoning Tasks18 Sep 2023 0 repositories listed
-
On the Potential of CLIP for Compositional Logical Reasoning30 Aug 2023 0 repositories listed
-
EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE23 Aug 2023 0 repositories listed
-
Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories21 Aug 2023 0 repositories listed
-
Towards Grounded Visual Spatial Reasoning in Multi-Modal Vision Language Models18 Aug 2023 0 repositories listed
-
Tree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual Reasoning18 Aug 2023 0 repositories listed
-
Multimodal Analysis Of Google Bard And GPT-Vision: Experiments In Visual Reasoning17 Aug 2023 0 repositories listed
-
Bridging the Gap: Exploring the Capabilities of Bridge-Architectures for Complex Visual Reasoning Tasks31 Jul 2023 0 repositories listed
-
LOIS: Looking Out of Instance Semantics for Visual Question Answering26 Jul 2023 0 repositories listed
-
Grounded Object Centric Learning18 Jul 2023 0 repositories listed
-
Does Visual Pretraining Help End-to-End Reasoning?17 Jul 2023 0 repositories listed
-
Look, Remember and Reason: Grounded reasoning in videos with language models30 Jun 2023 0 repositories listed
-
PhD Thesis: Exploring the role of (self-)attention in cognitive and computer vision architecture26 Jun 2023 0 repositories listed
-
A Domain-Independent Agent Architecture for Adaptive Operation in Evolving Open Worlds9 Jun 2023 0 repositories listed
-
Leveraging Large Language Models for Scalable Vector Graphics-Driven Image Understanding9 Jun 2023 0 repositories listed
-
11 May 2023 0 repositories listed
-
Incorporating Structured Representations into Pretrained Vision & Language Models Using Scene Graphs10 May 2023 0 repositories listed
-
The role of object-centric representations, guided attention, and external memory on generalizing visual relations14 Apr 2023 0 repositories listed
-
Boosting Cross-task Transferability of Adversarial Patches with Visual Relations11 Apr 2023 0 repositories listed
-
CAVL: Learning Contrastive and Adaptive Representations of Vision and Language10 Apr 2023 0 repositories listed
-
Explainable AI And Visual Reasoning: Insights From Radiology6 Apr 2023 0 repositories listed
-
Navigating to Objects Specified by Images3 Apr 2023 0 repositories listed
-
Curriculum Learning for Compositional Visual Reasoning27 Mar 2023 0 repositories listed
-
3D Concept Learning and Reasoning from Multi-View Images20 Mar 2023 0 repositories listed
-
Understanding and Constructing Latent Modality Structures in Multi-modal Representation Learning10 Mar 2023 0 repositories listed
-
Abstract Visual Reasoning Enabled by Language7 Mar 2023 0 repositories listed
-
Visual Analytics of Neuron Vulnerability to Adversarial Attacks on Convolutional Neural Networks6 Mar 2023 0 repositories listed
-
Jointly Visual- and Semantic-Aware Graph Memory Networks for Temporal Sentence Localization in Videos2 Mar 2023 0 repositories listed
-
Explicit3D: Graph Network with Spatial Inference for Single Image 3D Object Detection13 Feb 2023 0 repositories listed
-
Learning to Agree on Vision Attention for Visual Commonsense Reasoning4 Feb 2023 0 repositories listed
-
A Divide-Align-Conquer Strategy for Program Synthesis8 Jan 2023 0 repositories listed
-
Graph Representation for Order-Aware Visual Transformation1 Jan 2023 0 repositories listed
-
Image as a Foreign Language: BEiT Pretraining for Vision and Vision-Language Tasks1 Jan 2023 0 repositories listed
-
Open Set Video HOI detection from Action-Centric Chain-of-Look Prompting1 Jan 2023 0 repositories listed
-
1 Jan 2023 0 repositories listed
-
EuclidNet: Deep Visual Reasoning for Constructible Problems in Geometry27 Dec 2022 0 repositories listed
-
VQA and Visual Reasoning: An Overview of Recent Datasets, Methods and Challenges26 Dec 2022 0 repositories listed
-
Towards Unsupervised Visual Reasoning: Do Off-The-Shelf Features Know How to Reason?20 Dec 2022 0 repositories listed
-
3 Dec 2022 0 repositories listed
-
29 Nov 2022 0 repositories listed
-
Reason from Context with Self-supervised Learning23 Nov 2022 0 repositories listed
-
21 Nov 2022 0 repositories listed
-
A survey on knowledge-enhanced multimodal learning19 Nov 2022 0 repositories listed
-
3 Nov 2022 0 repositories listed
-
MAMO: Masked Multimodal Modeling for Fine-Grained Vision-Language Representation Learning9 Oct 2022 0 repositories listed
-
Zero-shot visual reasoning through probabilistic analogical mapping29 Sep 2022 0 repositories listed
-
Deep Neural Networks for Visual Reasoning24 Sep 2022 0 repositories listed
-
One for All: One-stage Referring Expression Comprehension with Dynamic Reasoning31 Jul 2022 0 repositories listed
-
3D Concept Grounding on Neural Fields13 Jul 2022 0 repositories listed
-
From Shallow to Deep: Compositional Reasoning over Graphs for Visual Question Answering25 Jun 2022 0 repositories listed
-
Interactive Visual Reasoning under Uncertainty18 Jun 2022 0 repositories listed
-
VL-BEiT: Generative Vision-Language Pretraining2 Jun 2022 0 repositories listed
-
Few-shot Subgoal Planning with Language Models28 May 2022 0 repositories listed
-
Guiding Visual Question Answering with Attention Priors25 May 2022 0 repositories listed
-
Continual learning on 3D point clouds with random compressed rehearsal16 May 2022 0 repositories listed
-
Introduction to Soar8 May 2022 0 repositories listed
-
Answer-Me: Multi-Task Open-Vocabulary Visual Question Answering2 May 2022 0 repositories listed
-
Co-VQA : Answering by Interactive Sub Question Sequence2 Apr 2022 0 repositories listed