Browse State-of-the-Art › Visual Question Answering (VQA) › Papers, page 19
Visual Question Answering (VQA)
Papers archive 2025-07-28
archive papers tagged: 2,167 · with a code link: 1,039 · where Syntology ran a sample: 359 (287 with a run with no instrument failure, 72 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (359 of 2,167 tagged: 287 with a run with no instrument failure, 72 where every run was a failure of Syntology's instrument)
Page 19 of 22: papers 1,801 to 1,900 of 2,167, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Visual Grounding Strategies for Text-Only Natural Language Processing25 Mar 2021 0 repositories listed
-
How to Design Sample and Computationally Efficient VQA Models22 Mar 2021 0 repositories listed
-
A Comprehensive Survey of Scene Graphs: Generation and Application17 Mar 2021 0 repositories listed
-
VMAF And Variants: Towards A Unified VQA13 Mar 2021 0 repositories listed
-
Characterizing Misclassifications of Deep NLP Models12 Mar 2021 0 repositories listed
-
RL-CSDia: Representation Learning of Computer Science Diagrams10 Mar 2021 0 repositories listed
-
Learning Reasoning Paths over Semantic Graphs for Video-grounded Dialogues1 Mar 2021 0 repositories listed
-
Learning Compositional Representation for Few-shot Visual Question Answering21 Feb 2021 0 repositories listed
-
An Empirical Study on the Generalization Power of Neural Representations Learned via Visual Guessing Games31 Jan 2021 0 repositories listed
-
Unanswerable Questions about Images and Texts25 Jan 2021 0 repositories listed
-
Visual Question Answering based on Local-Scene-Aware Referring Expression Generation22 Jan 2021 0 repositories listed
-
Understanding in Artificial Intelligence17 Jan 2021 0 repositories listed
-
Latent Variable Models for Visual Question Answering16 Jan 2021 0 repositories listed
-
Reasoning over Vision and Language: Exploring the Benefits of Supplemental Knowledge15 Jan 2021 0 repositories listed
-
Recent Advances in Video Question Answering: A Review of Datasets and Methods15 Jan 2021 0 repositories listed
-
Understanding the Role of Scene Graphs in Visual Question Answering14 Jan 2021 0 repositories listed
-
Predicting Relative Depth between Objects from Semantic Features12 Jan 2021 0 repositories listed
-
Transformers in Vision: A Survey4 Jan 2021 0 repositories listed
-
Differentiable End-to-End Program Executor for Sample and Computationally Efficient VQA1 Jan 2021 0 repositories listed
-
Erasure for Advancing: Dynamic Self-Supervised Learning for Commonsense Reasoning1 Jan 2021 0 repositories listed
-
Hierarchical Graph Attention Network for Few-Shot Visual-Semantic Learning1 Jan 2021 0 repositories listed
-
Linguistically Routing Capsule Network for Out-of-Distribution Visual Question Answering1 Jan 2021 0 repositories listed
-
Unshuffling Data for Improved Generalization in Visual Question Answering1 Jan 2021 0 repositories listed
-
Seeing is Knowing! Fact-based Visual Question Answering using Knowledge Graph Embeddings31 Dec 2020 0 repositories listed
-
Object-Centric Diagnosis of Visual Reasoning21 Dec 2020 0 repositories listed
-
20 Dec 2020 0 repositories listed
-
Trying Bilinear Pooling in Video-QA18 Dec 2020 0 repositories listed
-
13 Dec 2020 0 repositories listed
-
WeaQA: Weak Supervision via Captions for Visual Question Answering4 Dec 2020 0 repositories listed
-
1 Dec 2020 0 repositories listed
-
Multimodal Graph Networks for Compositional Generalization in Visual Question Answering1 Dec 2020 0 repositories listed
-
Modular Graph Attention Network for Complex Visual Relational Reasoning22 Nov 2020 0 repositories listed
-
Logically Consistent Loss for Visual Question Answering19 Nov 2020 0 repositories listed
-
Generating Natural Questions from Images for Multimodal Assistants17 Nov 2020 0 repositories listed
-
Reasoning Over History: Context Aware Visual Dialog2 Nov 2020 0 repositories listed
-
Can Pre-training help VQA with Lexical Variations?1 Nov 2020 0 repositories listed
-
CapWAP: Image Captioning with a Purpose1 Nov 2020 0 repositories listed
-
ISAAQ - Mastering Textbook Questions with Pre-trained Transformers and Bottom-Up and Top-Down Attention1 Nov 2020 0 repositories listed
-
Representation, Learning and Reasoning on Spatial Language for Downstream NLP Tasks1 Nov 2020 0 repositories listed
-
STL-CQA: Structure-based Transformers with Localization and Encoding for Chart Question Answering1 Nov 2020 0 repositories listed
-
Leveraging Visual Question Answering to Improve Text-to-Image Synthesis28 Oct 2020 0 repositories listed
-
Beyond VQA: Generating Multi-word Answer and Rationale to Visual Questions24 Oct 2020 0 repositories listed
-
Answer-checking in Context: A Multi-modal FullyAttention Network for Visual Question Answering17 Oct 2020 0 repositories listed
-
New Ideas and Trends in Deep Multimodal Content Understanding: A Review16 Oct 2020 0 repositories listed
-
Does my multimodal model learn cross-modal interactions? It's harder to tell than you might think!13 Oct 2020 0 repositories listed
-
Interpretable Neural Computation for Real-World Compositional Visual Question Answering10 Oct 2020 0 repositories listed
-
8 Oct 2020 0 repositories listed
-
Finding the Evidence: Localization-aware Answer Prediction for Text Visual Question Answering6 Oct 2020 0 repositories listed
-
Pathological Visual Question Answering6 Oct 2020 0 repositories listed
-
Attention Guided Semantic Relationship Parsing for Visual Question Answering5 Oct 2020 0 repositories listed
-
CAPTION: Correction by Analyses, POS-Tagging and Interpretation of Objects using only Nouns2 Oct 2020 0 repositories listed
-
ISAAQ -- Mastering Textbook Questions with Pre-trained Transformers and Bottom-Up and Top-Down Attention1 Oct 2020 0 repositories listed
-
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network30 Sep 2020 0 repositories listed
-
Spatial Attention as an Interface for Image Captioning Models29 Sep 2020 0 repositories listed
-
Regularizing Attention Networks for Anomaly Detection in Visual Question Answering21 Sep 2020 0 repositories listed
-
A Multimodal Memes Classification: A Survey and Open Research Issues17 Sep 2020 0 repositories listed
-
Cross-modal Knowledge Reasoning for Knowledge-based Visual Question Answering31 Aug 2020 0 repositories listed
-
Visual Question Answering on Image Sets27 Aug 2020 0 repositories listed
-
Document Visual Question Answering Challenge 202020 Aug 2020 0 repositories listed
-
Linguistically-aware Attention for Reducing the Semantic-Gap in Vision-Language Tasks18 Aug 2020 0 repositories listed
-
Graph Edit Distance Reward: Learning to Edit Scene Graph15 Aug 2020 0 repositories listed
-
Assisting Scene Graph Generation with Self-Supervision8 Aug 2020 0 repositories listed
-
Interpretable Visual Reasoning via Probabilistic Formulation under Natural Supervision1 Aug 2020 0 repositories listed
-
TRRNet: Tiered Relation Reasoning for Compositional Visual Question Answering1 Aug 2020 0 repositories listed
-
Contrastive Visual-Linguistic Pretraining26 Jul 2020 0 repositories listed
-
Reducing Language Biases in Visual Question Answering with Visually-Grounded Question Encoder13 Jul 2020 0 repositories listed
-
Image Captioning with Compositional Neural Module Networks10 Jul 2020 0 repositories listed
-
Eliminating Catastrophic Interference with Biased Competition3 Jul 2020 0 repositories listed
-
Visual Question Answering as a Multi-Task Problem3 Jul 2020 0 repositories listed
-
Scene Graph Reasoning for Visual Question Answering2 Jul 2020 0 repositories listed
-
The Impact of Explanations on AI Competency Prediction in VQA2 Jul 2020 0 repositories listed
-
Aligned Dual Channel Graph Convolutional Network for Visual Question Answering1 Jul 2020 0 repositories listed
-
Multimodal Neural Graph Memory Networks for Visual Question Answering1 Jul 2020 0 repositories listed
-
Towards Visual Dialog for Radiology1 Jul 2020 0 repositories listed
-
30 Jun 2020 0 repositories listed
-
Improving VQA and its Explanations by Comparing Competing Explanations28 Jun 2020 0 repositories listed
-
Self-Segregating and Coordinated-Segregating Transformer for Focused Deep Multi-Modular Network for Visual Question Answering25 Jun 2020 0 repositories listed
-
Mucko: Multi-Layer Cross-Modal Knowledge Reasoning for Fact-based Visual Question Answering16 Jun 2020 0 repositories listed
-
ORD: Object Relationship Discovery for Visual Dialogue Generation15 Jun 2020 0 repositories listed
-
Exploring Weaknesses of VQA Models through Attribution Driven Insights11 Jun 2020 0 repositories listed
-
Estimating semantic structure for the VQA answer space10 Jun 2020 0 repositories listed
-
Counterfactual Vision and Language Learning1 Jun 2020 0 repositories listed
-
Multimodal grid features and cell pointers for Scene Text Visual Question Answering1 Jun 2020 0 repositories listed
-
TA-Student VQA: Multi-Agents Training by Self-Questioning1 Jun 2020 0 repositories listed
-
On the Value of Out-of-Distribution Testing: An Example of Goodhart's Law19 May 2020 0 repositories listed
-
Visual Relationship Detection using Scene Graphs: A Survey16 May 2020 0 repositories listed
-
Visual Question Answering with Prior Class Semantics4 May 2020 0 repositories listed
-
A Corpus for Visual Question Answering Annotated with Frame Semantic Information1 May 2020 0 repositories listed
-
Image Position Prediction in Multimodal Documents1 May 2020 0 repositories listed
-
A Novel Attention-based Aggregation Function to Combine Vision and Language27 Apr 2020 0 repositories listed
-
Visual Question Answering Using Semantic Information from Image Descriptions23 Apr 2020 0 repositories listed
-
Learning What Makes a Difference from Counterfactual Examples and Gradient Supervision20 Apr 2020 0 repositories listed
-
Are we pretraining it right? Digging deeper into visio-linguistic pretraining19 Apr 2020 0 repositories listed
-
Knowledge-Based Visual Question Answering in Videos17 Apr 2020 0 repositories listed
-
Rephrasing visual questions by specifying the entropy of the answer distribution10 Apr 2020 0 repositories listed
-
Understanding Knowledge Gaps in Visual Question Answering: Implications for Gap Identification and Testing8 Apr 2020 0 repositories listed
-
Generating Rationales in Visual Question Answering4 Apr 2020 0 repositories listed
-
27 Mar 2020 0 repositories listed
-
Linguistically Driven Graph Capsule Network for Visual Question Reasoning23 Mar 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.