Browse State-of-the-Art › Visual Question Answering › Papers, page 20
Visual Question Answering
Papers archive 2025-07-28
archive papers tagged: 2,177 · with a code link: 1,042 · where Syntology ran a sample: 378 (308 with a run with no instrument failure, 70 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (378 of 2,177 tagged: 308 with a run with no instrument failure, 70 where every run was a failure of Syntology's instrument)
Page 20 of 22: papers 1,901 to 2,000 of 2,177, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
CapWAP: Image Captioning with a Purpose1 Nov 2020 0 repositories listed
-
ISAAQ - Mastering Textbook Questions with Pre-trained Transformers and Bottom-Up and Top-Down Attention1 Nov 2020 0 repositories listed
-
Representation, Learning and Reasoning on Spatial Language for Downstream NLP Tasks1 Nov 2020 0 repositories listed
-
Leveraging Visual Question Answering to Improve Text-to-Image Synthesis28 Oct 2020 0 repositories listed
-
Beyond VQA: Generating Multi-word Answer and Rationale to Visual Questions24 Oct 2020 0 repositories listed
-
Answer-checking in Context: A Multi-modal FullyAttention Network for Visual Question Answering17 Oct 2020 0 repositories listed
-
New Ideas and Trends in Deep Multimodal Content Understanding: A Review16 Oct 2020 0 repositories listed
-
Does my multimodal model learn cross-modal interactions? It's harder to tell than you might think!13 Oct 2020 0 repositories listed
-
Interpretable Neural Computation for Real-World Compositional Visual Question Answering10 Oct 2020 0 repositories listed
-
8 Oct 2020 0 repositories listed
-
Finding the Evidence: Localization-aware Answer Prediction for Text Visual Question Answering6 Oct 2020 0 repositories listed
-
Pathological Visual Question Answering6 Oct 2020 0 repositories listed
-
Attention Guided Semantic Relationship Parsing for Visual Question Answering5 Oct 2020 0 repositories listed
-
CAPTION: Correction by Analyses, POS-Tagging and Interpretation of Objects using only Nouns2 Oct 2020 0 repositories listed
-
ISAAQ -- Mastering Textbook Questions with Pre-trained Transformers and Bottom-Up and Top-Down Attention1 Oct 2020 0 repositories listed
-
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network30 Sep 2020 0 repositories listed
-
Spatial Attention as an Interface for Image Captioning Models29 Sep 2020 0 repositories listed
-
Regularizing Attention Networks for Anomaly Detection in Visual Question Answering21 Sep 2020 0 repositories listed
-
A Multimodal Memes Classification: A Survey and Open Research Issues17 Sep 2020 0 repositories listed
-
Cross-modal Knowledge Reasoning for Knowledge-based Visual Question Answering31 Aug 2020 0 repositories listed
-
Visual Question Answering on Image Sets27 Aug 2020 0 repositories listed
-
Document Visual Question Answering Challenge 202020 Aug 2020 0 repositories listed
-
Assisting Scene Graph Generation with Self-Supervision8 Aug 2020 0 repositories listed
-
Interpretable Visual Reasoning via Probabilistic Formulation under Natural Supervision1 Aug 2020 0 repositories listed
-
TRRNet: Tiered Relation Reasoning for Compositional Visual Question Answering1 Aug 2020 0 repositories listed
-
Reducing Language Biases in Visual Question Answering with Visually-Grounded Question Encoder13 Jul 2020 0 repositories listed
-
Image Captioning with Compositional Neural Module Networks10 Jul 2020 0 repositories listed
-
Eliminating Catastrophic Interference with Biased Competition3 Jul 2020 0 repositories listed
-
Visual Question Answering as a Multi-Task Problem3 Jul 2020 0 repositories listed
-
Scene Graph Reasoning for Visual Question Answering2 Jul 2020 0 repositories listed
-
The Impact of Explanations on AI Competency Prediction in VQA2 Jul 2020 0 repositories listed
-
Aligned Dual Channel Graph Convolutional Network for Visual Question Answering1 Jul 2020 0 repositories listed
-
Multimodal Neural Graph Memory Networks for Visual Question Answering1 Jul 2020 0 repositories listed
-
Towards Visual Dialog for Radiology1 Jul 2020 0 repositories listed
-
Improving VQA and its Explanations by Comparing Competing Explanations28 Jun 2020 0 repositories listed
-
Self-Segregating and Coordinated-Segregating Transformer for Focused Deep Multi-Modular Network for Visual Question Answering25 Jun 2020 0 repositories listed
-
Mucko: Multi-Layer Cross-Modal Knowledge Reasoning for Fact-based Visual Question Answering16 Jun 2020 0 repositories listed
-
ORD: Object Relationship Discovery for Visual Dialogue Generation15 Jun 2020 0 repositories listed
-
Exploring Weaknesses of VQA Models through Attribution Driven Insights11 Jun 2020 0 repositories listed
-
Estimating semantic structure for the VQA answer space10 Jun 2020 0 repositories listed
-
Counterfactual Vision and Language Learning1 Jun 2020 0 repositories listed
-
Multimodal grid features and cell pointers for Scene Text Visual Question Answering1 Jun 2020 0 repositories listed
-
TA-Student VQA: Multi-Agents Training by Self-Questioning1 Jun 2020 0 repositories listed
-
On the Value of Out-of-Distribution Testing: An Example of Goodhart's Law19 May 2020 0 repositories listed
-
Visual Relationship Detection using Scene Graphs: A Survey16 May 2020 0 repositories listed
-
Visual Question Answering with Prior Class Semantics4 May 2020 0 repositories listed
-
A Corpus for Visual Question Answering Annotated with Frame Semantic Information1 May 2020 0 repositories listed
-
Image Position Prediction in Multimodal Documents1 May 2020 0 repositories listed
-
A Novel Attention-based Aggregation Function to Combine Vision and Language27 Apr 2020 0 repositories listed
-
Visual Question Answering Using Semantic Information from Image Descriptions23 Apr 2020 0 repositories listed
-
Learning What Makes a Difference from Counterfactual Examples and Gradient Supervision20 Apr 2020 0 repositories listed
-
Knowledge-Based Visual Question Answering in Videos17 Apr 2020 0 repositories listed
-
Rephrasing visual questions by specifying the entropy of the answer distribution10 Apr 2020 0 repositories listed
-
Understanding Knowledge Gaps in Visual Question Answering: Implications for Gap Identification and Testing8 Apr 2020 0 repositories listed
-
Generating Rationales in Visual Question Answering4 Apr 2020 0 repositories listed
-
27 Mar 2020 0 repositories listed
-
Linguistically Driven Graph Capsule Network for Visual Question Reasoning23 Mar 2020 0 repositories listed
-
Visual Question Answering for Cultural Heritage22 Mar 2020 0 repositories listed
-
Normalized and Geometry-Aware Self-Attention Network for Image Captioning19 Mar 2020 0 repositories listed
-
RSVQA: Visual Question Answering for Remote Sensing Data16 Mar 2020 0 repositories listed
-
A Study on Multimodal and Interactive Explanations for Visual Question Answering1 Mar 2020 0 repositories listed
-
Unshuffling Data for Improved Generalization27 Feb 2020 0 repositories listed
-
On the General Value of Evidence, and Bilingual Scene-Text Visual Question Answering24 Feb 2020 0 repositories listed
-
VQA-LOL: Visual Question Answering under the Lens of Logic19 Feb 2020 0 repositories listed
-
CQ-VQA: Visual Question Answering on Categorized Questions17 Feb 2020 0 repositories listed
-
Component Analysis for Visual Question Answering Architectures12 Feb 2020 0 repositories listed
-
Uncertainty based Class Activation Maps for Visual Question Answering23 Jan 2020 0 repositories listed
-
Accuracy vs. Complexity: A Trade-off in Visual Question Answering Models20 Jan 2020 0 repositories listed
-
10 Jan 2020 0 repositories listed
-
Multi-Layer Content Interaction Through Quaternion Product For Visual Question Answering3 Jan 2020 0 repositories listed
-
Vision and Language: from Visual Perception to Content Creation26 Dec 2019 0 repositories listed
-
Deep Exemplar Networks for VQA and VQG19 Dec 2019 0 repositories listed
-
Towards Causal VQA: Revealing and Reducing Spurious Correlations by Invariant and Covariant Semantic Editing16 Dec 2019 0 repositories listed
-
9 Dec 2019 0 repositories listed
-
Weak Supervision helps Emergence of Word-Object Alignment and improves Vision-Language Tasks6 Dec 2019 0 repositories listed
-
Deep Bayesian Active Learning for Multiple Correct Outputs2 Dec 2019 0 repositories listed
-
A Free Lunch in Generating Datasets: Building a VQG and VQA System with Attention and Humans in the Loop30 Nov 2019 0 repositories listed
-
Assessing the Robustness of Visual Question Answering Models30 Nov 2019 0 repositories listed
-
Unsupervised Keyword Extraction for Full-sentence VQA23 Nov 2019 0 repositories listed
-
Explanation vs Attention: A Two-Player Game to Obtain Attention for VQA19 Nov 2019 0 repositories listed
-
Question-Conditioned Counterfactual Image Generation for VQA14 Nov 2019 0 repositories listed
-
Open-Ended Visual Question Answering by Multi-Modal Domain Adaptation11 Nov 2019 0 repositories listed
-
Multimodal Intelligence: Representation Learning, Information Fusion, and Applications10 Nov 2019 0 repositories listed
-
Are we asking the right questions in MovieQA?8 Nov 2019 0 repositories listed
-
Representing Movie Characters in Dialogues1 Nov 2019 0 repositories listed
-
YouMakeup: A Large-Scale Domain-Specific Multimodal Dataset for Fine-Grained Semantic Comprehension1 Nov 2019 0 repositories listed
-
Learning Rich Image Region Representation for Visual Question Answering29 Oct 2019 0 repositories listed
-
Enforcing Reasoning in Visual Commonsense Reasoning21 Oct 2019 0 repositories listed
-
Good, Better, Best: Textual Distractors Generation for Multiple-Choice Visual Question Answering via Reinforcement Learning21 Oct 2019 0 repositories listed
-
Neural Memory Plasticity for Anomaly Detection12 Oct 2019 0 repositories listed
-
Multi-modal Deep Analysis for Multimedia11 Oct 2019 0 repositories listed
-
Modulated Self-attention Convolutional Network for VQA8 Oct 2019 0 repositories listed
-
From Strings to Things: Knowledge-Enabled VQA Model That Can Read and Reason1 Oct 2019 0 repositories listed
-
SegEQA: Video Segmentation Based Visual Attention for Embodied Question Answering1 Oct 2019 0 repositories listed
-
On Incorporating Semantic Prior Knowledge in Deep Learning Through Embedding-Space Constraints30 Sep 2019 0 repositories listed
-
Learning to Recognize the Unseen Visual Predicates25 Sep 2019 0 repositories listed
-
On Incorporating Semantic Prior Knowlegde in Deep Learning Through Embedding-Space Constraints25 Sep 2019 0 repositories listed
-
UNITER: Learning UNiversal Image-TExt Representations25 Sep 2019 0 repositories listed
-
Why Does the VQA Model Answer No?: Improving Reasoning through Visual and Linguistic Inference25 Sep 2019 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.