Datasets › VQA-CP

VQA-CP

Introduced by Aishwarya Agrawal et al. in Don't Just Assume; Look and Answer: Overcoming Priors for Visual Question Answering1 Jan 2018 archive 2025-07-28

The VQA-CP dataset was constructed by reorganizing VQA v2 such that the correlation between the question type and correct answer differs in the training and test splits. For example, the most common answer to questions starting with What sport… is tennis in the training set, but skiing in the test set. A model that guesses an answer primarily from the question will perform poorly.

Source: Unshuffling Data for Improved Generalization Image Source: https://arxiv.org/pdf/1712.00377.pdf

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Visual Question Answering (VQA) VQA-CP CSS Score 58.95 Counterfactual Samples Synthesizing for Robust Visual... yanxinzju/CSS-VQA +1 10 Compare

Papers archive 2025-07-28

9 shown of 9 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 13. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Greedy Gradient Ensemble for Robust Visual Question Answering 1 1 27 Jul 2021 ran 1 of 2 samples (1 unverified)
Removing Bias in Multi-modal Classifiers: Regularization by Maximizing Functional Entropies 1 2 21 Oct 2020 ran 0 of 1 samples (1 unverified; 1 pointer-only for licence)
Counterfactual Samples Synthesizing for Robust Visual Question Answering 2 1 14 Mar 2020 ran 3 of 3 samples (0 unverified; 3 pointer-only for licence)
Don't Take the Easy Way Out: Ensemble Based Methods for Avoiding Known Dataset Biases 3 1 9 Sep 2019 ran 3 of 6 samples (3 unverified; 3 pointer-only for licence)
Learning by Abstraction: The Neural State Machine 4 1 9 Jul 2019 ran 3 of 21 samples (18 unverified; 2 pointer-only for licence)
RUBi: Reducing Unimodal Biases in Visual Question Answering 1 1 24 Jun 2019 not harvested
Self-Critical Reasoning for Robust Visual Question Answering 1 1 24 May 2019 not harvested
MUREL: Multimodal Relational Reasoning for Visual Question Answering 1 1 25 Feb 2019 not harvested
Learning Visual Question Answering by Bootstrapping Hard Attention 1 1 1 Aug 2018 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • VQA-CP

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections