Papers › ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

19 Mar 2022Findings (ACL) 2022 5arXiv:2203.10244archive 2025-07-28

Ahmed Masry, Do Xuan Long, Jia Qing Tan, Shafiq Joty, Enamul Hoque

Charts are very popular for analyzing data. When exploring charts, people often ask a variety of complex reasoning questions that involve several logical and arithmetic operations. They also commonly refer to visual features of a chart in their questions. However, most existing datasets do not focus on such complex reasoning questions as their questions are template-based and answers come from a fixed-vocabulary. In this work, we present a large-scale benchmark covering 9.6K human-written questions as well as 23.1K questions generated from human-written chart summaries. To address the unique challenges in our benchmark involving visual and logical reasoning over charts, we present two transformer-based models that combine visual features and the data table of the chart in a unified way to answer questions. While our models achieve the state-of-the-art results on the previous datasets as well as on our benchmark, the evaluation also reveals several challenges in answering complex reasoning questions.

PaperPDFConference PDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

vis-nlp/chartqa officialmentioned in paperpytorchGPL-3.0 report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Chart Question AnsweringLogical ReasoningQuestion Answering

Datasets

Introduced by this paper, per the archive.

ChartQA

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Chart Question Answering ChartQA VisionTapas-OCR 1:1 Accuracy 45.5 #25 of 27 Archive leaderboard report
Chart Question Answering PlotQA VL-T5-OCR 1:1 Accuracy 66.0 #4 of 6 Archive leaderboard report
Chart Question Answering PlotQA VisionTapas-OCR 1:1 Accuracy 53.9 #6 of 6 Archive leaderboard report
Chart Question Answering RealCQA crct - baseline 1:1 Accuracy 0.178733575026565 #1 of 5 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections