Papers › The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents

The StatCan Dialogue Dataset: Retrieving Data Tables through Conversations with Genuine Intents

3 Apr 2023arXiv:2304.01412archive 2025-07-28

Xing Han Lu, Siva Reddy, Harm de Vries

We introduce the StatCan Dialogue Dataset consisting of 19,379 conversation turns between agents working at Statistics Canada and online users looking for published data tables. The conversations stem from genuine intents, are held in English or French, and lead to agents retrieving one of over 5000 complex data tables. Based on this dataset, we propose two tasks: (1) automatic retrieval of relevant tables based on a on-going conversation, and (2) automatic generation of appropriate agent responses at each turn. We investigate the difficulty of each task by establishing strong baselines. Our experiments on a temporal data split reveal that all models struggle to generalize to future conversations, as we observe a significant drop in performance across both tasks when we move from the validation to the test set. In addition, we find that response generation models struggle to decide when to return a table. Considering that the tasks pose significant challenges to existing models, we encourage the community to develop models for our task, which can be directly used to help knowledge workers find relevant tables for live chat users.

PaperPDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

McGill-NLP/statcan-dialogue-dataset officialmentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Dialogue GenerationRetrievalTable Retrieval

Datasets

Introduced by this paper, per the archive.

Statcan Dialogue Dataset

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Table Retrieval Statcan Dialogue Dataset DPR (retrieving basic info + member items) Recall@10 46.2 #1 of 5 Archive leaderboard report
Table Retrieval Statcan Dialogue Dataset DPR (retrieving basic info) Recall@10 45.0 #2 of 5 Archive leaderboard report
Table Retrieval Statcan Dialogue Dataset DPR (retrieving title) Recall@10 43.8 #3 of 5 Archive leaderboard report
Table Retrieval Statcan Dialogue Dataset TAPAS-NQ (retrieving truncated table) Recall@10 30.0 #4 of 5 Archive leaderboard report
Table Retrieval Statcan Dialogue Dataset TAPAS (retrieving truncated table) Recall@10 22.1 #5 of 5 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

AdafactorAttentionAttention DropoutBPEDense ConnectionsDropoutGated Linear UnitInverse Square Root ScheduleLayer NormalizationLinear LayerMulti-Head AttentionResidual ConnectionSentencePieceSoftmaxT5

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections