Papers › Cause and Effect in Governmental Reports: Two Data Sets for Causality Detection in Swedish

Cause and Effect in Governmental Reports: Two Data Sets for Causality Detection in Swedish

1 Jun 2022PoliticalNLP (LREC) 2022 6archive 2025-07-28

Luise Dürlich, Sebastian Reimann, Gustav Finnveden, Joakim Nivre, Sara Stymne

Causality detection is the task of extracting information about causal relations from text. It is an important task for different types of document analysis, including political impact assessment. We present two new data sets for causality detection in Swedish. The first data set is annotated with binary relevance judgments, indicating whether a sentence contains causality information or not. In the second data set, sentence pairs are ranked for relevance with respect to a causality query, containing a specific hypothesized cause and/or effect. Both data sets are carefully curated and mainly intended for use as test data. We describe the data sets and their annotation, including detailed annotation guidelines. In addition, we present pilot experiments on cross-lingual zero-shot and few-shot causality detection, using training data from English and German.

PaperPDFCode

Code

uppsalanlp/sou-corpus officialmentioned in paper report
uppsalanlp/swedish-causality-datasets officialmentioned in paper report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Sentence

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections