{"url":"/dataset/gepade","name":"GePaDe","full_name":null,"description_markdown":"This dataset encompasses 265 speeches (over 200,000 tokens) from the German Bundestag, primarily from the 19th legislative term (2017-2021), given by 195 distinct speakers representing 6 political parties.\r\n\r\nThe data was annotated to perform a semantic role labeling task, namely to identify who said what to whom (speaker attribution). Cues (triggers) were annotated that are associated with events of speech, writing, or thought. Additionally, the arguments (roles) of each trigger have been annotated, encompassing the SOURCE, ADDRESSEE, MESSAGE, MEDIUM, TOPIC, and EVIDENCE related to the speech event.\r\n\r\nThe dataset was introduced in the international GermEval 2023 Shared Task on Speaker Attribution in Newswire and Parliamentary Debates (SpkAtt-2023) to evaluate the quality of systems for automated identification of cues and associated roles.\r\n\r\nReference\r\n\r\nRehbein, I. et al, Overview of the GermEval 2023 Shared Task on Speaker Attribution in Newswire and Parliamentary Debates, https://github.com/umanlp/SpkAtt-2023/blob/master/doc/SpkAtt2023-proceedings.pdf","description_withheld":null,"homepage":"https://github.com/umanlp/SpkAtt-2023/tree/master","introduced_date":"2023-04-01","introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Semantic Role Labeling","url":"/task/semantic-role-labeling","datasets_with_task":"/datasets/task/semantic-role-labeling"},{"name":"Speaker Attribution in German Parliamentary Debates (GermEval 2023, subtask 1)","url":"/task/speaker-attribution-in-german-parliamentary","datasets_with_task":"/datasets/task/speaker-attribution-in-german-parliamentary"},{"name":"Speaker Attribution in German Parliamentary Debates (GermEval 2023, subtask 2)","url":"/task/speaker-attribution-in-german-parliamentary-1","datasets_with_task":"/datasets/task/speaker-attribution-in-german-parliamentary-1"}],"languages":[{"name":"German","url":"/datasets/language/german"}],"variants":["GePaDe"],"data_loaders":[],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[{"leaderboard":"/sota/speaker-attribution-in-german-parliamentary","task":"Speaker Attribution in German Parliamentary Debates (GermEval 2023, subtask 1)","dataset_variant":"GePaDe","rows":1,"metrics":["F1"],"first_row_in_archive_order":{"model":"Llama 2 70 B QLoRa adapted","paper":"/paper/speaker-attribution-in-german-parliamentary","metrics":{"F1":"0.813"},"code_links":[{"title":"umanlp/spkatt-2023","url":"https://github.com/umanlp/spkatt-2023"},{"title":"dslaborg/germeval2023","url":"https://github.com/dslaborg/germeval2023"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/speaker-attribution-in-german-parliamentary-1","task":"Speaker Attribution in German Parliamentary Debates (GermEval 2023, subtask 2)","dataset_variant":"GePaDe","rows":1,"metrics":["F1"],"first_row_in_archive_order":{"model":"Llama 2 70 B QLoRa adapted","paper":"/paper/speaker-attribution-in-german-parliamentary","metrics":{"F1":"0.891"},"code_links":[{"title":"umanlp/spkatt-2023","url":"https://github.com/umanlp/spkatt-2023"},{"title":"dslaborg/germeval2023","url":"https://github.com/dslaborg/germeval2023"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"}],"papers_with_a_benchmark_row":[{"paper":"/paper/speaker-attribution-in-german-parliamentary","title":"Speaker attribution in German parliamentary debates with QLoRA-adapted large language models","date":"2023-09-18","rows_on_this_dataset":2,"code_links":2,"syntology":null}],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}