{"url":"/dataset/swefaq-2-0","name":"SweFAQ","full_name":null,"description_markdown":"The **SweFAQ** dataset is a collection of frequently asked questions from Swedish authorities' websites with shuffled answers. It was created by Aleksandrs Berdicevskis and is published by Språkbanken Text. The dataset is a part of the SuperLim collection.\r\n\r\nHere are some key details about the dataset:\r\n- It contains **976 question-answer pairs** and **100 categories**.\r\n- The number of QA pairs in a category varies.\r\n- The answers are randomly shuffled within each category.\r\n- The dataset is in **Swedish**.\r\n- The format is JSON Lines, with one item per line. Each item contains a category ID, a question, an array of all answers in this category (this value is the same for all items within a category), and a label index that indicates the correct answer (numbering starts at 0).\r\n- The dataset is intended for tasks such as Machine Learning, Question Answering, and Evaluation of language models.\r\n- The task is to match questions and answers.","description_withheld":null,"homepage":"https://spraakbanken.gu.se/en/resources/swefaq","introduced_date":null,"introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[],"tasks":[],"languages":[],"variants":["SweFAQ"],"data_loaders":[],"num_papers_in_archive":0,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}