{"url":"/dataset/persianqa","name":"PersianQA","full_name":"Persian Question Answering Dataset","description_markdown":"# PersianQA: a dataset for Persian Question Answering\r\n\r\nPersian Question Answering (PersianQA) Dataset is a reading comprehension\r\ndataset on [Persian Wikipedia](https://fa.wikipedia.org/). The crowd-sourced\r\nthe dataset consists of more than 9,000 entries. Each entry can be either an\r\n_impossible-to-answer_ or a question with one or more answers spanning in the\r\npassage (the _context_) from which the questioner proposed the question.\r\nMuch like the SQuAD2.0 dataset, the impossible or _unanswerable_ questions can be\r\nutilized to create a system which \"knows that it doesn't know the answer\".\r\n\r\nMoreover, the dataset has 900 test data available. On top of that, the very\r\nfirst models trained on the dataset, Transformers, are available online.\r\n\r\nAll the crowd workers of the dataset are native Persian speakers. Also, it worth\r\nmentioning that the contexts are collected from all categories of the Wiki\r\n(Historical, Religious, Geography, Science, etc).","description_withheld":null,"homepage":"https://github.com/sajjjadayobi/PersianQA","introduced_date":"2021-04-29","introduced_date_note":null,"introduced_by":null,"license":{"name":"GPL","url":"https://github.com/sajjjadayobi/PersianQA/blob/main/LICENSE"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Question Answering","url":"/task/question-answering","datasets_with_task":"/datasets/task/question-answering"},{"name":"Reading Comprehension","url":"/task/reading-comprehension","datasets_with_task":"/datasets/task/reading-comprehension"},{"name":"Machine Reading Comprehension","url":"/task/machine-reading-comprehension","datasets_with_task":"/datasets/task/machine-reading-comprehension"}],"languages":[{"name":"Persian","url":"/datasets/language/persian"},{"name":"Iranian Persian","url":"/datasets/language/iranian-persian"}],"variants":["PersianQA"],"data_loaders":[{"repo":"https://github.com/sajjjadayobi/PersianQA","url":"https://github.com/sajjjadayobi/PersianQA","frameworks":[]}],"num_papers_in_archive":1,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}