{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/qasports-a-question-answering-dataset-about","title":"QASports: A Question Answering Dataset about Sports","arxiv_id":null,"date":"2023-09-25","proceeding":"Simpósio Brasileiro de Bancos de Dados - Dataset Showcase Workshop 2023 9","authors":["Pedro Calciolari Jardim","Leonardo Mauro Pereira Moraes","Cristina Dutra Aguiar"],"abstract":"Sport is one of the most popular and revenue-generating forms of entertainment. Therefore, analyzing data related to this domain introduces several opportunities for Question Answering (QA) systems, such as supporting tactical decision-making. But, to develop and evaluate QA systems, researchers and developers need datasets that contain questions and their corresponding answers. In this paper, we focus on this issue. We propose QASports, the first large sports question answering dataset for extractive answer questions. QASports contains more than 1.5 million triples of questions, answers, and context about three popular sports: soccer, American football, and basketball. We describe the QASports processes of data collection and questions and answers generation. We also describe the characteristics of the QASports data. Furthermore, we analyze the sources used to obtain raw data and investigate the usability of QASports by issuing \"wh-queries\". Moreover, we describe scenarios for using QASports, highlighting its importance for training and evaluating QA systems.","url_abs":"https://doi.org/10.5753/dsw.2023.233602","url_pdf":"https://sol.sbc.org.br/index.php/dsw/article/view/25500/25320","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"qasports-a-question-answering-dataset-about","repo_url":"https://github.com/leomaurodesenv/qasports-dataset-scripts","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[{"task_slug":"decision-making","task_name":"Decision Making"},{"task_slug":"question-answering","task_name":"Question Answering"}],"methods":[{"method_slug":null,"method_name":"American"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}