{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/spoken-squad-a-study-of-mitigating-the-impact","title":"Spoken SQuAD: A Study of Mitigating the Impact of Speech Recognition Errors on Listening Comprehension","arxiv_id":"1804.00320","date":"2018-04-01","proceeding":null,"authors":["Chia-Hsuan Li","Szu-Lin Wu","Chi-Liang Liu","Hung-Yi Lee"],"abstract":"Reading comprehension has been widely studied. One of the most representative\nreading comprehension tasks is Stanford Question Answering Dataset (SQuAD), on\nwhich machine is already comparable with human. On the other hand, accessing\nlarge collections of multimedia or spoken content is much more difficult and\ntime-consuming than plain text content for humans. It's therefore highly\nattractive to develop machines which can automatically understand spoken\ncontent. In this paper, we propose a new listening comprehension task - Spoken\nSQuAD. On the new task, we found that speech recognition errors have\ncatastrophic impact on machine comprehension, and several approaches are\nproposed to mitigate the impact.","url_abs":"http://arxiv.org/abs/1804.00320v1","url_pdf":"http://arxiv.org/pdf/1804.00320v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"spoken-squad-a-study-of-mitigating-the-impact","repo_url":"https://github.com/chiahsuan156/Spoken-SQuAD","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok","spdx":"NOASSERTION"}},{"paper_slug":"spoken-squad-a-study-of-mitigating-the-impact","repo_url":"https://github.com/chia-hsuan-lee/spoken-squad","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok","spdx":"NOASSERTION"}},{"paper_slug":"spoken-squad-a-study-of-mitigating-the-impact","repo_url":"https://github.com/maikezuefle/contr-pretraining","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"ok","spdx":"Apache-2.0"}}],"tasks":[{"task_slug":"question-answering","task_name":"Question Answering"},{"task_slug":"reading-comprehension","task_name":"Reading Comprehension"},{"task_slug":"speech-recognition","task_name":"Speech Recognition"},{"task_slug":"spoken-language-understanding","task_name":"Spoken Language Understanding"},{"task_slug":"speech-recognition-1","task_name":"speech-recognition"}],"methods":[],"datasets_introduced":[{"slug":"spoken-squad","name":"Spoken-SQuAD","full_name":null}],"methods_introduced":[],"results":[{"leaderboard":"/sota/spoken-language-understanding-on-spoken-squad","task":"Spoken Language Understanding","dataset":"Spoken-SQuAD","model":"Baseline","rank_in_archive_order":4,"of":4,"metrics":{"F1 score":"58.71"},"uses_additional_data":false}],"syntology":{"syntology_url":null,"atlas_url":"https://app.syntology.ai/?focus=1804.00320","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}