{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/manually-annotated-spelling-error-corpus-for","title":"Manually Annotated Spelling Error Corpus for Amharic","arxiv_id":"2106.13521","date":"2021-06-25","proceeding":null,"authors":["Andargachew Mekonnen Gezmu","Tirufat Tesifaye Lema","Binyam Ephrem Seyoum","Andreas Nürnberger"],"abstract":"This paper presents a manually annotated spelling error corpus for Amharic, lingua franca in Ethiopia. The corpus is designed to be used for the evaluation of spelling error detection and correction. The misspellings are tagged as non-word and real-word errors. In addition, the contextual information available in the corpus makes it useful in dealing with both types of spelling errors.","url_abs":"https://arxiv.org/abs/2106.13521v1","url_pdf":"https://arxiv.org/pdf/2106.13521v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"manually-annotated-spelling-error-corpus-for","repo_url":"https://github.com/andmek/ErrorCorpus","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[],"methods":[],"datasets_introduced":[{"slug":"amharic-error-corpus","name":"Amharic Error Corpus","full_name":""}],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}