{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/to-err-is-human-but-llamas-can-learn-it-too","title":"To Err Is Human, but Llamas Can Learn It Too","arxiv_id":"2403.05493","date":"2024-03-08","proceeding":null,"authors":["Agnes Luhtaru","Taido Purason","Martin Vainikko","Maksym Del","Mark Fishel"],"abstract":"This study explores enhancing grammatical error correction (GEC) through artificial error generation (AEG) using language models (LMs). Specifically, we fine-tune Llama 2-based LMs for error generation and find that this approach yields synthetic errors akin to human errors. Next, we train GEC Llama models with the help of these artificial errors and outperform previous state-of-the-art error correction models, with gains ranging between 0.8 and 6 F0.5 points across all tested languages (German, Ukrainian, and Estonian). Moreover, we demonstrate that generating errors by fine-tuning smaller sequence-to-sequence models and prompting large commercial LMs (GPT-3.5 and GPT-4) also results in synthetic errors beneficially affecting error generation models.","url_abs":"https://arxiv.org/abs/2403.05493v2","url_pdf":"https://arxiv.org/pdf/2403.05493v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"to-err-is-human-but-llamas-can-learn-it-too","repo_url":"https://github.com/TartuNLP/gec-llm","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"ok"}}],"tasks":[{"task_slug":"grammatical-error-correction","task_name":"Grammatical Error Correction"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/grammatical-error-correction-on-estgec-l2","task":"Grammatical Error Correction","dataset":"EstGEC-L2","model":"Llama + 1M BT + gold","rank_in_archive_order":1,"of":1,"metrics":{"F0.5":"69.97"},"uses_additional_data":true},{"leaderboard":"/sota/grammatical-error-correction-on-falko-merlin","task":"Grammatical Error Correction","dataset":"Falko-MERLIN","model":"Llama + 1M BT + gold","rank_in_archive_order":1,"of":6,"metrics":{"F0.5":"76.75"},"uses_additional_data":true},{"leaderboard":"/sota/grammatical-error-correction-on-ua-gec","task":"Grammatical Error Correction","dataset":"UA-GEC","model":"Llama + 1M BT + gold","rank_in_archive_order":1,"of":5,"metrics":{"F0.5":"74.09"},"uses_additional_data":true}],"syntology":{"syntology_url":null,"atlas_url":"https://app.syntology.ai/?focus=2403.05493","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}