{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/convabuse-data-analysis-and-benchmarks-for","title":"ConvAbuse: Data, Analysis, and Benchmarks for Nuanced Abuse Detection in Conversational AI","arxiv_id":"2109.09483","date":"2021-09-20","proceeding":null,"authors":["Amanda Cercas Curry","Gavin Abercrombie","Verena Rieser"],"abstract":"We present the first English corpus study on abusive language towards three conversational AI systems gathered \"in the wild\": an open-domain social bot, a rule-based chatbot, and a task-based system. To account for the complexity of the task, we take a more `nuanced' approach where our ConvAI dataset reflects fine-grained notions of abuse, as well as views from multiple expert annotators. We find that the distribution of abuse is vastly different compared to other commonly used datasets, with more sexually tinted aggression towards the virtual persona of these systems. Finally, we report results from bench-marking existing models against this data. Unsurprisingly, we find that there is substantial room for improvement with F1 scores below 90%.","url_abs":"https://arxiv.org/abs/2109.09483v1","url_pdf":"https://arxiv.org/pdf/2109.09483v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"convabuse-data-analysis-and-benchmarks-for","repo_url":"https://github.com/amandacurry/convabuse","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"none","reach":null}],"tasks":[{"task_slug":"abuse-detection","task_name":"Abuse Detection"},{"task_slug":"abusive-language","task_name":"Abusive Language"},{"task_slug":"chatbot","task_name":"Chatbot"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2109.09483","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}