{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/safespeech-a-comprehensive-and-interactive","title":"SafeSpeech: A Comprehensive and Interactive Tool for Analysing Sexist and Abusive Language in Conversations","arxiv_id":"2503.06534","date":"2025-03-09","proceeding":null,"authors":["Xingwei Tan","Chen Lyu","Hafiz Muhammad Umer","Sahrish Khan","Mahathi Parvatham","Lois Arthurs","Simon Cullen","Shelley Wilson","Arshad Jhumka","Gabriele Pergola"],"abstract":"Detecting toxic language including sexism, harassment and abusive behaviour, remains a critical challenge, particularly in its subtle and context-dependent forms. Existing approaches largely focus on isolated message-level classification, overlooking toxicity that emerges across conversational contexts. To promote and enable future research in this direction, we introduce SafeSpeech, a comprehensive platform for toxic content detection and analysis that bridges message-level and conversation-level insights. The platform integrates fine-tuned classifiers and large language models (LLMs) to enable multi-granularity detection, toxic-aware conversation summarization, and persona profiling. SafeSpeech also incorporates explainability mechanisms, such as perplexity gain analysis, to highlight the linguistic elements driving predictions. Evaluations on benchmark datasets, including EDOS, OffensEval, and HatEval, demonstrate the reproduction of state-of-the-art performance across multiple tasks, including fine-grained sexism detection.","url_abs":"https://arxiv.org/abs/2503.06534v1","url_pdf":"https://arxiv.org/pdf/2503.06534v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"abusive-language","task_name":"Abusive Language"},{"task_slug":null,"task_name":"Conversation Summarization"}],"methods":[{"method_slug":"focus","method_name":"Focus"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2503.06534","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2503.06534"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"deterministic:regex_extraction","url":"https://github.com/fastapi/fastapi","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"unverified":3},"by_repo_kind":{"found_in_text":{"samples":3,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"d746f15d84581faa","entry":"Default","repo":"fastapi/fastapi","repo_kind":"found_in_text","path":"fastapi/datastructures.py","file_url":"https://github.com/fastapi/fastapi/blob/HEAD/fastapi/datastructures.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d746f15d84581faa"}},{"code_sha256_prefix":"c73d8d79c6958378","entry":"decimal_encoder","repo":"fastapi/fastapi","repo_kind":"found_in_text","path":"fastapi/encoders.py","file_url":"https://github.com/fastapi/fastapi/blob/HEAD/fastapi/encoders.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"c73d8d79c6958378"}},{"code_sha256_prefix":"b28e9069aad6a6ff","entry":"isoformat","repo":"fastapi/fastapi","repo_kind":"found_in_text","path":"fastapi/encoders.py","file_url":"https://github.com/fastapi/fastapi/blob/HEAD/fastapi/encoders.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"b28e9069aad6a6ff"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}