Browse State-of-the-Art › Abusive Language
Abusive Language
51 papers with code · 0 benchmarks · 9 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
9 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 51 papers with code (166 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
29 May 2019 5 repositories listed Syntology ran 0 of 4 samples · 4 unverifiedTechnologies for abusive language detection are being developed and applied with little consideration of their potential biases.
-
30 Aug 2018 4 repositories listedHowever, this dataset has not been comprehensively studied to its potential.
-
17 May 2025 1 repository listedThis work presents a large-scale human-annotated multi-task benchmark dataset for abusive language detection in Tigrinya social media with joint annotations for three tasks: abusiveness, sentiment, and topic…
-
14 Jan 2025 1 repository listedThese limitations are mainly due to the lack of high-quality data in the local languages and the failure to include local communities in the collection, annotation, and moderation processes.
-
18 Dec 2024 1 repository listedAutomatic detection of hate and abusive language is essential to combat its online spread.
-
2 Dec 2024 1 repository listedWe investigate the potential of pre-trained audio representations for detecting abusive language in low-resource languages, in this case, in Indian languages using Few Shot Learning (FSL).
-
30 Sep 2024 1 repository listedTo address this gap, we present a comprehensive dataset of user replies from Telegram channels associated with political extremism or the cyberbullying of public officials in the United States.
-
28 Apr 2024 1 repository listedHowever, little or none has been done in the detection of abusive language and hate speech in Nigeria.
-
2 Apr 2024 1 repository listedOnline gender-based harassment is a widespread issue limiting the free expression and participation of women and marginalized genders in digital spaces.
-
4 Jul 2023 1 repository listedClassifiers tend to learn a false causal relationship between an over-represented concept and a label, which can result in over-reliance on the concept and compromised classification accuracy.
-
30 Nov 2022 1 repository listed Syntology ran 3 of 4 samples · 1 unverifiedWe introduce two rationale-integrated BERT-based architectures (the RGFS models) and evaluate our systems over five different abusive language datasets, finding that in the few-shot classification setting, RGFS-based…
-
21 Sep 2022 1 repository listedAnnotating abusive language is expensive, logistically complex and creates a risk of psychological harm.
-
14 Jul 2022 1 repository listedIn this paper, we present two shared tasks of abusive and threatening language detection for the Urdu language which has more than 170 million speakers worldwide.
-
1 Jul 2022 1 repository listedWe address the task of distinguishing implicitly abusive sentences on identity groups (“Muslims contaminate our planet”) from other group-related negative polar sentences (“Muslims despise terrorism”).
-
1 Jul 2022 1 repository listedThis paper presents a comprehensive corpus for the study of socially unacceptable language in Dutch.
-
29 May 2022 1 repository listedIn recent times, the detection of hate-speech, offensive, or abusive language in online media has become an important topic in NLP research due to the exponential growth of social media and the propagation of such…
-
1 May 2022 1 repository listedTo address the automatic detection of abusive languages in online platforms, this paper describes the models submitted by our team - MUCIC to the shared task on “Abusive Comment Detection in Tamil-ACL 2022”.
-
1 May 2022 1 repository listedTamil has also emerged as a popular language for use on social media platforms due to the increasing penetration of vernacular media like Sharechat and Moj, which focus more on local Indian languages than English and…
-
Data Bootstrapping Approaches to Improve Low Resource Abusive Language Detection for Indic Languages26 Apr 2022 1 repository listedIn this paper, to bridge the gap, we demonstrate a large-scale analysis of multilingual abusive speech in Indic languages.
-
5 Apr 2022 1 repository listedRobustness of machine learning models on ever-changing real-world data is critical, especially for applications affecting human well-being such as content moderation.
-
27 Nov 2021 1 repository listedIn this FIRE 2021 shared task - "HASOC- Abusive and Threatening language detection in Urdu" the organizers propose an abusive language detection dataset in Urdu along with threatening language detection.
-
1 Nov 2021 1 repository listedAs users in online communities suffer from severe side effects of abusive language, many researchers attempted to detect abusive texts from social media, presenting several datasets for such detection.
-
20 Sep 2021 1 repository listedWe find that the distribution of abuse is vastly different compared to other commonly used datasets, with more sexually tinted aggression towards the virtual persona of these systems.
-
1 Sep 2021 1 repository listed
-
1 Sep 2021 1 repository listedA prevalent form of bias in hate speech and abusive language datasets is annotator bias caused by the annotator’s subjective perception and the complexity of the annotation task.
-
1 Aug 2021 1 repository listedOnline misogyny, a category of online abusive language, has serious and harmful social consequences.
-
1 Aug 2021 1 repository listedAs socially unacceptable language become pervasive in social media platforms, the need for automatic content moderation become more pressing.
-
1 Aug 2021 1 repository listedOur evaluation shows that, although BERT-based classifiers achieve high accuracy levels on a variety of natural language processing tasks, they perform very poorly as regards fairness and bias, in particular on samples…
-
1 Aug 2021 1 repository listedMainstream research on hate speech focused so far predominantly on the task of classifying mainly social media posts with respect to predefined typologies of rather coarse-grained hate speech categories.
-
1 Aug 2021 1 repository listedThese standards include definitions of what is considered abusive language, annotation guidelines and reporting on the process.
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections