Browse State-of-the-Art › Hate Speech Detection
Hate Speech Detection
203 papers with code · 15 benchmarks · 47 datasets archive 2025-07-28
Hate speech detection is the task of detecting if communication such as text, audio, and so on contains hatred and or encourages violence towards a person or a group of people. This is usually based on prejudice against 'protected characteristics' such as their ethnicity, gender, sexual orientation, religion, age et al. Some example benchmarks are ETHOS and HateXplain. Models can be evaluated with metrics like the F-score or F-measure.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
15 leaderboard tables shown for this task, 15 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 15 until expanded.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
47 datasets whose archive record lists this task, ordered by the archive's paper count. 30 shown of 47 until expanded.
Subtasks archive 2025-07-28
3 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 203 papers with code (507 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
2 Oct 2019 37 repositories listed Syntology ran 19 of 27 samples · 8 unverifiedAs Transfer Learning from large-scale pre-trained models becomes more prevalent in Natural Language Processing (NLP), operating these large models in on-the-edge and/or under constrained computational training or…
-
2 May 2022 11 repositories listed Syntology ran 6 of 24 samples · 18 unverified · 16 pointer-only (licence)Large language models, which are often trained for hundreds of thousands of compute days, have shown remarkable capabilities for zero- and few-shot learning.
-
11 Mar 2017 8 repositories listed Syntology ran 0 of 12 samples · 12 unverifiedWe train a multi-class classifier to distinguish between these different categories.
-
18 Dec 2020 6 repositories listedWe also observe that models, which utilize the human rationales for training, perform better in reducing unintended bias towards target communities.
-
30 Aug 2018 4 repositories listedHowever, this dataset has not been comprehensively studied to its potential.
-
17 Jun 2024 3 repositories listedThis paper serves as a crucial resource, guiding researchers and practitioners in harnessing the potential of LLMs for data annotation, thereby fostering advancements in this critical field.
-
31 Dec 2020 3 repositories listedDetecting online hate is a difficult task that even state-of-the-art models struggle with.
-
14 Apr 2020 3 repositories listedHate speech detection is a challenging problem with most of the datasets available in only one language: English.
-
11 Dec 2024 2 repositories listedThis paper explores hate speech detection in Devanagari-scripted languages, focusing on Hindi and Nepali, for Subtask B of the CHIPSAL@COLING 2025 Shared Task.
-
28 Mar 2024 2 repositories listedFinally, owing to the modest performance of HSD systems in real-world conditions, we find that content moderators would need to review about ten thousand Nigerian tweets flagged as hateful daily to moderate 60% of all…
-
9 Feb 2024 2 repositories listedThis study details our approach for the CASE 2024 Shared Task on Climate Activism Stance and Hate Event Detection, focusing on Hate Speech Detection, Hate Speech Target Identification, and Stance Detection as…
-
19 Jan 2024 2 repositories listedWith the recent surge and exponential growth of social media usage, scrutinizing social media content for the presence of any hateful content is of utmost importance.
-
29 Nov 2023 2 repositories listedWide us of this language on social media platforms such as Twitter, Instagram, or Tiktok and strategic position of the country in the world politics makes it appealing for the social network researchers and industry.
-
6 Oct 2023 2 repositories listedWith the growth of online services, the need for advanced text classification algorithms, such as sentiment analysis and biased text detection, has become increasingly evident.
-
2 Aug 2022 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedWith ferret, users can visualize and compare transformers-based models output explanations using state-of-the-art XAI methods on any free-text or existing XAI corpora.
-
3 Jan 2022 2 repositories listedDetecting and labeling stance in social media text is strongly motivated by hate speech detection, poll prediction, engagement forecasting, and concerted propaganda detection.
-
20 Dec 2021 2 repositories listedLarge-scale generative language models such as GPT-3 are competitive few-shot learners.
-
23 Mar 2021 2 repositories listedGiven this capacity, we are interested in whether large language models can be used to identify hate speech and classify text as sexist or racist.
-
22 Mar 2021 2 repositories listedOn social medias, hate speech has become a critical problem for social network users.
-
31 Dec 2020 2 repositories listedWe provide a new dataset of ~40, 000 entries, generated and labelled by trained annotators over four rounds of dynamic data creation.
-
Multilingual Twitter Corpus and Baselines for Evaluating Demographic Bias in Hate Speech Recognition24 Feb 2020 2 repositories listed Syntology ran 2 of 13 samples · 11 unverifiedExisting research on fairness evaluation of document classification models mainly uses synthetic monolingual data without ground truth for author demographic attributes.
-
28 Oct 2019 2 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedTo address these needs, in this study we introduce a novel transfer learning approach based on an existing pre-trained language model called BERT (Bidirectional Encoder Representations from Transformers).
-
17 Dec 2018 2 repositories listedWith the online proliferation of hate speech, there is an urgent need for systems that can detect such harmful content.
-
12 Sep 2018 2 repositories listedHate speech is commonly defined as any communication that disparages a target group of people based on some characteristic such as race, colour, ethnicity, gender, sexual orientation, nationality, religion, or other…
-
20 Oct 2017 2 repositories listedIn the wake of a polarizing election, the cyber world is laden with hate speech.
-
26 May 2025 1 repository listedImplicit hate speech detection is challenging due to its subtlety and reliance on contextual interpretation rather than explicit offensive words.
-
22 May 2025 1 repository listedSocial media and online forums are increasingly becoming popular.
-
4 May 2025 1 repository listedIn this paper, we examine various state-of-the-art LLMs to understand their behaviour in different personalisation scenarios, specifically focusing on hate speech.
-
7 Mar 2025 1 repository listedAlgorithmic hate speech detection faces significant challenges due to the diverse definitions and datasets used in research and practice.
-
26 Feb 2025 1 repository listedIn this study, we systematically investigate the performance of LLMs on detecting hate speech across multilingual datasets and diverse geographic contexts.
Syntology lines on 6 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections