Papers › Aggressive Language Identification Using Word Embeddings and Sentiment Features

Aggressive Language Identification Using Word Embeddings and Sentiment Features

1 Aug 2018COLING 2018 8archive 2025-07-28

Constantin Or{\u{a}}san

This paper describes our participation in the First Shared Task on Aggression Identification. The method proposed relies on machine learning to identify social media texts which contain aggression. The main features employed by our method are information extracted from word embeddings and the output of a sentiment analyser. Several machine learning methods and different combinations of features were tried. The official submissions used Support Vector Machines and Random Forests. The official evaluation showed that for texts similar to the ones in the training dataset Random Forests work best, whilst for texts which are different SVMs are a better choice. The evaluation also showed that despite its simplicity the method performs well when compared with more elaborated methods.

PaperPDFCode

Code

dinel/aggression_identification officialmentioned in paper report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Aggression IdentificationBIG-bench Machine LearningLanguage IdentificationWord Embeddings

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

CharacterBERT

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections