HateBERT: Retraining BERT for Abusive Language Detection in English

Tommaso Caselli, Valerio Basile, Jelena Mitrović, Michael Granitzer

2020-10-23ACL (WOAH) 2021 8Abusive Language Hate Speech Detection Language Modelling

Abstract

In this paper, we introduce HateBERT, a re-trained BERT model for abusive language detection in English. The model was trained on RAL-E, a large-scale dataset of Reddit comments in English from communities banned for being offensive, abusive, or hateful that we have collected and made available to the public. We present the results of a detailed comparison between a general pre-trained language model and the abuse-inclined version obtained by retraining with posts from the banned communities on three English datasets for offensive, abusive language and hate speech detection tasks. In all datasets, HateBERT outperforms the corresponding general BERT model. We also discuss a battery of experiments comparing the portability of the generic pre-trained language model and its corresponding abusive language-inclined counterpart across the datasets, indicating that portability is affected by compatibility of the annotated phenomena.

Results

Task	Dataset	Metric	Value	Model
Abuse Detection	HatEval	Macro F1	0.494	HateBERT
Abuse Detection	HatEval	Macro F1	0.48	BERT
Abuse Detection	AbusEval	Macro F1	0.742	HateBERT
Abuse Detection	AbusEval	Macro F1	0.724	BERT
Abuse Detection	OffensEval 2019	Macro F1	0.805	HateBERT
Abuse Detection	OffensEval 2019	Macro F1	0.803	BERT
Hate Speech Detection	HatEval	Macro F1	0.494	HateBERT
Hate Speech Detection	HatEval	Macro F1	0.48	BERT
Hate Speech Detection	AbusEval	Macro F1	0.742	HateBERT
Hate Speech Detection	AbusEval	Macro F1	0.724	BERT
Hate Speech Detection	OffensEval 2019	Macro F1	0.805	HateBERT
Hate Speech Detection	OffensEval 2019	Macro F1	0.803	BERT

Related Papers

Visual-Language Model Knowledge Distillation Method for Image Quality Assessment2025-07-21 Making Language Model a Hierarchical Classifier and Generator2025-07-17 VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning2025-07-17 The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations2025-07-17 Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities2025-07-17 Assay2Mol: large language model-based drug design using BioAssay context2025-07-16 Describe Anything Model for Visual Question Answering on Text-rich Images2025-07-16 InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing2025-07-16