TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/HateBERT: Retraining BERT for Abusive Language Detection i...

HateBERT: Retraining BERT for Abusive Language Detection in English

Tommaso Caselli, Valerio Basile, Jelena Mitrović, Michael Granitzer

2020-10-23ACL (WOAH) 2021 8Abusive LanguageHate Speech DetectionLanguage Modelling
PaperPDFCode(official)

Abstract

In this paper, we introduce HateBERT, a re-trained BERT model for abusive language detection in English. The model was trained on RAL-E, a large-scale dataset of Reddit comments in English from communities banned for being offensive, abusive, or hateful that we have collected and made available to the public. We present the results of a detailed comparison between a general pre-trained language model and the abuse-inclined version obtained by retraining with posts from the banned communities on three English datasets for offensive, abusive language and hate speech detection tasks. In all datasets, HateBERT outperforms the corresponding general BERT model. We also discuss a battery of experiments comparing the portability of the generic pre-trained language model and its corresponding abusive language-inclined counterpart across the datasets, indicating that portability is affected by compatibility of the annotated phenomena.

Results

TaskDatasetMetricValueModel
Abuse DetectionHatEvalMacro F10.494HateBERT
Abuse DetectionHatEvalMacro F10.48BERT
Abuse DetectionAbusEvalMacro F10.742HateBERT
Abuse DetectionAbusEvalMacro F10.724BERT
Abuse DetectionOffensEval 2019Macro F10.805HateBERT
Abuse DetectionOffensEval 2019Macro F10.803BERT
Hate Speech DetectionHatEvalMacro F10.494HateBERT
Hate Speech DetectionHatEvalMacro F10.48BERT
Hate Speech DetectionAbusEvalMacro F10.742HateBERT
Hate Speech DetectionAbusEvalMacro F10.724BERT
Hate Speech DetectionOffensEval 2019Macro F10.805HateBERT
Hate Speech DetectionOffensEval 2019Macro F10.803BERT

Related Papers

Visual-Language Model Knowledge Distillation Method for Image Quality Assessment2025-07-21Making Language Model a Hierarchical Classifier and Generator2025-07-17VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning2025-07-17The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations2025-07-17Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities2025-07-17Assay2Mol: large language model-based drug design using BioAssay context2025-07-16Describe Anything Model for Visual Question Answering on Text-rich Images2025-07-16InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing2025-07-16