Towards Generalized Offensive Language Identification

Alphaeus Dmonte; Tejas Arya; Tharindu Ranasinghe; Marcos Zampieri

arXiv:2407.18738·cs.CL·July 29, 2024·1 cites

Towards Generalized Offensive Language Identification

Alphaeus Dmonte, Tejas Arya, Tharindu Ranasinghe, Marcos Zampieri

PDF

Open Access

TL;DR

This paper evaluates how well offensive language detection models and datasets perform across diverse domains, highlighting their generalizability and robustness in real-world applications.

Contribution

It introduces a novel benchmark to empirically assess the cross-domain generalizability of offensive language detection models and datasets.

Findings

01

Models vary significantly in cross-domain performance

02

Certain datasets improve generalizability when used for training

03

The benchmark reveals gaps in current offensive language detection systems

Abstract

The prevalence of offensive content on the internet, encompassing hate speech and cyberbullying, is a pervasive issue worldwide. Consequently, it has garnered significant attention from the machine learning (ML) and natural language processing (NLP) communities. As a result, numerous systems have been developed to automatically identify potentially harmful content and mitigate its impact. These systems can follow two approaches; (1) Use publicly available models and application endpoints, including prompting large language models (LLMs) (2) Annotate datasets and train ML models on them. However, both approaches lack an understanding of how generalizable they are. Furthermore, the applicability of these systems is often questioned in off-domain and practical environments. This paper empirically evaluates the generalizability of offensive language detection models and datasets across a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsHate Speech and Cyberbullying Detection · Swearing, Euphemism, Multilingualism

MethodsSoftmax · Attention Is All You Need