Computer Science > Computation and Language

arXiv:2407.18738 (cs)

[Submitted on 26 Jul 2024]

Title:Towards Generalized Offensive Language Identification

Authors:Alphaeus Dmonte, Tejas Arya, Tharindu Ranasinghe, Marcos Zampieri

Abstract:The prevalence of offensive content on the internet, encompassing hate speech and cyberbullying, is a pervasive issue worldwide. Consequently, it has garnered significant attention from the machine learning (ML) and natural language processing (NLP) communities. As a result, numerous systems have been developed to automatically identify potentially harmful content and mitigate its impact. These systems can follow two approaches; (1) Use publicly available models and application endpoints, including prompting large language models (LLMs) (2) Annotate datasets and train ML models on them. However, both approaches lack an understanding of how generalizable they are. Furthermore, the applicability of these systems is often questioned in off-domain and practical environments. This paper empirically evaluates the generalizability of offensive language detection models and datasets across a novel generalized benchmark. We answer three research questions on generalizability. Our findings will be useful in creating robust real-world offensive language detection systems.

Comments:	Accepted to ASONAM 2024
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2407.18738 [cs.CL]
	(or arXiv:2407.18738v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2407.18738

Submission history

From: Alphaeus Dmonte [view email]
[v1] Fri, 26 Jul 2024 13:50:22 UTC (383 KB)

Computer Science > Computation and Language

Title:Towards Generalized Offensive Language Identification

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Towards Generalized Offensive Language Identification

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators