Towards Generalized Offensive Language Identification

Computing and Communications

Electronic data

2407.18738v1
Accepted author manuscript, 482 KB, PDF document

View graph of relations

Research output: Contribution in Book/Report/Proceedings - With ISBN/ISSN › Conference contribution/Paper › peer-review

E-pub ahead of print

Alphaeus Dmonte
Tejas Arya
Tharindu Ranasinghe
Marcos Zampieri

More...

Publication date	5/09/2024
Host publication	Proceedings of the the 16th International Conference on Advances in Social Networks Analysis and Mining
Place of Publication	Cham
Publisher	Springer Nature
<mark>Original language</mark>	English
Event	The 16th International Conference on Advances in Social Networks Analysis and Mining - University of Calabria, Rende (CS), Calabria, Italy Duration: 2/09/2024 → 5/09/2024 https://asonam.cpsc.ucalgary.ca/2024/

Conference

Conference	The 16th International Conference on Advances in Social Networks Analysis and Mining
Abbreviated title	ASONAM-2024
Country/Territory	Italy
City	Calabria
Period	2/09/24 → 5/09/24
Internet address	https://asonam.cpsc.ucalgary.ca/2024/

Conference

Conference	The 16th International Conference on Advances in Social Networks Analysis and Mining
Abbreviated title	ASONAM-2024
Country/Territory	Italy
City	Calabria
Period	2/09/24 → 5/09/24
Internet address	https://asonam.cpsc.ucalgary.ca/2024/

Abstract

The prevalence of offensive content on the internet, encompassing hate speech and cyberbullying, is a pervasive issue worldwide. Consequently, it has garnered significant attention from the machine learning (ML) and natural language processing (NLP) communities. As a result, numerous systems have been developed to automatically identify potentially harmful content and mitigate its impact. These systems can follow two approaches; (1) Use publicly available models and application endpoints, including prompting large language models (LLMs) (2) Annotate datasets and train ML models on them. However, both approaches lack an understanding of how generalizable they are. Furthermore, the applicability of these systems is often questioned in off-domain and practical environments. This paper empirically evaluates the generalizability of offensive language detection models and datasets across a novel generalized benchmark. We answer three research questions on generalizability. Our findings will be useful in creating robust real-world offensive language detection systems.

Research

Electronic data