The effect of gender bias on hate speech detection

Autor:	Furkan Şahinuç, Eyup Halit Yilmaz, Cagri Toraman, Aykut Koç
Přispěvatelé:	Koç, Aykut
Jazyk:	angličtina
Rok vydání:	2022
Předmět:	Gender identity Signal Processing Debiased embedding Deep learning Electrical and Electronic Engineering Hate speech Language model
Zdroj:	Signal, Image and Video Processing
Popis:	Hate speech against individuals or communities with different backgrounds is a major problem in online social networks. The domain of hate speech has spread to various topics, including race, religion, and gender. Although there are many efforts for hate speech detection in different domains and languages, the effects of gender identity are not solely examined in hate speech detection. Moreover, hate speech detection is mostly studied for particular languages, specifically English, but not low-resource languages, such as Turkish. We examine gender identity-based hate speech detection for both English and Turkish tweets. We compare the performances of state-of-the-art models using 20 k tweets per language. We observe that transformer-based language models outperform bag-of-words and deep learning models, while the conventional bag-of-words model has surprising performances, possibly due to offensive or hate-related keywords. Furthermore, we analyze the effect of debiased embeddings for hate speech detection. We find that the performance can be improved by removing the gender-related bias in neural embeddings since gender-biased words can have offensive or hateful implications.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_dedup___::6b8ced97d9b02d123d3d698b67e341cb https://hdl.handle.net/11693/111608 Zobrazit plný text záznamu