DeepHateExplainer: Explainable Hate Speech Detection in Under-resourced Bengali Language (2012.14353v4)

Published 28 Dec 2020 in cs.CL and cs.LG

Abstract: The exponential growths of social media and micro-blogging sites not only provide platforms for empowering freedom of expressions and individual voices, but also enables people to express anti-social behaviour like online harassment, cyberbullying, and hate speech. Numerous works have been proposed to utilize textual data for social and anti-social behaviour analysis, by predicting the contexts mostly for highly-resourced languages like English. However, some languages are under-resourced, e.g., South Asian languages like Bengali, that lack computational resources for accurate NLP. In this paper, we propose an explainable approach for hate speech detection from the under-resourced Bengali language, which we called DeepHateExplainer. Bengali texts are first comprehensively preprocessed, before classifying them into political, personal, geopolitical, and religious hates using a neural ensemble method of transformer-based neural architectures (i.e., monolingual Bangla BERT-base, multilingual BERT-cased/uncased, and XLM-RoBERTa). Important(most and least) terms are then identified using sensitivity analysis and layer-wise relevance propagation(LRP), before providing human-interpretable explanations. Finally, we compute comprehensiveness and sufficiency scores to measure the quality of explanations w.r.t faithfulness. Evaluations against machine learning~(linear and tree-based models) and neural networks (i.e., CNN, Bi-LSTM, and Conv-LSTM with word embeddings) baselines yield F1-scores of 78%, 91%, 89%, and 84%, for political, personal, geopolitical, and religious hates, respectively, outperforming both ML and DNN baselines.

PDF Abstract

Summarize Bookmark Chat (Pro)

Authors (9)

Md. Rezaul Karim (16 papers)
Sumon Kanti Dey (4 papers)
Tanhim Islam (4 papers)
Sagor Sarker (4 papers)
Mehadi Hasan Menon (4 papers)
Kabir Hossain (2 papers)
Bharathi Raja Chakravarthi (26 papers)
Md. Azam Hossain (5 papers)
Stefan Decker (24 papers)

Citations (72)

View on Semantic Scholar

DeepHateExplainer: Explainable Hate Speech Detection in Under-resourced Bengali Language (2012.14353v4)

Related Papers