Skip to content
Review Open access

Deep Learning Approaches for Hate Speech Detection in Resource-Constrained Social Media Texts: A Comparative Evaluation

Jul 2026 · Artificial Intelligence and Applications · 0 citations · 37 references

TL;DR

This paper further provides empirical evidence regarding the ability of the BERT model in detecting hate speech in a rough text environment and comparing directly with other deep learning techniques, suggesting that BERT has great potential as a multilingual content moderation tool to be applied in informal and unstructured digital environments.

Abstract

Hate speech that occurred in social media platforms presents critical safety challenges, especially in language environments with minimal source where text is often written in an informal form, abbreviated, and context dependent. This paper further provides empirical evidence regarding the ability of the BERT (Bidirectional Encoder Representations from Transformers) model in detecting hate speech in a rough text environment and comparing directly with other deep learning techniques.Three Bidirectional Long Short–Term Memory (BiLSTM) variants using Word2Vec, FastText, and TF-IDF representations were compared with BERT, which employs contextual language representations. The data is separated by using a train-test ratio of 80:20, and the performance is evaluated according to their accuracy, precision, recall, F1-score, and area under the curve (AUC). The result shows that BERT outperforms all variations of BiLSTM by achieving an accuracy of 87.18%, an F1-score of 87.08%, and an AUC matrix of 0.9244. According to the efficiency matters, BiLSTM and Term Frequency–Inverse Document Frequency (TF-IDF) are the most ineffective for their uneven classification distribution, while BiLSTM with Word2Vec and FastText shows moderate effectivity. This finding conclusively demonstrates the benefit of transformer-based models to capture the subtleties of language with noisy textual data, which is commonly found in low-resource settings. To conclude the above discussion, this finding suggests that BERT has great potential as a multilingual content moderation tool to be applied in informal and unstructured digital environments.   Received: 25 August 2025 | Revised: 29 January 2026 | Accepted: 23 June 2026   Conflicts of Interest The authors declare that they have no conflicts of interest to this work.   Data Availability Statement The data that support the findings of this study are openly available in GitHub at https://github.com/okkyibrohim/id-multi-label-hate-speech-and-abusive-language-detection, reference number [29].   Author Contribution Statement Yosia Immanuel Bastian: Conceptualization, Methodology, Formal analysis, Resources, Data curation. Aditiya Hermawan: Software, Supervision, Project administration, Writing – original draft, Writing – review & editing, Visualization. Ardiane Rossi Kurniawan Maranto: Validation, Investigation. Benny Daniawan: Validation, Investigation. Junaedi Junaedi: Validation, Investigation.

Read PDF

Similar papers

Review Open access Jul 2026

Detecting Hate Speech in Hindi Digital Discourse Using Transformer–Long Short-Term Memory Models

Evaluating advanced language models for analysing hate speech in Hindi social media content illustrates that language-specific computational tools can be used for both platform governance and communication research, provided that the cultural context is considered.

Rachna Narula, Vedika Gupta, Jawad Khan et al. · 0 citations
Open access Aug 2026

A Context-Aware and Target-Adaptive Multilingual Framework for Hate Speech Detection in Code-Switched Social Media Text

A Context-Aware and Target-Adaptive Multilingual Hate Speech Detection model that combines multilingual transformer-based embeddings with a context-aware attention mechanism to capture semantic dependencies in text and reduces false positives is introduced.

K. Shruthi, K. Shivanna · 0 citations
#artificial intelligence Preprint Aug 2026

Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu

It is challenging to detect hate speech in Low Resource Languages (LRLs) because of the absence of annotated data, the informality of its language structure, and the lack of standardized grammar. A good example of such a challenge is Roman Urdu which is broadly used by South Asians on social media and has a high variat...

Toneema Zubair, Muhammad Asif, F. Kamiran et al. · 0 citations
Open access 2026

From Binary to Multi-Class: LLM-Judged Synthetic Annotation Applied to Hate Speech Detection

The proposed strategy offers a versatile solution for nuanced classification tasks beyond hate speech, providing a valuable technique for detailed categorisation in various domains.

Antonio Moreno-Cediel, Antonio Garcia-Cabot, Eva García-López · 0 citations
Open access Aug 2026

Indonesian Hate Speech Detection Across Diverse Domains Using Parameter-Efficient Fine-Tuning with IndoBERT and LoRA

The findings indicate that IndoBERT+LoRA provides a promising and resource-efficient approach for multi-domain Indonesian hate and abusive speech classification, while stricter leave-one-domain-out evaluation remains an important direction for future work.

Fergie Joanda Kaunang, Bhustomy Hakim, A. P. Thenata · 0 citations
Open access Jul 2026

Low-Resource Hate Speech Detection in English-Swahili Code-Switched Text Using Fine-Tuning of Pre-trained Language Models

This study explores a low-resource approach to detecting hate speech in English and Swahili code-switched text by fine-tuning pre-trained language models, and shows that fine-tuning modern language models can offer a practical and scalable solution for hate speech detection in multilingual environments.

Kipkebut Andrew, Jepkemei Betty · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.