Jul 2026· Journal of Information System Exploration and Research· Vol 4, pp. 281-290· 0 citations· 28 references
TL;DR
This study contributes to Indonesian clickbait detection research by demonstrating that ensemble aggregation of diverse transformer architectures yields more reliable performance than reliance on any single model.
Abstract
Clickbait is an increasingly prevalent phenomenon in Indonesian online news media, where headlines are crafted to attract clicks without accurately reflecting article content. This study proposes and compares eight classification models: three shallow learning algorithms — Naive Bayes, Support Vector Machine (SVM), and Logistic Regression — and five transformer-based deep learning models: IndoBERT-p1, IndoBERT-p2, XLM-RoBERTa, mBERT, and DistilBERT. The dataset used is CLICK-ID, consisting of 15,000 labeled headlines from 12 Indonesian news portals, expanded to 25,138 samples via semi-supervised pseudo labeling with a confidence threshold of 0.85. All deep learning models were trained with Focal Loss (α=0.25, γ=2.0) to address class imbalance and Automatic Mixed Precision (AMP) for GPU efficiency. Results show that IndoBERT-p1, IndoBERT-p2, XLM-RoBERTa, mBERT, and DistilBERT achieve comparable performance, with macro F1-scores ranging from 86.57% to 88.54%. Among shallow learning models, SVM performs best with 83.51% F1-score. An average ensemble of all five transformer models achieves the best overall performance at 90.14% accuracy and 89.00% F1-score, outperforming every individual model. This study contributes to Indonesian clickbait detection research by demonstrating that ensemble aggregation of diverse transformer architectures yields more reliable performance than reliance on any single model.
Misinformation propagation across online platforms continues to pose serious risks to informed public discourse and media credibility. To address this, we design and evaluate a fully integrated fake news detection pipeline built upon the FakeNewsNet benchmark, drawing from both PolitiFact and Buz-zFeed corpora. This wo...
G. Sai, R. B. Kumar, Yalavarthi Sai Eswari· International Conference on...· 0 citations
Stance detection has become a fundamental task in natural language processing (NLP), yet it remains under-explored for low-resource languages such as Sorani Kurdish. Building on the previously released Bochun dataset, the present work focuses exclusively on the comprehensive evaluation of stance detection models and pr...
P. S. Rostam, Rebwar M. Nabi· ARO. The Scientific Journal...· 0 citations
Across Indonesian online platforms, fabricated news spreads faster than fact-checkers can confirm. Because much of the literature relies on resource-intensive deep models, one applied question stays unsettled: which lighter, more transparent classifier best detects Indonesian hoaxes? We assessed four algorithms, Random...
Dedi Irawan, Sudarmaji· Journal of Information Syste...· 0 citations
Non-performing loan (NPL) detection is inherently a class-imbalance problem because defaulting borrowers represent a persistent minority. Standard gradient boosting often favors the majority class. This paper proposes AugLog-LightGBM, an extension of LightGBM that improves initialization through Log-Based Feature Augme...
H. Azizah, E. Sumarminingsih, A. Fernandes· International Journal of Adv...· 0 citations
Nowadays, Natural Language Processing, or NLP, is a key component of many programs that analyze and comprehend human language. The sentiment analysis of mobile product reviews collected from the Kaggle repository—more especially, the 20,710-review Amazon Mobile evaluations dataset—is the main emphasis of this research....
Dhananchezhiyan R, M. Rameshkumar· International journal of com...· 0 citations
The authors suggest a computationally efficient multimodal deep learning framework using Bidirectional Encoder Representations of Transformers (BERT) to extract textual features and convolutional neural networks to learn visual representations that offers a computational scaling alternative to attention-based models, w...
P. Jadhav, R. K. Shukla· International Research Journ...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.