Skip to content
Open access

Deep Learning-Based Detection of Machine-Generated Tweets with FastText Word Embeddings

Jul 2026 · International Journal of Data Science and IoT Management System · Vol 5, pp. 550-557 · 0 citations · 2 references

TL;DR

A deep learning approach for detecting machinegenerated tweets using FastText word embeddings and a Convolutional Neural Network and demonstrates better performance than conventional machine learning methods.

Abstract

The rapid growth of social media has made it easier for information to spread quickly, but it has also increased the circulation of machine-generated and misleading content. Advanced language models can now produce tweets that closely resemble human writing, making it difficult to identify fake content through manual inspection. This project presents a deep learning approach for detecting machinegenerated tweets using FastText word embeddings and a Convolutional Neural Network (CNN). The collected tweet dataset is first preprocessed by removing unwanted characters, stop words, and noise to improve text quality. FastText is then used to convert the cleaned text into meaningful vector representations that preserve semantic information. These embeddings are provided as input to the CNN model for classification. The proposed approach effectively distinguishes human-written tweets from machine-generated ones and demonstrates better performance than conventional machine learning methods. The developed system can support social media platforms in reducing the spread of automated misinformation and improving the reliability of online communication.

Read PDF

Similar papers

Open access Aug 2026

Fake News Detection Using Machine Learning and LLM Embeddings: A Comparative Study of TF-IDF and BERT Representations on the Welfake Dataset

The proposed framework highlights the potential of integrating transformer-based language models with classical machine learning algorithms to build robust and scalable fake news detection systems.

Umme Noor Us Saqa, S. R. · 0 citations
Review Aug 2026

A Review of Deep Learning-Based Text Classification Research

The exponential growth of textual data on social media and information networks poses a significant challenge to extracting valuable information. Text classification, a core task in Natural Language Processing (NLP), is essential for organizing and categorizing such data. Deep learning has emerged as an effective a...

Ran Jin, Ya Wang, Tianzi Wu et al. · 0 citations
Open access Aug 2026

Multilingual Fake News Detection Using Machine Learning with Contextual-Based Feature Extraction

The proposed approach provides a simple and efficient solution for multilingual fake news detection in data-scarce environments with ensemble-based classifiers such as Random Forest and Gradient Boosting achieving reliable performance across both languages.

Nikita Garg, Pritam Singh Negi · 0 citations
Open access Aug 2026

Performance Evaluation of Word2Vec and FastText Embeddings in a CNN-BiLSTM Model for Sentiment Classification of the LPDP Alumni Controversy

This study aims to analyze public sentiment toward the LPDP alumni controversy on social media using a deep learning approach. The research data consist of YouTube user comments related to the LPDP issue, which were processed through text preprocessing and automatically labeled using IndoBERT into three sentiment class...

Dwi Erzalianti, Joice Junansi Tandirerung, C. Suhaeni et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.