Skip to content
Open access

Interpretable AI-Enhanced Machine Learning for Intelligent Spambot and Fake Followers Identification

Aug 2026 · International Journal of Innovative Research in Engineering · 0 citations

TL;DR

The experimental results indicate that feature selection combined with CatBoost provides an effective approach for spambot prediction while explainability methods can provide additional insight into model decisions.

Abstract

Social spambots are automated accounts that imitate normal users and can be used to generate misleading, manipulative, or unwanted activity on social networking platforms. Accurate identification of such accounts is challenging because bot and human accounts may share similar profile and behavioural characteristics. This work presents a machine-learning based framework for spambot prediction using account-level and behavioural features. Genuine-user and social-spambot datasets were merged and assigned Human and Bot labels. Missing values were replaced with zero, labels were encoded numerically, features were shuffled and normalized using Min-Max scaling, and the resulting dataset was divided into 80% training and 20% testing subsets. Several machine-learning classifiers, including Random Forest, SVM, Decision Tree, XGBoost, LightGBM, Logistic Regression, Extra Trees, Naive Bayes, and AdaBoost, were evaluated using accuracy, precision, recall, and F1-score. As an extension, Chi-Square SelectKBest feature selection was used to retain 20 features before training a CatBoost classifier. The extended CatBoost model achieved 99.567% accuracy, precision, recall, and F1-score on the test set. To improve model interpretability, LIME and SHAP were incorporated to examine the contribution of input features to predictions. The experimental results indicate that feature selection combined with CatBoost provides an effective approach for spambot prediction while explainability methods can provide additional insight into model decisions.

Read PDF

Similar papers

Open access Aug 2026

INTERPRETABLE MACHINE LEARNING FOR DETECTION OF SPAMBOTS AND FAKE FOLLOWERS ON SOCIAL NETWORKS USING FEATURE BASED AND TEXT BASED

A combination of feature selection, advanced resampling, and ensemble learning algorithms and interpretable methods of robust social network automatic recognition of accounts is demonstrated to be working.

K. S. Prasad, S. V. Achutha Rao · 0 citations
Open access Sep 2026

Spam Email Classification Using TF-IDF and Classical Machine Learning on the Enron-Spam Corpus

Unsolicited email is a persistent operational and security burden and many high-performing neural approaches require computational and deployment costs that are not necessary for resource-constrained filtering systems. This paper tries to fill this gap by providing a rigorous and reproducible comparison of lightweight...

N. K. Hadi, Aymen Adil · 0 citations
Open access Sep 2026

Comparative Evaluation of Machine Learning Classifiers for SMS Spam Detection with Optimized Feature Extraction

It is demonstrated that lightweight probabilistic classifiers, when paired with effective feature engineering, can achieve near-perfect spam detection performance and offer a practical, interpretable foundation for real-world SMS filtering systems.

O. Ogunnusi · 0 citations
Open access Aug 2026

Email Security: Predictive Analysis of Spam Detection Using Machine Learning

The proposed system utilizes textual features such as word frequency, message structure, and content patterns to classify emails as spam or legitimate (ham) through supervised learning techniques, and is developed using Python and Scikit-learn.

M. K, S. Nandhini · 0 citations
Conference Aug 2026

A Context-Aware NLP Model for Automated Spam Message Identification

Spam communications remain a persistent cybersecurity threat, consuming network bandwidth, reducing productivity, and delivering malicious content such as phishing links and malware. Traditional rule-based and keyword-centric filtering methods struggle against modern spam that employs contextual manipulation and obfusc...

D. A, K. M, Tarun M et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.