Skip to content
Open access

BotCHF: camouflage-heterogeneity-aware fusion for social bot detection

Jul 2026 · Journal of King Saud University: Computer and Information Sciences · Vol 38 · 0 citations · 54 references
Computer Science

TL;DR

A camouflage-heterogeneity-aware decision fusion framework for social bot detection (BotCHF), which adopts an alternating collaborative optimization strategy that periodically injects fine-tuned semantic features into the graph encoder, thereby progressively aligning semantic encoding with structural learning.

Abstract

Social bot detection is crucial for maintaining the security and integrity of online social networks (OSNs). Although graph-based methods have achieved state-of-the-art performance, rapid advances in large language models have made bots increasingly similar to humans in the textual modality, while some bots also adopt diverse camouflage strategies in the structural modality. As a result, social bots exhibit pronounced individual-level heterogeneity in camouflage behavior, causing the discriminative power of different modalities to vary substantially across accounts. Existing multimodal methods typically rely on unified fusion strategies, which are insufficient to handle such sample-specific variation and may lead to misclassification when one modality is heavily camouflaged. Moreover, they generally lack an effective mechanism for jointly optimizing semantic and structural representations. To address these issues, we propose a camouflage-heterogeneity-aware decision fusion framework for social bot detection (BotCHF). In the encoding stage, BotCHF adopts an alternating collaborative optimization strategy that periodically injects fine-tuned semantic features into the graph encoder, thereby progressively aligning semantic encoding with structural learning. In the fusion stage, it maintains separate text and graph branches and adaptively weights their predictions for each account, enabling decision fusion that responds to account-specific variation in modality reliability. Extensive experiments on three real-world datasets demonstrate that BotCHF consistently outperforms strong baselines. Further analysis of fusion weights reveals substantial cross-dataset differences in modality preference, highlighting the necessity of explicitly modeling camouflage heterogeneity for robust social bot detection.

Read PDF

Similar papers

Book Open access Aug 2026

OTPCL: Optimal Transport Driven Pseudo-Labeling with Contrastive Learning for Social Bot Detection

OTPCL (Optimal Transport Driven Pseudo-Labeling with Contrastive Learning), a plug-in framework for GNN-based social bot detection, first employs contrastive learning to obtain well-separated node representations and formulates pseudolabel assignment as an optimal transport problem, which simultaneously generates pseudo-labels and quantifies their reliability via transport scores.

Ruixuan Xu, Mengting Hu, Xinqi Yang et al. · 0 citations
Book Open access Aug 2026

Rethinking Generalization in Graphs: A Hierarchical Interaction Perspective for Generalist Detection

With the increasing heterogeneity of social networks and online interaction systems, generalist graph anomaly detection (GAD) has become essential for identifying abnormal and fraudulent behaviors in complex environments. However, most existing GAD approaches rely heavily on domain-specific semantic alignment, which substantially restricts their ability to learn transferable node representations and often leads to poor generalization on unseen graph domains. To address this challenge, we propose HIerarchical Interaction MOdeling for zero-shot generalist GAD (termed HIMO-GAD). HIMO-GAD enables anomaly detection across diverse graph domains without retraining or access to target-domain supervision by modeling the evolutionary trajectories of node representations across hierarchical structural depths, thereby capturing interaction patterns that exhibit strong cross-domain stability. Specifically, HIMO-GAD integrates two core components: (1) a Dynamic Interaction Modeling Module that characterizes cross-layer interaction evolution to extract transferable representations, and (2) an Anomaly-Aware Regulation Mechanism that combines gradient immunity and centralization regularization to suppress overfitting and stabilize cross-domain generalization. Extensive experiments on multiple real-world graph datasets demonstrate that HIMO-GAD consistently outperforms state-of-the-art baselines in strict zero-shot settings, achieving up to a 10% improvement in key evaluation metrics and exhibiting strong generalization across heterogeneous graph domains.

Xiangping Zheng, Xuan Feng, Bo Wu et al. · 0 citations
Open access Jul 2026

Can LLMs Keep Up? Evaluating Phishing Detection on Telegram

The findings in this study highlight the potential and current limitations of LLMs for phishing detection in dynamic instant messaging environments and emphasize the superior performance of platform-tailored models.

Md Erfan, Paula Branco, Guy-Vincent Jourdan · 0 citations
Open access Jul 2026

Fake News Identification Using Hybrid Transformer Ensemble Approach

A hybrid transformer-based ensemble model for automated fake news identification using the FakeNewsNet dataset is proposed and Experimental results show that the ensemble model achieves an accuracy of approximately 93%, outperforming the individual constituent models.

E. C. Babu, G. Sukanya · 0 citations
Open access Aug 2026

Phishing GAT: Adversarial-Hardened Phishing Email Detection via Semantic-Structural Fusion and Graph Attention Networks

PhishingGAT, a detector that fuses word-level semantic features with structural ones and is hardened against adversarial perturbation, is presented, a detector that fuses word-level semantic features with structural ones and is hardened against adversarial perturbation.

R. Kodali, Siva Rama Krishna T Dr · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.