Skip to content
Conference

Hybrid Machine Learning Framework for Phishing Website Detection

Aug 2026 · International Conference on Information Security and Cryptology · pp. 1977-1982 · 0 citations · 18 references

Abstract

Phishing websites continue to be a major cybersecurity threat because attackers create deceptive web pages that imitate trusted banking, e-commerce, social media, and government platforms to collect sensitive user information. Traditional blacklist-based and rule-based detection methods are limited because they mainly identify known threats and require frequent manual updates. This paper proposes a hybrid machine learning framework for phishing website detection using character-level Term Frequency-Inverse Document Frequency (TF-IDF), Truncated Singular Value Decomposition (SVD), K-Means clustering, Random Forest, Extreme Gradient Boosting (XGBoost), and an Artificial Neural Network (ANN). In the proposed approach, TF-IDF and SVD generate compact URL representations, while K-Means cluster labels and centroid-distance features enhance the feature space before supervised classification. The prediction probabilities generated by Random Forest, XGBoost, and ANN are combined through a weighted soft-voting ensemble. Experimental results on a balanced URL dataset demonstrate that the hybrid model achieves strong classification performance with an accuracy of 93.76%, precision of 90.70%, recall of 80.55%, F1-score of 85.32%, and ROC-AUC of 97.65%. The proposed framework provides a practical and scalable approach for URL-based phishing detection and real-time web security applications.

View source

Similar papers

Open access Aug 2026

A Framework for Optimising Phishing Websites Detection via Ensemble Learning and Synthetic Minority Oversampling Techniques

Phishing websites remain one of the most prevalent forms of cyber-attacks. They often exploit users, through deceptive web interfaces, to steal sensitive information such as login credentials, financial data, and personal records. The increasing sophistication of phishing strategies has reduced the effectiveness of tra...

D. Asuquo, Fredrick Umoh, K. Attai et al. · 0 citations

Explainable Phishing Website Detection Using Comparative Machine Learning and SHAP

An integrated comparative evaluation that combines six-algorithm benchmarking, leakage-free hyperparameter optimization, and SHAP-based interpretation on a public phishing dataset, offering practical guidance for security analysts is offered.

Juni Ismail, Raja Anan Nasution, Muhammad Nasri Gea · 0 citations
Open access Sep 2026

A High Accuracy Classifier-Based Approach for Detecting Phishing on Social Media

The increasing prevalence of phishing attacks on social media platforms poses a serious challenge to online security and user trust. Cybercriminals exploit the openness and anonymity of these platforms to deceive users into revealing sensitive information or downloading malicious content. This study presents a high-...

Olayinka Oluwaseun Olaiya · 0 citations
Open access Aug 2026

Real-Time Phishing URL Detection Using a Hybrid Stacking Ensemble: Gradio and Browser Extension Deployment

Phishing attacks remain a prevalent and rapidly evolving cybersecurity threat, leveraging deceptive Uniform Resource Locators (URLs) and fraudulent websites to steal sensitive user data, financial credentials, and personal information. Traditional detection mechanisms, such as blacklist-based and heuristic approaches,...

E. Kavya, A. S. Chakravarthy · 0 citations
Conference Aug 2026

An Intelligent TabNet-based Framework for Phishing Website Detection using URL and Webpage Features

One of the most common cyber security threats today is through phishing websites, which appear legitimate to trick users into providing their details, including login information, banking details, and personal information. The sophistication of phishing attacks has grown and so has the application of machine learning a...

T. Srinivas, T. Shanthi, Chennaiah Kate et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.