Skip to content
Review Open access

Small Language Models for Phishing Website Detection: A Review of Cost, Performance, and Privacy Trade-Offs

Jul 2026 · International Journal of Innovations in Science, Engineering And Management · pp. 120-123 · 0 citations · 9 references

TL;DR

It is concluded that, besides the number of parameters, the deployability of an SLM depends on its architecture and on the reliability of its output format, which does not depend only on its classification skill.

Abstract

This review investigates Goldenits et al.'s (2025) empirical comparison of fifteen open small language models (SLMs) for phishing website detection, which aimed to evaluate whether locally hosted models can achieve the same accuracy as proprietary large language models (LLMs), without the prohibitive costs or privacy concerns. The authors test each model with a stratified sample of 1,000 labelled websites from a pool of 10,395 websites and measure the accuracy, precision, recall and F1 score of the results. The best local model, llama3.3:70b, achieves an F1 score of 0.893 and recall of 0.948, which is close to, but still lower than, the F1 scores above 0.95 achieved by the largest proprietary systems (Goldenits et al. 2025). This review restates the three research questions posed in this paper, assesses the evidence provided for each of the questions and situates the evidence in the context of the existing literature on cost-aware deployment (Irugalbandara et al. 2024; Kavya & Sumathi, 2024) and LLM-based phishing detection (Koide et al. 2024). It concludes that, besides the number of parameters, the deployability of an SLM depends on its architecture and on the reliability of its output format, which does not depend only on its classification skill.

Read PDF

Similar papers

Conference Jul 2026

Performance and Explainability of Open-Weight Large Language Models for Spam Email Detection

Despite the advancements made by researchers, spam emails remain one of the biggest challenges in the field of cybersecurity. Spam emails can serve as phishing emails or carry viruses that compromise the security of an organization's system. Current detection techniques depend on supervised learning or rely on cloud-ba...

Vusal Shahbazov · 0 citations
Open access Jul 2026

Can LLMs Keep Up? Evaluating Phishing Detection on Telegram

The findings in this study highlight the potential and current limitations of LLMs for phishing detection in dynamic instant messaging environments and emphasize the superior performance of platform-tailored models.

Md Erfan, Paula Branco, Guy-Vincent Jourdan · 0 citations
Review Open access Sep 2026

Beyond Reported Accuracy: A Verifiability-Focused, Multi-Axis Review of Content and Context Based Fake News Detection

Fake news undermines information integrity, and the proliferation of large language models (LLMs) has intensified both its generation and its detection. Prior surveys catalogue methods but rarely verify the performance figures they synthesize, restrict comparison to same-benchmark settings, or assess risk of bias, so r...

Long Dinh Tuan, Le Ngoc An, Duong Dinh Thai · 0 citations
Review Open access Aug 2026

Large Language Models and Social Media Information Integrity: Opportunities, Challenges, and Research Directions

This comprehensive review examines the dual role of LLMs in both facilitating and mitigating various information integrity challenges, including misinformation, disinformation, fake news, social bots, and privacy concerns, and demonstrates critical gaps in current approaches.

Jun-Jie Xiong, Zheng-Yuan Jiang, Xiao-Ran Xu et al. · 0 citations
Open access Jul 2026

Linguistic Complexity and Readability in Corporate Privacy Policies: A Corpus-Based Comparison of Apple and Samsung

Corporate privacy policies are legally binding documents that govern how technology companies collect, process, and share user data. Despite their legal significance, these documents are frequently characterised by dense, inaccessible language that hinders consumer comprehension. This study employs a corpus-based appro...

Zainab Shakoor, Fariha Yasmeen, Syed Faizan Hussain Zaidi · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.