Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Preprint Open access Oct 2026

In Vino Veritas and Vulnerabilities: Examining LLM Safety via Drunk Language Inducement

Humans are susceptible to undesirable behaviours and privacy leaks under the influence of alcohol. This paper investigates drunk language, i.e., text written under the influence of alcohol, as a driver for safety failures in large language models (LLMs). We investigate three mechanisms for inducing drunk language in LL...

Anudeex Shetty, Aditya Joshi, Salil S. Kanhere · 0 citations
#artificial intelligence Preprint Oct 2026

KaliBench: A Fine-Grained Benchmark for Cybersecurity Tool Use on Kali Linux with Runtime-Free Verifiable Rewards

LLMs are increasingly applied to cybersecurity workflows, where they are expected to translate analysts'intent into tool invocations. However, existing evaluations focus on knowledge-based assessments or end-to-end agentic tasks, and do not directly measure LLMs'ability to generate executable commands for real-world cy...

Peng-Fei Li, Naufal Suryanto, Si-Cheng Zhang et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

A Hybrid Approach to Malware Detection: Integrating Few-Shot Model-Agnostic Meta-Learning with Autoencoders

Ransomware has emerged as a major cybersecurity threat, with incidents increasing in frequency and impact across critical sectors. These attacks are typically launched through phishing emails, malicious downloads, or exploitation of software vulnerabilities to gain system access. Once inside, the malware encrypts files...

Emmanuela Andam, Yasir Abbas Zaidi, Abdelali Hadir et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

A Structured State Space Sequence Model for Multi-Class Classification of Malware

By 2030, Internet of Things (IoT) devices are projected to reach 40 billion, with fast-paced technological advancements in fields such as industry, healthcare, agriculture, automobiles, and building/home automation systems. This expansion has created a large attack surface for cybercrime, as the majority of these devic...

Emmanuela Andam, Rana Shaaban, Emanuel Grant et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

From Network Intrusion Detection to Blockchain-Backed Endpoint Detection and Response: Mapping the Landscape of Decentralized Detection-and-Response Architectures

While the literature on blockchain-assisted intrusion detection and prevention systems (IDS/IPS) for Internet of Things (IoT) and Industrial Internet of Things (IIoT) networks is mature, existing systematic reviews suffer from two critical limitations: they overlook the structural shift toward modern Endpoint Detection...

Yahya Shahsavari, Sara Rouhani, Kaiwen Zhang · 0 citations
#artificial intelligence Preprint Oct 2026

Walking the Embedding Space: Datastore Extraction from Multimodal RAG

Multimodal Retrieval-Augmented Generation (MRAG) has emerged as a reliable and cost-effective technique of grounding the generative capabilities of Multimodal Large Language Models (MLLMs) into relevant, up-to-date, external knowledge. Despite presenting several benefits, such as reducing hallucinatory behavior, they a...

Maria Carmen Jica, Ali Satvaty, Suzan Verberne et al. · 0 citations
#artificial intelligence Preprint Oct 2026

SoK: Decentralized Agent Economic Infrastructure

Decentralized agent economies increasingly build a single task from protocols that were designed and secured separately. This creates a simple problem: a workflow can look correct at each step and still produce the wrong outcome. For example, a correct escrow may release payment on an authorized approval that provides...

Rui Sun, Xi-Han Xiong, Qin Wang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Chaining Skills to Hijack LLM Agents

LLM agents use skills to improve performance on specialized tasks. To complete a user request, an agent may invoke several skills in sequence, allowing information produced under one skill to guide the next. Because skills may come from open-source repositories, this handoff can also carry attacker-controlled claims in...

Tian Dong, Zi-Xuan Ma, Haodong Zhao et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

False Floors: LLM Safety Routing Evaluations Break Under Distribution Shift

Safety routers send each request to one of several models and are judged against the best single model. A major routing benchmark picks that comparator on the evaluation data. In the benchmark's own setting this is harmless, but under distribution shift it is not. On HELM Safety the selection cost is 0.003-0.030 of har...

Amit Singh Bhatti, Vishal Vaddina · 0 citations
#artificial intelligence Preprint Oct 2026

PACE: Provenance-Aware Capability Enforcement for Tool-Using LLM Agents

Tool-using large language model (LLM) agents turn generated text into real side effects, so poisoned tool metadata, retrieved pages, memory, and reusable skills can steer the next call. Vetting an artifact before admission does not settle this. A safe variant and a leaking variant can produce the same admission evidenc...

Feng-Peng Li, Qi-Zhou Wang, Yu-Ke Hu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

ReCast: Contract-Preserving Protection for Fixed-Interface Multimodal Reasoning

Remote multimodal models offer strong numerical reasoning capabilities over charts and speech, but sending private inputs risks exposing sensitive content. Text-only sanitization cannot directly satisfy fixed media interfaces, while identity anonymization leaves the underlying task content exposed. We introduce ReCast,...

Bingchen Pei, Lichong Chen, Bingxi Zhao et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Jev-IDS: System One Models for Network Intrusion Detection

Machine-learning Network Intrusion Detection Systems (IDS) depend on substantial labeled datasets and task-specific training, whereas Large Language Models (LLMs) detection can analyze flow records directly but incurs higher inference cost and latency, with less constrained outputs. This paper presents JEV-IDS, an open...

Paulo Severo, Silvio E. Quincozes, Amanda Dias · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.