Skip to content

Category

cybersecurity

1,032 papers

#artificial intelligence Preprint Open access Oct 2026

MARS: Malware Analysis with Rule-Based Scoring of LLM Claims

Large language models can triage malware through direct verdicts or behavioral claims scored by an external policy. We present MARS, a malware triage framework, and compare direct classification with single-pass claim scoring using the same evidence collector and identical static evidence bundles for each model. The ev...

Hyeongjun Choi · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Secure-CUA: Controlling Untrusted Influence in Computer-Use Agents

Computer-use agents (CUAs) perform tasks across applications (such as desktops, mobile apps, and web browsers) by observing graphical interfaces and issuing commands such as clicks and keystrokes. These interfaces combine trusted controls and content with untrusted content needed for legitimate tasks. An adversary cont...

Sarthak Choudhary, Mihai Christodorescu, Ashish Hooda et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Package Hallucination Attacks on Coding Agents through Prompt Injection in Rule Files

Modern agentic coding frameworks increasingly rely on community-shared rule files (e.g., AGENTS.md or .cursorrules) to guide autonomous code generation, yet the security risks of this pipeline remain underexplored. To bridge this gap, we introduce the package hallucination attack, where an attacker injects malicious pr...

Yupu Wang, Zhengyuan Jiang, Reachal Wang et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Adversarial Images Hijack Web Agents from Visual Grounding to Browser Execution

Modern web agents built on large vision-language models process webpages, select relevant UI elements, and translate model outputs into browser actions. Existing visual red-teaming approaches use adversarial visual content to manipulate this process. However, they primarily target model inference and do not explicitly...

Wanjing Han, Levi Taiji Li, Mu Zhang et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

SwarmReconGuard: Black-Box Detection of Distributed Collective Reconnaissance by Individually Benign-Looking Agent Populations

Autonomous and agentic clients can distribute reconnaissance across many identities so that each request remains valid, low-rate, and benign-looking while the population collectively acquires broad system knowledge. We formalize this threat as Distributed Collective Reconnaissance (DCR) and present SwarmReconGuard, a r...

Vahid Tavakkoli, Kabeh Mohsenzadegan, Kyandoghere Kyamakya · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Quad-State Safety Evaluation of Open-Weight Large Language Models on Non-Canonical Inputs

Standard safety evaluations of large language models assess harmful requests written in canonical plain text, while models in real-world deployment routinely receive inputs containing emojis, altered spellings, encoded strings, and character-level variations. This work introduces the Adversarial Surface-Form Robustness...

Pavan Maddula · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Visual Memory Attacks Can Persist Through The KV Cache

Modern language model systems operate autonomously over increasingly long contexts containing untrusted text and images. Can an adversarial input continue to steer a model even after that input is removed from its context? We show that attacks can be trained to persist through the key/value (KV) cache of subsequent tok...

David Dobre, Leo Schwinn, Gauthier Gidel et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Multi-Aspect Runtime Verification for Simulation-Based V&V of LLM-Enabled Autonomous Agents

LLM-based agents are entering decision-support roles in defence staff work, where the obligations they must respect are already written down and binding, and where retraining is not available as a control because models arrive as procured components. What can be placed under engineering control is the interface between...

Nikolaos Kekatos, Dimitrios Nikou, Anastasios Temperekidis et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Contextualization of Third-Party Cloud Security Findings

Finding severity is the main driver of how security teams prioritize remediation. For third-party cloud security findings, that severity is static: the rule that raised the finding assigns it before the rule meets any environment, so it reflects the risk of the condition in general rather than the risk the finding pose...

Leon Goldberg, Gal Engelberg · 0 citations
#artificial intelligence Preprint Open access Oct 2026

CredLeakBench: Evaluating Credential Leakage and Recovery in LLM Agents

Language model agents are increasingly deployed to automate everyday digital chores from managing emails and social media to handling banking and bills allowing users to step away from supervision. However, this capability also exposes sensitive information to phishing. Safe execution requires distinguishing malicious...

Rafid Ahmed, Joseph Fioresi, Mubarak Shah et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Sigma-Hunter: A Domain-Specific Language Model for Threat Hunting and Detection Engineering

Detection engineers must translate threat reports, forensic observations, and hunt hypotheses into precise, testable rules. General-purpose large language models (LLMs) can draft such rules, but often produce invalid YAML, incorrect log sources, unsupported fields, or overly broad detection logic. This paper presents \...

Kemal Davaslioglu, Sastry Kompella · 0 citations
#artificial intelligence Preprint Oct 2026

How Fragile Is On-Device Language Model Safety? Localizing Safety-Critical Parameters for Sparse Fault Analysis

As small language models (SLMs) are increasingly deployed on resource-constrained and on-device platforms, including as components of agentic systems, the integrity of locally stored model parameters becomes an important safety concern. We investigate whether safety-sensitive behavior in LLaMA-2-7B-Chat is concentrated...

M. Karamat, Christian García · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.