Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Preprint Open access Sep 2026

Structurally Close, Temporally Distant: Measuring Security Exposure in Long-Horizon LLM Agents

Long-horizon LLM agents interact with untrusted content, persistent memory, external state, and sensitive tools. Existing analyses often characterize attacks by the number of execution steps between malicious input and a downstream action. We show that temporal remoteness can overstate security separation in stateful a...

Md Jafrin Hossain, Nur Al Hasan Haldar · 0 citations
#artificial intelligence Preprint Sep 2026

SAFEGuard: Detect Optimization-Based Jailbreak Attacks Through Harmful Semantic Analysis and Fluency Measurement

Despite the significant efforts devoted to aligning large language models (LLMs) with human values and ensuring safe deployment, recent work has revealed that LLMs remain vulnerable to adversarial jailbreak attacks that can bypass safety guardrails and elicit harmful responses. Many defense methods are proposed to dete...

Quoc le Viet Vo, Trung Le, D. Ranasinghe et al. · 0 citations
#artificial intelligence Preprint Sep 2026

XAI-SDN: An Explainable Entropy-Guided Machine Learning Framework for Real-Time DDoS Detection in Software Defined Networks

One of the biggest risks faced by Software Defined Networks (SDN) is the Distributed Denial of Service (DDoS) attack in which a compromised controller can make an entire network unusable. To address these challenges, we suggest an entropy-guided machine learning framework, called XAI-SDN, for real-time DDoS detection i...

Adeel Ahmad, Ali Akarma, Ahmad Ali et al. · 0 citations
#artificial intelligence Preprint Sep 2026

ResidualAuth: What Authorization State Must Language Agents Preserve under Revocable Delegation?

Tool-using language agents can delegate and revoke permissions while acting through external services. We show that two authorization histories can have identical current permissions and identical all-pairs reachability yet require opposite decisions after the same direct-edge revocation. We formalize the information n...

M. Choi, Seokho Jeong, Seunggeun Lee · 1 citation
#artificial intelligence Preprint Sep 2026

A Translational Note on AI Safety Evaluation

The AI-safety version of the threat-model coverage gap is called the AI-safety version the threat-model coverage gap, and it is found that it persists in a current open-weight model, where harms surface in non-English prompts that English benchmarks miss.

Madhava Gaikwad · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Multimodal Resource-Exhaustion Attacks on Vision-Language Models via Joint Pixel-Prompt Optimization

Resource-exhaustion attacks against autoregressive vision-language models (VLMs) typically assume unimodal threat models, treating the image branch as the primary optimization surface while holding user-visible prompts fixed. Even recent loop-centric variants remain confined to this single-channel paradigm, leaving the...

Zhaoxiong Ni, Yatie Xiao, Chi-Man Pun et al. · 0 citations

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

Neutral Prompting Attack (NPA), a highly stealthy attack paradigm in which semantically benign instructions, such as encouraging imagination and exhaustiveness, increase package hallucination propensity without containing explicit malicious intent is introduced.

Chia-Yi Hsu, Chi-Yu Li, Chun-Ying Huang et al. · 1 citation
#machine learning Open access Sep 2026

Conformal Prediction for Offensive Security

Initial findings in two key areas of offensive security: Privacy-Preserving Machine Learning, and network traffic analysis are presented, by presenting initial findings in two key areas of offensive security.

Giovanni Cherubin · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.