Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Preprint Open access Sep 2026

Policy-Backed Selective Regeneration under Tainted Inter-Agent Communication

Inter-agent communication is essential to multi-agent language-model systems, yet a single message may combine task-critical information with instructions not authorized by the original request. Prompt-based defenses leave enforcement to models exposed to adversarial messages, while indiscriminate message removal disca...

Jinghan Xu, Longze Fan, Zeyuan Wang et al. · 0 citations
#artificial intelligence Review Sep 2026

Toward Responsible AI-Augmented Cyber Defense: Pattern Recognition, Defense-in-Depth, and the Case for Human-AI Collaboration

These results give the widely repeated qualitative recommendation of "balanced human-AI collaboration" a precise, testable form and suggest an interior-optimum capacity ratio as a concrete design target for security operations centers (SOCs), including those securing IT/OT-converged critical infrastructure.

Mustafa S. Aljumaily, Hayder Kareem Abed, Nawar S. Alseelawi · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Benchmarking Neural Defend ARCAS 1B: A Foundational Multimodal Deepfake Detection Model

AI-generated imagery evolves faster than benchmark-specific detector evaluations, making a single score an incomplete account of generalization. This paper evaluates Neural Defend ARCAS 1B across benchmark families without benchmark-specific parameter updates. We retain native aggregation and supplement it with record-...

Sivashankar Selvarajan, Piyush Verma, Sumit Kumar et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Not All 4-bit Quantizers Are Equal: Deployment-Time Mitigation of PII Leakage in Fine-Tuned Small Language Models

Organizations fine-tune small language models on private data and then compress them to 4 bits for resource-efficient deployment. We show that the compression method also affects privacy. What separates the methods is not the bit width but whether they tune their rounding on a small sample of text, the calibration corp...

Cristhian Kapelinski, Diego Kreutz · 0 citations
#artificial intelligence Preprint Aug 2026

Selection-Invariant Communication Compilers for Privacy-Aware Multi-Agent LLM Workflows

Structured multi-agent workflows exchange intermediate messages whose content and form can reveal private state even when the final output is safe. We identify selection-channel leakage: after authorization fixes what may be released, a private-state-aware choice among semantically valid realizations creates an additio...

Jing-Heng Xu, Long-Ze Fan, Ze-Yuan Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Evaluating Coding Agents on Kernel Exploit Generation

This work introduces KEX-bench, a benchmark for evaluating coding agents on exploit primitive generation against real operating-system kernels, and evaluates state-of-the-art coding agents paired with frontier and open-weight models under fixed tool-call budgets.

Junyoung Jang, Gwanhyun Lee, Hwiwon Lee et al. · 0 citations
#artificial intelligence Preprint Sep 2026

ZeroGate: Trust-Preserving Fast Paths for Governed AI Agent Runtimes

A conditional decision-preservation proposition: successful local admission implies that a specified synchronous policy would authorize the same action at the admission point, provided approval is sound, all policy dependencies are represented and current, observations are faithful, and consumption is atomic.

Ze-Xu Wang · 0 citations
#artificial intelligence Preprint Aug 2025

BridgeShield: Risk-Aware Graph Modeling for Cross-Chain Bridge Attack Detection

BridgeShield is presented, a graph-based framework for detecting cross-chain bridge attacks through risk-aware modeling of cross-chain execution behaviors, and consistently outperforms existing rule-based and graph-based baselines in cross-chain attack detection.

Dan-yan Lin, Shun-Feng Lu, Zi-Yan Liu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Decoding Guardrails: XAI-Guided Perturbation Analysis of Prompt Injection Detection

Large language models (LLMs) are increasingly deployed in production systems, raising concerns about their exposure to adversarial manipulation through prompt injection and jailbreak attacks. Classifier-based guardrails, such as Prompt Guard 2, are widely used as a first line of defense against such attacks, but their...

Fernando Outeda, Gustavo Betarte, J. Campo et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Reasoning Topology Matters: A Controlled Study of LLM-Based Cybersecurity Analysis

Large Language Models (LLMs) are increasingly used in cybersecurity, where accurate analysis often requires multi-step and context-dependent reasoning over complex and heterogeneous data. However, existing prompting approaches typically focus on eliciting reasoning without explicitly considering how intermediate reason...

Ji-Ling Zhou, Aisvarya Adeseye, Antti Hakkala et al. · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.