Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Review Sep 2026

Trust in Edge-Enabled IoT Security: Features, Challenges and Research Directions

Providing autonomous intelligence, pervasive connectivity and usability to human life and industry has led to the emergence of the Internet of Things (IoT). To support time-sensitive and resource-constrained applications, IoT systems nowadays increasingly rely on edge computing. This brings computation and decision-mak...

Esin Ece Aydin, Şerif Bahtiyar, Gürkan Gür · 0 citations
#artificial intelligence Preprint Sep 2026

Beyond Predictable Paths: Redefining AI Security Incident Reporting for Agents

Several open research questions are identified, including how to efficiently record incidents and how to determine whether vulnerabilities and incidents generalize, and privacy requirements are summarized and research directions for the secure and trustworthy deployment of AI agents are outlined.

Anastasia Pustozerova, E. Bagdasarian, Luca Beurer-Kellner et al. · 0 citations
#artificial intelligence Preprint Sep 2026

ActGov: Governing LLM Agent Actions via Policy-Constrained Validation

Large language model (LLM) agents increasingly execute long-horizon workflows through external tools, allowing untrusted outputs to influence subsequent actions and exceed user authorization. Existing defenses isolate injected content or constrain execution with predefined plans and static policies, but these approache...

Kai-Yuan Zhang, Yu-Ke Peng, Ke Jiang et al. · 1 citation · ⚡1
#artificial intelligence Preprint Open access Sep 2026

Dissecting Agentic Forensics: The Role of Triage, Prompting, and Evidence Arbitration in Open-World Fake Image Detection

Image forensics is increasingly an open-world problem: manipulations range from fully synthetic images to localized edits, splicing and swapping, while most forensic detectors remain specialized to a single manipulation family. Agentic AI has recently emerged as a promising solution. In principle, such systems can asse...

Xianlong Li (IMT School for Advanced Studies Lucca, Italy), Pietro Bongini (University of Siena et al. · 0 citations
#artificial intelligence Preprint Sep 2026

From Bits to Beliefs: Recoverable Semantic Fingerprints for Black-Box Verification of Large Language Models

Open-weight large language models (LLMs) can be copied, modified, and redeployed behind black-box APIs, making post-release ownership verification difficult. Existing black-box fingerprints often rely on secret query-key pairs that reproduce predefined responses, and can therefore be easily disrupted by fine-tuning, pr...

Jia-Xin Hong, Yu-Xin Peng, Hong-Yao Yu et al. · 0 citations
#artificial intelligence Review Sep 2026

Connecting the Dots in Agentic AI Security: A Cross-Dimensional Threat Taxonomy, Evaluation Maturity, and Open Challenges

Agentic AI extends LLM security beyond generated content to persistent state, autonomous actions, tool use, and interactions with humans and other agents. Existing threat classifications often emphasize individual dimensions, obscuring connections among entry points, affected components, and security consequences. The...

Heewon Baek, Alsharif Abuadbba, Kristen Moore et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SyzHarness: Patch-Based Kernel Bug Reproduction with LLM-Synthesized Fuzzing Harnesses

SyzHarness is a framework that combines LLM reasoning with coverage-guided fuzzing for patch-based Linux kernel vulnerability reproduction and achieves a 73% bug reproduction success rate, substantially outperforming prior directed greybox fuzzing.

Xing-Yu Li, Jue-Fei Pu, Hao-Nan Li et al. · 0 citations
#artificial intelligence Preprint Sep 2026

When the Agent Becomes the Kernel: A Systematization of Security on the Path to AI-Native Operating Systems

Large language model agents are now privileged principals that take consequential actions: editing code repositories, operating inboxes, completing purchases. Their authority is kernel-grade, but it comes without what classical systems security requires: a trusted mediator interposed on every access. Operating-system v...

Li Zhang, Yang Sun, Jie Shi · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Dual-Locking Learned AI Models: A PIN-Based Sparse QIM Watermarking and Adaptive Index Permutation Approach

We present a dual-locking method for securing trained neural networks that combines key-driven index permutation with PIN-based watermarking based on Sparse Quantization Index Modulation (QIM). Cryptographic randomness is introduced by independently applying a uniform random permutation to each row of adaptively select...

Iva Vasic, Jes\'us Mu\~noz-C\'adiz, Bata Vasic · 0 citations
#artificial intelligence Review Sep 2026

When Agentic Trust Crosses Organizational Boundaries: Structural Externalization and a Reference Model for Trust Evidence

Agentic systems increasingly invoke tools, services, data, and other agents across organizational boundaries, yet a relying party cannot assess a delegated action solely from producing-domain controls and records. This paper develops Trustworthiness as a Service (TaaS) through a synthesis of trustworthy-AI governance,...

Hua-Fu Li, Ji'an Xia · 0 citations
#artificial intelligence Preprint Sep 2026

The Price of Safety: Benign-Case Utility and Token Overhead of Memory-Poisoning Defenses in LLM Agents

Memory-poisoning defenses for LLM agents are typically evaluated by their ability to prevent attacks. However, the traffic they process is rarely adversarial. The cost of implementing a defense is paid with each interaction, while its benefits are only seen in a small percentage of cases. We developed a measurement set...

Pritom Bhowmik · 0 citations
#artificial intelligence Preprint Sep 2026

SelfOp: An Optimization Algorithm for Self-Improving Security Agents

LLM agents are increasingly used for security tasks: vulnerability discovery, exploit reproduction, and patch generation. Improving them at the model level demands expert demonstrations or computable rewards, which security tasks rarely offer: traces are costly, failures hard to diagnose, rewards sparse, and non-comput...

Saad Ullah, Yiğitcan Kaya, Christopher Kruegel et al. · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.