Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Preprint Sep 2026

Forgetting Without Restarting: Execution-State Unlearning for Stateful LLM Agents

Long-running LLM agents are stateful: beyond the transcript they accrete compressed summaries, plaintext memory, pending tool plans, and, under every serving API, a KV cache. Yet today's"forget"operations delete a plaintext memory record and stop, leaving every artifact derived from the revoked information intact. We f...

Chao Yao, Yang-Bo Wei, Zhen Huang et al. · 1 citation
#artificial intelligence Preprint Open access Sep 2026

Repeat-After-Me: Black-Box Adaptive Visual Prompt Injection

Prompt injection is widely recognized as a major security threat to AI agents that interact with untrusted external data, such as websites, documents, and emails. Prior work has shown that, in the text domain, black-box prompt injection can achieve near-perfect attack success rates (ASRs). In the image domain, however,...

Sizhe Chen, Yu-Lin Tsai, Ivan Evtimov et al. · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Blockchain-Enabled Secure Logging for Fiscal Electronic Mechanisms: Evaluation of the Greek eSEND and myDATA Tax Systems

This paper analyzes the implementation of blockchain-based integrity mechanisms in Greek Fiscal Electronic Mechanisms (FEMs) and the central tax information system eSEND. The study examines the cryptographic architecture of fiscal devices, including Electronic Cash Registers, Fiscal Printers, Fiscal Signing Machines, a...

Panagiotis Mavridis, Anargyros Baklezos, Christos Nikolopoulos · 0 citations
#artificial intelligence Preprint Sep 2026

Rethinking Indirect Prompt Injection as a Test-Time Search Problem

This work identifies the attacker's adaptive search over the system attack surfaces as an important and underexplored security risk for tool-using agents and suggests that agentic security evaluations should characterize both the attacker's search procedure and compute budget.

D. M. Nguyen, Joon Sik Kim, Blazej Manczak et al. · 1 citation
#artificial intelligence Open access Sep 2026

Blockchain-Enabled Artificial Intelligence and AI Agents for Secure Data Sharing and Cybersecurity Applications

This paper presents a meta-synthesis that draws together four constituent studies covering adversarial machine learning, AI-powered anomaly detection in cloud environments, automated vulnerability patching by multi-agent large language model (LLM) pipelines, and the broader landscape of securing AI systems across their...

Harsh Verma · 0 citations
#natural language process... Preprint Sep 2026

Rent-a-RAG: Embedding-Space Watermarks for Auditing Third-Party RAG

DirBucket is the only method that consistently achieves strong target detection with no non-target activation, detecting non-compliance in every audit within 23 audited answers on the authors' primary benchmark, and the results suggest that embedding-space watermarking can make document reuse in third-party RAG statist...

Alexandr Goultiaev Tolstokorov, K. Mouratidis, Javad Dogani et al. · 0 citations
#machine learning Preprint Open access Sep 2026

Safety Training Modulates Harmful Misalignment Under On-Policy RL, But Direction Depends on Environment Design

Specification gaming under Reinforcement Learning (RL) is known to cause LLMs to develop sycophantic, manipulative, or deceptive behavior, yet the conditions under which this occurs remain unclear. We train 11 instruction-tuned LLMs (0.5B-14B) with on-policy RL across 3 environments and find that model size acts as a s...

Leon Eshuijs, Shihan Wang, Antske Fokkens · 0 citations
#machine learning Preprint Sep 2026

Flip, Don't Shuffle: Watermarking LLMs at the Speed of Inference

We introduce Stateless Bernoulli Watermarking (SBW), a new statistical watermark for Large Language Models that determines green list membership through independent per-token Bernoulli trials. Unlike KGW's vocabulary permutation or SynthID's multi-layer tournament, SBW requires only a single comparison per token agains...

Simone Ceppi, Ignacio Sanchez · 0 citations
#machine learning Preprint Sep 2026

Spruce: Scalable Private Outsourced Retrieval Using Compact Embeddings

Spruce learns compact binary codes that preserve candidates for full-precision reranking, replacing corpus-wide embedding scoring with efficient Hamming-distance computation under two-server multi-party computation under two-server multi-party computation (MPC).

Pei-Chun Hua, Yun-Ming Xiao · 4 citations
#machine learning Preprint Sep 2026

CRAW: Codec Robust Audio Watermarking

CRAW is introduced, a codec-robust audio watermarking framework that jointly improves robustness against neural re-synthesis while maintaining high perceptual quality and achieves state-of-the-art robustness against neural codecs, denoisers, and vocoders.

D. Chernin, Ethan Fetaya · 1 citation · ⚡1

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.