Jul 2026
Bad Memory: Evaluating Prompt Injection Risks from Memory in Agentic Systems
This work evaluates two agentic systems, Anthropic Claude Code and OpenAI Codex, across four models and shows that persistent memory changes the threat model for prompt injection and motivate defenses that protect memory updates without removing useful agent adaptation.
Soham U. Gadgil, David Alexander, S. Sunku et al.
· arXiv.org · 2 citations
· ⚡1