When Malicious Instructions Persist: Persistent Memory Poisoning Attack on Harness-Based Agents
A targeted prompt-level defense is evaluated and finds that it can reduce memory injection in many settings, but provides limited protection once the persistent memory has been poisoned.
Shu-Huai Huang, Jing-Feng Zhang, Hong Jia
· 1 citation