MemPoison: Bypassing Selective Memory Mechanisms to Plant Backdoors in LLM Agents
MemPoison is proposed, a novel memory poisoning attack that bypasses selective memory mechanisms in LLM agents, where an attacker can inject triggerable backdoors into the agent's long-term memory through dialogue interactions, thereby misleading its subsequent responses.