#computer vision
Jan 2026
EMemBench: Interactive Benchmarking of Episodic Memory for VLM Agents
This work introduces EMemBench, a programmatic benchmark generator for evaluating long-term episodic memory of agents through interactive games, and evaluates memory agents with strong LMs/VLMs as backbones, using in-context prompting as baselines.
Xinze Li, Zi-Yue Zhu, Siyuan Liu et al.
· arXiv.org · 7 citations
· ⚡1