Skip to content

Author

Hao Peng

We have 2 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#software testing Preprint Sep 2026

Covert Assistance: Helpful LLM Agents Evade Oversight in Multi-Agent Systems

This work emulates a software-engineering workflow in which a planner represents a company hiring an external developer and reads the nondisclosure rule as banning plaintext, not character codes or riddles, and shows that benign agents can cross the same boundaries without adversarial incentives.

Deema Alnuhait, Geng-Yu Wang, Muhammad Khalifa et al. · 0 citations

Useful Memories Become Faulty When Continuously Updated by LLMs

This work traces the regression to the consolidation step rather than the underlying experience: the same trajectories yield qualitatively different memories under different update schedules, and an episodic-only control that simply retains those trajectories remains competitive with the consolidators the authors test.

Dylan Zhang, Yan-Shan Lin, Zheng Wu et al. · 10 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.