Jul 2026
AgentS4D: Benchmarking Runtime Risks across the Execution Lifecycle of LLM-Based Workspace Agents
This work introduces AgentS4D, a sandboxed benchmark for lifecycle-wide runtime safety evaluation and evaluates all 20 combinations of four harnesses and five LLM backends, finding that the observed safety of an agent system varies with both its harness-LLM pairing and how risk is introduced.
Jiajun Zhou, Zhaoxuan Ke, Jihang Ye et al.
· arXiv.org · 1 citation