Terminal-agent capability depends jointly on model weights and the runtime harness that formats prompts, binds tools, and handles error recovery. Existing harness-model co-evolution approaches improve both components, yet often treat trajectories produced during harness search as an undifferentiated replay buffer. This...
Ji-Xuan Chen, Jia-Xin Zhang, Qinyuan Ye et al.· 0 citations
The Workshop on Trustworthy Natural Language Processing (TrustNLP), co-located with major ACL conferences since 2021, has grown from 8 proceedings papers to 41 over six editions, documenting a field-wide transition from post-hoc interpretability of static models to mechanistic understanding and proactive control of gen...
Rahul Gupta, Abhinav Mohanty, Anaelia Ovalle et al.· 0 citations
A comprehensive re-evaluation of two memory-based methods for self-improving agents is conducted, broadening the scope of evaluation along two axes and hypothesizing that task and environment underspecification contribute to this fragility.
Qinyuan Ye, Yu Li, Yada Pruksachatkun et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.