Embodied agents now take on ever longer tasks. For long tasks, knowing only whether a task finally succeeds or fails says little; the steps along the way matter. Progress Reward Models (PRMs) score how far a task has come at every step, and serve as dense rewards, verifiers and monitors. Yet in long tasks the current f...
Jian-Shu Zhang, Ke-Liang Wu, Cheng-Xuan Qian et al.· 0 citations
A unified view of progress reward modeling for robotic learning is provided in three connected steps that connect what a progress model is, how it is built, and how its quality is validated.
Jian-Shu Zhang, Ke-Liang Wu, Haoran Lu et al.· arXiv.org· 5 citations
By foregrounding decision intelligence in complex systems, Enactive AI expands the frontier of AI from model capability to system-aware action, opening new possibilities for scalable, governable, and socially valuable AI deployment.
Zuo-Jun Max Shen, Yuan Qu, Pu-Jun Zhang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.