Reinforcement learning for embodied control remains constrained by the difficulty of reward specification. Although recent large language model (LLM)-based methods can synthesize reward functions from natural-language descriptions, they often fail to capture subtle behavioral properties that humans care about, such as...
Eren Sadikoglu, Aditya Taparia, Xin-Yuan Liu et al.· 0 citations
Reliable video world models could provide scalable predictive environments for robot learning, planning, and evaluation. However, generated robot videos can violate physical principles and complete tasks through physically implausible behavior, limiting their reliability for robot learning and planning. Current video-g...
Isaiah Milkey, Som Sagar, Aditya Taparia et al.· 0 citations
ARC (Agentic Resource&Configuration learner), a lightweight hierarchical policy that dynamically selects query-specific agent configurations, consistently improves over budget-matched tool-augmented LLMs, demonstrating that learning per-query agent configurations is a powerful alternative to"one size fits all"designs.
Aditya Taparia, Som Sagar, Ransalu Senanayake· arXiv.org· 2 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.