This work introduces a physical intelligence framework in which distributed compliant interactions jointly reveal task-relevant information and organize manipulation behavior and demonstrates this principle through blind whole-arm grasping with a hybrid rigid-soft robotic arm that is equip with IMUs embedded directly within its compliant structure.
Abstract
In animals such as elephants and octopuses, acquiring non-visual information about an object and physically engaging with it are inseparable processes mediated by rich, large-area interactions between compliant appendages and the environment. Soft robots provide a natural platform for translating this principle into engineered systems. Yet current robotic intelligence makes limited use of physical interaction, treating it primarily as a disturbance to be rejected or, at best, as a means of compensating for object misalignment. Here, we introduce a physical intelligence framework in which distributed compliant interactions jointly reveal task-relevant information and organize manipulation behavior. This results in an intrinsically partially observable problem: key task-relevant information is never measured directly, but must instead be inferred from the history of physical interactions. We propose a reinforcement-learning architecture that addresses this challenge by learning a memory-based control policy end-to-end. The key innovations making this possible are (i) a pretrained exploration policy that provides a reference for broad workspace exploration, (ii) joint optimization that integrates exploration and grasping objectives within a single recurrent policy, and (iii) a two-stage sim-to-real adaptation including observation mapping and policy fine-tuning. We demonstrate this principle through blind whole-arm grasping with a hybrid rigid-soft robotic arm that we equip with IMUs embedded directly within its compliant structure, providing its only source of proprioceptive sensing. The learned policy successfully identifies and grasps various objects by autonomously coordinating workspace exploration, object encounter and localization, inference of grasp-relevant properties, and stable whole-arm wrapping.
This work presents RobotMover, a complete learning-based system for large-object manipulation that leverages human–object interaction demonstrations to train robot control policies and achieves strong performance in terms of capability, robustness, and controllability, outperforming both learned and teleoperation basel...
Tian-Yu Li, Joanne Truong, Tsung-Yen Yang et al.· IEEE Transactions on robotic...· 2 citations
GeniWorld is presented, an interactive world model for robots that generalizes robustly across unseen scenarios by explicitly decoupling embodiment kinematics from environmental dynamics, and generates diverse manipulation trajectories within the world model, improving downstream policy performance and robustness in co...
This work proposes an LfD method that explicitly uses what physical interactions take place where and when, and discusses how robustness, generalization, and adaptivity can be explicitly implemented, which is generally lacking in the LfD literature.
A. H. G. Overbeek, H. van der Kooij, M. Vlutters· 0 citations
Zeva is presented, the first framework that enables in-context learning from a robot's own physical interaction experience while keeping the policy model frozen, and achieves the best performance among the compared frontier VLAs and WAMs and enables self-evolution during deployment without gradient updates.
Fu Chen, Xin Ding, Bing-Jia Huang et al.· 3 citations
This study investigates the combination of a state-of-the-art reinforcement learning (RL) algorithm with human demonstrations to learn how to open a door with minimal task-specific engineering on an articulated soft robot arm and shows that combining LfD with RL results in both better performance and more robust behavi...
Laurenz Elstner, Erik Kyrkjebø, M. Stoelen· Frontiers in Robotics and AI· 0 citations
Experimental results demonstrate that a single policy enables a Unitree G1 humanoid to complete the full task using only onboard depth sensing and proprioception, while generalizing robustly across task variations and transferring effectively from simulation to reality.
Ze-Jie Tian, Rui-Bing Hou, Bin Ma et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.