This work introduces a physical intelligence framework in which distributed compliant interactions jointly reveal task-relevant information and organize manipulation behavior and demonstrates this principle through blind whole-arm grasping with a hybrid rigid-soft robotic arm that is equip with IMUs embedded directly within its compliant structure.
Abstract
In animals such as elephants and octopuses, acquiring non-visual information about an object and physically engaging with it are inseparable processes mediated by rich, large-area interactions between compliant appendages and the environment. Soft robots provide a natural platform for translating this principle into engineered systems. Yet current robotic intelligence makes limited use of physical interaction, treating it primarily as a disturbance to be rejected or, at best, as a means of compensating for object misalignment. Here, we introduce a physical intelligence framework in which distributed compliant interactions jointly reveal task-relevant information and organize manipulation behavior. This results in an intrinsically partially observable problem: key task-relevant information is never measured directly, but must instead be inferred from the history of physical interactions. We propose a reinforcement-learning architecture that addresses this challenge by learning a memory-based control policy end-to-end. The key innovations making this possible are (i) a pretrained exploration policy that provides a reference for broad workspace exploration, (ii) joint optimization that integrates exploration and grasping objectives within a single recurrent policy, and (iii) a two-stage sim-to-real adaptation including observation mapping and policy fine-tuning. We demonstrate this principle through blind whole-arm grasping with a hybrid rigid-soft robotic arm that we equip with IMUs embedded directly within its compliant structure, providing its only source of proprioceptive sensing. The learned policy successfully identifies and grasps various objects by autonomously coordinating workspace exploration, object encounter and localization, inference of grasp-relevant properties, and stable whole-arm wrapping.
Moving large objects, such as furniture or appliances, is a critical capability for robots operating in human environments. This task presents unique challenges, including whole-body coordination to avoid collisions and managing the underactuated dynamics of bulky, heavy objects. In this work, we present RobotMover, a...
Tian-Yu Li, Joanne Truong, Tsung-Yen Yang et al.· IEEE Transactions on robotic...· 2 citations
GeniWorld is presented, an interactive world model for robots that generalizes robustly across unseen scenarios by explicitly decoupling embodiment kinematics from environmental dynamics, and generates diverse manipulation trajectories within the world model, improving downstream policy performance and robustness in co...
This work proposes an LfD method that explicitly uses what physical interactions take place where and when, and discusses how robustness, generalization, and adaptivity can be explicitly implemented, which is generally lacking in the LfD literature.
A. H. G. Overbeek, H. van der Kooij, M. Vlutters· 0 citations
Zeva is presented, the first framework that enables in-context learning from a robot's own physical interaction experience while keeping the policy model frozen, and achieves the best performance among the compared frontier VLAs and WAMs and enables self-evolution during deployment without gradient updates.
Fu Chen, Xin Ding, Bing-Jia Huang et al.· 3 citations
This study investigates the combination of a state-of-the-art reinforcement learning (RL) algorithm with human demonstrations to learn how to open a door with minimal task-specific engineering on an articulated soft robot arm and shows that combining LfD with RL results in both better performance and more robust behavi...
Laurenz Elstner, Erik Kyrkjebø, M. Stoelen· Frontiers in Robotics and AI· 0 citations
Humanoid loco-manipulation requires adaptive whole-body coordination to seamlessly integrate locomotion and physical interaction. Despite recent advances, learning autonomous loco-manipulation remains challenging due to the scarcity of diverse, physically executable robot-object interaction data and the difficulty of l...
Ze-Jie Tian, Rui-Bing Hou, Bin Ma et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.