Human demonstrations capture diverse scenes and rich whole-body skills without requiring robot teleoperation. Prior work on egocentric transfer has emphasized scene generalization in loco-manipulation under decoupled control, leaving direct transfer of coordinated whole-body skills less explored. We present EgoHumanoid...
Jin Chen, Yi-Ming Jiang, Chong-Yang Xu et al.· 0 citations
Robotic manipulation with dexterous hands is a cornerstone of Embodied AI, yet its progress is stifled by the high cost of collecting embodiment-aware teleoperation data. While abundant egocentric videos of human hands offer a scalable alternative, the profound discrepancies in appearance, articulation, and camera view...
Zhen-Jie Yang, Xingyu Jiao, Guopeng Zhong et al.· 5 citations
This work presents JITOMA (Just-In-Time On-demand Memory Activation), a closed-loop framework that unifies task reasoning, perception, and memory into a just-in-time growth process, and introduces JITOMA-Bench, a comprehensive suite for long-horizon multi-tasking and complex multi-step reasoning.