Skip to content

A Minimalist Retargeting-Guided Reinforcement Learning Recipe for Dexterous Manipulation

Jul 2026 · arXiv.org · Vol abs/2607.11874 · 1 citation · 52 references
Computer Science

TL;DR

Through systematic hardware experiments, this work identifies and analyze the key factors that govern sim-to-real transfer in dexterous manipulation, offering practical guidance for retargeting-based learning in contact-rich settings.

Abstract

Recent work in humanoid whole-body control has found success with a simple recipe: retarget human motion to robot kinematic references, then train policies via reinforcement learning (RL) to track them. But how does this recipe transfer to dexterous manipulation? The answer is not obvious, as manipulation involves complex, contact-rich dynamics and requires delicate regulation of contact modes and forces. We present REGRIND, a minimalist retargeting-guided RL pipeline that learns dexterous manipulation policies from a single human demonstration. REGRIND retargets human hand-object motion to a robot reference that preserves hand-object spatial and contact relationships, trains a residual RL policy in simulation to track object-centric keypoints along that reference, and transfers the resulting policy zero-shot to hardware with careful system identification. The resulting policies produce fluid, human-like behavior on two different multi-fingered hands across contact-rich tool-use tasks, including operating a pair of scissors and turning a screwdriver. Through systematic hardware experiments, we identify and analyze the key factors that govern sim-to-real transfer in dexterous manipulation, offering practical guidance for retargeting-based learning in contact-rich settings. Videos and code are available at https://yunhaifeng.com/REGRIND.

View source

Similar papers

Preprint Aug 2026

ReForce: Learning Force-aware Retargeting for Dexterous Manipulation

ReForce is presented, a Force-aware Retargeting method that turns human motion and forces into robot actions that reproduce the intended contact and supports both online force-aware teleoperation and offline data translation.

Yuhang Wu, Ling-Qi Zeng, Chang-Wei Jing et al. · 3 citations
Preprint Aug 2026

DexMani: Human-Derived Manipulability Guidance for Dexterous Rotation

This work introduces DexMani, a framework that transfers human demonstrations as contact-conditioned manipulability evolution and uses it to guide downstream reinforcement learning, enabling rotation skills to be acquired across robot embodiments with distinct kinematics and active-contact configurations.

Xiaoyang Chen, Shengcheng Luo, Haoran Guo et al. · 0 citations

Lightweight Adaptation of Pretrained Robot Manipulation Systems: Two Approaches

Two systematic attempts to improve large pretrained models with minimal or zero modification to their weights via reinforcement learning on a frozen OpenVLA-7B using binary task-success rewards on LIBERO-Goal reveal a common ceiling.

Adam Lalani, Chen Sun, Hui Wang · 0 citations
Preprint Aug 2026

NestDex: Nested Policy Learning with Copilot Assisted Teleoperation for Dexterous Manipulation

Dexterous manipulation promises substantially richer robot interaction with the physical world, but learning these behaviours remains constrained by the difficulty of collecting consistent, complete-task demonstrations. Unlike parallel-jaw manipulation, dexterous tasks require the operator to coordinate arm motion with...

James Zhao, Jin-He Tang, Ming-Yuan Ba et al. · 1 citation
Preprint Sep 2026

One Demonstration, Many Objects: Generalizing Manipulation via Local Contact Geometry

DemoMimic (Dexterous Motion Mimic), a policy that manipulates objects by focusing on their geometry local to the contact points to improve sim-to-real consistency, yielding a single real-world policy that transfers across objects of varying shape, scale, mass, and friction wherever local contact structure is preserved.

Satvik Sharma, Samrat Sahoo, Huang Huang et al. · 3 citations · ⚡1
#artificial intelligence Preprint Aug 2026

ADEPT: Accelerating Dexterity via Pre-Training and Post-Training using Reinforcement Learning

We introduce Accelerating Dexterity via Pre-Training (ADEPT), a large-scale reinforcement learning (RL) framework for learning sim-to-real transferable dexterity across high degree-of-freedom (DoF) robot embodiments that can solve long-horizon tasks directly from raw visuo-tactile perception. ADEPT pretrains a dexterou...

Jayjun Lee, Jessica Yin, Asif Rana et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.