Skip to content
Conference Open access

Enhancing Sim2Real Transfer for Torque-Controlled Robots through Real2Sim Dynamics Estimation and Reinforcement Learning

Jul 2026 · 2026 IEEE/ASME International Conference on Advanced Intelligent Mechatronics (AIM) · pp. 1-6 · 0 citations · 23 references
Computer Science Engineering

TL;DR

This work proposes a Real2Sim2Real pipeline that improves Sim2Real transfer for torque-controlled robotic arms by combining trajectory matching, parameter optimization via genetic algorithms, and domain randomization, and demonstrates a significant improvement in tracking accuracy and policy robustness after parameter tuning.

Abstract

Transferring reinforcement learning policies from simulation to Real-World robots remains a major challenge, particularly when dealing with low-level torque control, where even small modelling inaccuracies can lead to unstable or unsafe behaviours. In this work, we propose a Real2Sim2Real pipeline that improves Sim2Real transfer for torque-controlled robotic arms by combining trajectory matching, parameter optimization via genetic algorithms, and domain randomization. Using the 7-DOF Franka Emika Panda robot, we first identify friction, inertia, and gravity compensation parameters by minimizing the error between real and simulated joint trajectories. These calibrated dynamics are then used to train a TQC-based reinforcement learning agent in simulation. The trained policy is evaluated in both Gazebo and MuJoCo environments, and finally deployed on the real robot. Our results demonstrate a significant improvement in tracking accuracy and policy robustness after parameter tuning, with smooth policy transfer from simulation to the Real-World across multiple target-reaching tasks. This work highlights the effectiveness of accurate physical modelling in enabling stable and generalizable torque-based reinforcement learning policies.

Read PDF

Similar papers

Open access Sep 2026

A transformer-based reinforcement learning framework for autonomous training of industrial robotic manipulators

Autonomous robotic manipulation in modern manufacturing requires control policies that simultaneously achieve precision, adaptation, operational safety, and multi-task capability. This paper proposes a Transformer-Based Reinforcement Learning Network (TRL-Net) that combines temporal state encoding, task-cond...

Balnur Kenjayeva, A. Galimov · 0 citations
Conference Open access 2026

Reinforcement Learning for Quadrupedal Robot Control: Taxonomy, Sim-to-Real, Robustness, and Emerging Trends

. Quadrupedal robots exhibit strong mobility in complex environments where wheeled platforms often perform poorly, but their control remains difficult. In recent years, reinforcement learning (RL) has received growing attention in quadrupedal locomotion, as it supports direct policy optimization without relying entirel...

Chen Chang · 0 citations
Preprint Aug 2026

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL

This work uses Sample-based Model Predictive Control entirely in simulation as an automated, rapidly tunable expert to generate massive offline datasets and validate the robustness of this sim-to-real framework by successfully deploying complex loco-manipulation skills across different morphologies.

Martin Schuck, Maks Sorokin, S. Manni et al. · 1 citation
Preprint Sep 2026

Real-World Reinforcement Learning with MPC Scaffolding for Dexterous Manipulation

Real-world reinforcement learning (RL) offers a promising route to dexterous manipulation policies that can adapt directly from physical interaction, but learning is hindered by inefficient early exploration and costly failures. We propose a framework that uses sampling-based model predictive control (MPC) as scaffoldi...

Emek Barış Küçüktabak, Karankumar Patel, Zhao-Dong Yang et al. · 3 citations
#reinforcement learning Open access Sep 2026

Enhancing Sim-to-Real Transfer for a High-Gear-Ratio Quadruped Robot via Extended Actuator Dynamics Identification

Reinforcement learning (RL) has become a powerful tool for quadrupedal locomotion, and a sim-to-real approach is widely adopted to avoid hardware damage during training. However, the “sim-to-real gap” remains a critical challenge, particularly for robots driven by high-gear-ratio actuators, in which nonlinear friction...

Hansol Kang, Hyunyong Lee, Jiman Park et al. · 0 citations
Preprint Sep 2026

GR2PO: Group Relative Return Policy Optimization for Continuous Robot Control

Actor-critic architecture has been widely used in continuous robot control. However, they rely on learning a value network, introducing additional computational overhead during training. Moreover, policy learning may also be affected by the approximation error of value estimation. Critic-free group relative policy opti...

Pengqin Wang, Qi-Ming Zhang, Shao-Jie Shen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.