Online GP-MPC Command Supervision for Robust Reinforcement Learning-Based Quadruped Locomotion
Reinforcement learning-based quadruped locomotion policies can exhibit command-tracking errors under terrain variations and unmodeled dynamics. This study proposes an online bounded Gaussian Process-enhanced model predictive control framework, termed Gaussian Process–Model Predictive Control–Reinforcement Learning(GP-M...