ReCo: Response-Consistent Locomotion with Policy-Aware MPC for Legged Manipulation

๐Ÿ“… 2026-10-01
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This study addresses the challenge of coordinating end-effector tracking with base motion in legged manipulation. To this end, it proposes a coupled framework integrating a response-shaping training strategy with closed-loop model predictive control (MPC). This approach pioneers the combination of response consistency and policy-aware MPC by unifying reinforcement learning, system identification, and response shaping techniques, effectively resolving commandโ€“response inconsistencies under dynamic loads. Simulation results demonstrate an approximate 28% reduction in position and orientation errors. Furthermore, the proposed method is successfully validated on a physical robot, exhibiting continuous coordinated manipulation capabilities between the mobile base and the robotic arm.
๐Ÿ“ Abstract
Continuous legged manipulation requires accurate end-effector tracking while the base keeps walking. Combining reinforcement learning (RL) with model predictive control (MPC) suits this task: the learned policy provides robust locomotion, while MPC coordinates the base and arm to compensate for tracking errors. However, MPC can compensate only for base motion that it can predict, and a learned policy's command response varies with gait phase, contact, and payload. We present ReCo, a framework that couples response-consistent locomotion with policy-aware MPC for legged manipulation. Response shaping trains the policy to respond to commands consistently and repeatably across randomized dynamics. An identified closed-loop response model then lets MPC jointly plan locomotion commands and arm motion. On the simulation benchmark, ReCo reduces position and orientation root-mean-square error (RMSE) by 28.7% and 27.4% relative to the best baseline for each metric. Real-world experiments demonstrate onboard continuous legged manipulation with coordinated base and arm motion.
Problem

Research questions and friction points this paper is trying to address.

legged manipulation
model predictive control
reinforcement learning
end-effector tracking
response inconsistency
Innovation

Methods, ideas, or system contributions that make the work stand out.

Legged Manipulation
Model Predictive Control
Reinforcement Learning
Response Shaping
Policy-Aware MPC
๐Ÿ”Ž Similar Papers
๐Ÿ’ผ Related Jobs
No related jobs found.
K
Kuankuan Sima
Department of Electrical and Computer Engineering, National University of Singapore, Singapore 119077, Singapore
Y
Yichao Gao
Department of Electrical and Computer Engineering, National University of Singapore, Singapore 119077, Singapore
C
Chenxi Gu
Department of Mechanical Engineering, National University of Singapore, Singapore 119077, Singapore
K
Kefan Zhao
Department of Mechanical Engineering, National University of Singapore, Singapore 119077, Singapore
Lin Zhao
Lin Zhao
Assistant Professor, National University of Singapore
control theoryreinforcement learningroboticsautonomous vehiclespower system