Search papers, labs, and topics across Lattice.
This study introduces Waypoint-Guided Reinforcement Learning (WGRL) to enable robust brachiation on a life-sized dual-arm robot, addressing the challenges of coordinated motion and precise timing control. By guiding the robot's behavior through sparsely specified waypoints and integrating task success rewards, the method achieves stable forward progression and effective failure recovery in complex environments. Evaluations in both simulated and real-world settings demonstrate the effectiveness of WGRL, paving the way for enhanced arm-based locomotion in robotics.
Achieving robust brachiation on a life-sized robot reveals how sparse waypoint guidance can enhance complex motion learning in challenging environments.
Brachiation is a form of locomotion in which primates move primarily using their arms, enabling traversal in environments without footholds. However, this motion requires highly coordinated whole-body movement and precise timing control for bar grasping and release. As a result, achieving robust behavior on life-sized robotic platforms remains challenging. In this study, we present a reinforcement learning-based method to realize brachiation on a life-sized dual-arm robot. The core of the proposed approach is Waypoint-Guided Reinforcement Learning (WGRL), a learning framework for inducing non-linear and complex motions. For high-difficulty tasks where imitation learning data are unavailable, WGRL guides behavior acquisition by sparsely specifying waypoints for the end-effector trajectory, while whole-body motion is generated through reinforcement learning. In addition, by integrating the waypoint-following guidance with rewards based on task success and mechanical energy, and training in an environment designed for Sim-to-Real transfer, the proposed method achieves both forward progression and motion stability. The acquired behavior is evaluated through Sim-to-Sim experiments under monkey-bar environments with geometric variations and hardware experiments, confirming robust brachiation including failure recovery behavior. This study provides effective learning design guidelines for realizing arm-based locomotion on life-sized robotic hardware and expanding the traversable workspace of robots.