Search papers, labs, and topics across Lattice.
This paper presents a reinforcement learning approach to non-prehensile throwing in robotics, focusing on the ability to transport large, heavy, and deformable objects without relying on traditional grasp-based methods. By formulating the task as a Markov Decision Process and optimizing joint-space trajectories directly, the method overcomes limitations of existing model-based approaches, achieving high success rates in both simulation and real-world scenarios. The results demonstrate a 99% success rate in simulation and a 97% success rate in real-world applications, showcasing the method's robustness and generalization capabilities across diverse object types and configurations.
Achieving a 97% success rate in real-world non-prehensile throwing, this approach redefines the limits of robotic object transport beyond traditional grasping techniques.
Robotic throwing enables fast object transport and extends a robot's reachable workspace beyond traditional pick-and-place. While prehensile (grasp-based) throwing works well for graspable items, non-prehensile (grasp-free) throwing is better suited for large, heavy, and/or deformable objects. Existing approaches rely on model-based optimization with simplified contact models (e.g., dynamic grasping) and low-dimensional trajectory parameterizations, which limit solution quality and reachable workspace. We propose a reinforcement learning approach that additionally leverages sliding and rolling contact modes and directly optimizes joint-space trajectories without analytical contact models or custom parameterizations. The Markov Decision Process (MDP) is formulated as a dynamical system that evolves the robot's joint state conditioned on the throwing target, object model, and initial configuration. Joint-jerk trajectories are planned offline at a low control rate and upsampled into smooth, high-rate velocity commands for deployment. For sim-to-real transfer, we minimize the robot-dynamics gap through minimum-jerk system identification and train uncertainty-aware policies to mitigate object-modeling errors, particularly sensitivity to dynamic friction. In simulation, the policy achieves 99% success across thousands of configurations and generalizes to unseen objects. Sensitivity analysis shows robustness to mass uncertainty but high sensitivity to dynamic friction, consistent with the sliding-based release mechanism. Deployed zero-shot on a UR5e operating near its physical limits (5 m/s end-effector velocity), our method throws diverse objects including heavy (790 g) and large (20x20x28 cm) items to targets up to 350 cm distance or 180 cm elevation, achieving a 97% real-world success rate.