Search papers, labs, and topics across Lattice.
3
0
5
The proposed method accelerates training and enables stable policies where matched raw-action PPO remains near failure, with successful sim-to-real transfer in in-hand and arm-hand tasks.