Search papers, labs, and topics across Lattice.
Affiliation:
1
0
3
Sampling-Guided Policy Search (SGPS), which couples recurring action-target refinement by sampling-based model-predictive control with first-order policy optimization with first-order policy optimization, and shows that refinement improves policy learning beyond initialization and tracking alone.