Search papers, labs, and topics across Lattice.
2
0
4
LLMs can transform sparse reward signals into structured learning paths, leading to better performance in complex manipulation tasks than traditional dense reward systems.
Forget expensive real-world robotics data collection: ExpertGen uses RL to turn noisy, simulated behavior priors (even from LLMs!) into expert policies that transfer to real robots.