Search papers, labs, and topics across Lattice.
2
0
5
RLPF transforms how code generation models are trained by prioritizing runtime efficiency alongside correctness, leading to a dramatic increase in both runnable solutions and execution speed.
3D policy learning can finally leverage powerful 3D perception models thanks to a new architecture that overcomes training instabilities and overfitting.