Search papers, labs, and topics across Lattice.
Affiliation:
2
5
5
9
Robots can now learn and adapt in real-time, overcoming inference delays that previously stymied reinforcement learning effectiveness.
Continuous scoring from LLM-as-a-Verifier leads to state-of-the-art verification accuracy and improved sample efficiency in reinforcement learning tasks.