Search papers, labs, and topics across Lattice.
4
0
9
17
Self-improving agents can evolve autonomously with minimal human input, reshaping our approach to AI adaptability and deployment.
Smaller LLMs can learn to predict when they'll fail, paving the way for efficient "ask for help" systems that rival the performance of much larger models.
Forget agents and world models – the future of computing could be learned directly from I/O traces, turning the model itself into the computer.
Scale up offline policy training for diffusion LLMs without breaking the bank: dTRPO slashes trajectory computation costs while boosting performance up to 9.6% on STEM tasks.