Search papers, labs, and topics across Lattice.
3
0
4
5
AutoSaddler achieves up to 10% performance improvement in LLM agents by automatically optimizing harnesses based on execution failure signals.
Knowing the *perfect* API to use or *exact* location to edit could drastically improve SWE agent performance, but knowing the perfect regression test result? Not so much.
Automating software repository build and testing across languages and platforms is now possible, unlocking scalable benchmarking and training for coding agents.