Search papers, labs, and topics across Lattice.
3
0
4
Current AI coding agents struggle with large-scale refactoring tasks, achieving only a 41.2% success rate on a newly curated benchmark designed to challenge their capabilities.
Forget task-specific overfitting: training coding agents on atomic skills unlocks surprisingly broad generalization to complex software engineering tasks.