Search papers, labs, and topics across Lattice.
2
0
5
2
Achieving top-tier performance in complex professional tasks with a model significantly smaller than its competitors reveals a new frontier in agentic intelligence.
Decomposing GUI agent trajectories into verifiable milestones and auditing the evidence chain yields a 10% boost in RL training performance, outperforming single-judge reward systems.