Search papers, labs, and topics across Lattice.
3
0
5
Native computer use can be achieved at scale, enabling agents to outperform leading systems while significantly enhancing security against adversarial attacks.
Current AI agents only manage to complete 20.6% of complex real-world tasks, revealing a stark gap in their capabilities compared to human users.
Forget hand-crafted benchmarks: CUA-Gym's auto-generated training data lets computer-use agents crush existing open-source models on real-world tasks.