Search papers, labs, and topics across Lattice.
3
1
6
3
Verification of coding agent outputs is now the bottleneck, not generation, and targeted design can significantly enhance performance while curbing reward hacking.
Qwen-AgentWorld achieves unprecedented simulation fidelity, outperforming existing models and enabling scalable agentic reinforcement learning across diverse real-world environments.
An 80B model that runs like a 3B? Qwen3-Coder-Next shows you can get competitive coding agent performance with a fraction of the active parameters, thanks to smart training.