Search papers, labs, and topics across Lattice.
3
0
5
Tmax sets a new standard for terminal agent performance with a surprisingly simple RL recipe that outshines larger models.
Chinese open language models aren't just catching up鈥攖hey've already left their U.S. counterparts in the dust.
Language models can get a 12% boost in multi-turn conversation quality from just 10k examples of multi-turn training data, highlighting the critical gap between single-turn and multi-turn capabilities.