Search papers, labs, and topics across Lattice.
2
0
6
Tmax sets a new standard for terminal agent performance with a surprisingly simple RL recipe that outshines larger models.
Forget passively analyzing model outputs – this new attack actively *trains* the model to regurgitate specific texts, revealing its training data with surprising accuracy.