Search papers, labs, and topics across Lattice.
2
0
5
9
Frontier LLMs can be induced to generate biologically hazardous sequences, with attack success rates reaching up to 100%.
Intrinsic reward signals in unsupervised RL for LLMs inevitably collapse due to sharpening of the model's prior, but external rewards grounded in computational asymmetries offer a path to sustained scaling.