Search papers, labs, and topics across Lattice.
3
0
7
9
Skill Optimizers trained through execution feedback can outperform traditional models by over 9 points, revealing a critical gap in agent learning methodologies.
LLMs can be tricked into revealing harmful content by iteratively refining their own understanding of safety boundaries, turning their consistency into a vulnerability.
LLMs often fail to anticipate ecological risks arising from seemingly harmless queries, revealing a critical blind spot in their safety alignment.