Search papers, labs, and topics across Lattice.
2
0
5
1
SkillSentry reveals that dynamic testing can uncover harmful agent behaviors that static analysis fails to detect, achieving unprecedented accuracy in safety evaluations.
Foundation models can be made intrinsically resistant to unauthorized fine-tuning by concentrating learning in a sparsely masked subnetwork that remains private.