Search papers, labs, and topics across Lattice.
1
0
2
HyperSafe slashes harmful response rates in fine-tuned language models to below 1% without sacrificing task performance, revolutionizing safety alignment strategies.