Search papers, labs, and topics across Lattice.
1
0
2
3
Alignment tuning installs distinct bias directions in LLMs, allowing for targeted debiasing that recovers unbiased answers while maintaining performance.