Search papers, labs, and topics across Lattice.
1
0
12
Directly verbalizing sparse autoencoder features from LLM representations transforms how we interpret model behavior, making explanations more efficient and insightful.