Search papers, labs, and topics across Lattice.
1
0
2
Over 70% of sparse autoencoder features exhibit fundamentally mismatched input and output semantics, breaking the common assumption that what activates a feature directly mirrors its downstream causal effect.