Search papers, labs, and topics across Lattice.
2
0
3
2
Internal representations can serve as powerful lie detectors, revealing discrepancies in LLM forecasts that CoT reasoning may obscure.
Sparse autoencoders, despite their popularity for extracting interpretable features, often fail to capture the underlying manifold structure of concepts, instead fragmenting them across multiple, diluted features.