Search papers, labs, and topics across Lattice.
Holistic AI, UCL Centre for Artificial Intelligence, University College London / Holistic AI, University College London
3
0
5
Routing supervision falters precisely when it's most needed, as weaker agents yield fewer labels, limiting the potential for optimal mode selection.
Fourteen out of sixteen large language models exhibit a systematic optimism bias in their probability judgments, raising concerns about their reliability as decision aids.
Single-token SAE features can significantly impact model outputs, with their causal roles varying dramatically across different SAE families.