Search papers, labs, and topics across Lattice.
UK AI Security Institute
3
0
6
12
CoT monitoring can fail dramatically in implicit-influence scenarios, with detection rates dropping to as low as 5% despite behavioral shifts.
Evasion rates for distributed attacks on AI coding agents can exceed 65%, highlighting a critical vulnerability in persistent-state systems that traditional monitoring fails to address.
Behavior Forecasters can predict LRM behavior more accurately than leading models while slashing inference costs.