Search papers, labs, and topics across Lattice.
University College London, University of London
1
0
3
Context compaction in agentic RL can run up to 5x faster simply by streaming the KV cache instead of flushing it鈥攁ccidentally turning standard LLMs into recurrent agents that preserve evicted context purely through RL.