Search papers, labs, and topics across Lattice.
Polytechnique Montréal
1
0
3
Context compaction in agentic RL can run up to 5x faster simply by streaming the KV cache instead of flushing itâaccidentally turning standard LLMs into recurrent agents that preserve evicted context purely through RL.