Search papers, labs, and topics across Lattice.
2
0
5
0
Reallocating optimization effort based on reward saturation can boost performance by up to 9.2% in complex reasoning tasks.
Skills stabilize agent execution by transforming noisy trajectories into procedural anchors, but they can fail under brittle assumptions and incompatible contexts.