Search papers, labs, and topics across Lattice.
4
0
6
2
Self-summarization in LLMs can enhance reasoning coherence and reduce context exhaustion, leading to a 4% performance boost with shorter rollouts.
Fine-grained decision-making in agentic RL can boost performance by nearly 4 points, challenging the reliance on coarse heuristics.
Bootstrapping LLM agents to co-evolve as both agent and environment can lead to significant performance gains, with an average improvement of over 4% on complex tasks.
LLM agents can now learn from *everyone's* experience, not just their own, leading to system-wide improvements without requiring additional user effort.