Search papers, labs, and topics across Lattice.
2
0
6
1
Kimi K3's innovative architecture achieves a 2.5x scaling efficiency improvement, enabling robust performance across diverse long-horizon tasks.
LLM agents get stuck in error feedback loops, but ProCeedRL's process-level critic and reflection-based demonstrations can actively break these cycles and substantially improve exploration.