Search papers, labs, and topics across Lattice.
Gaoling School of Artificial Intelligence, Renmin University of China
2
0
5
19
Black-box RL can boost agent performance by nearly 15 points on complex tasks, revealing a new frontier for scalable optimization.
CalibForge reveals that adversarial calibration can dramatically enhance the effectiveness of training data for terminal agents, leading to unprecedented performance improvements on standard benchmarks.