Search papers, labs, and topics across Lattice.
Shanghai Jiao Tong University
2
0
3
Optimizing multiple moments of failure probabilities can dramatically enhance LLM reasoning performance, outperforming traditional single-moment approaches.
Allocating rollout budgets based on state informativeness allows LLM agents to achieve superior performance in complex decision-making tasks without increasing computational costs.