Search papers, labs, and topics across Lattice.
Tencent
2
0
3
Weak policies can achieve up to 18.6% better performance with PATS, a training framework that dynamically adapts guidance based on evolving policy needs.
LLM agents can learn to use tools more efficiently and accurately by explicitly learning when *not* to use them, leading to a 25% increase in tool productivity.