Search papers, labs, and topics across Lattice.
University of Toronto
2
0
4
3
Skill-switching accuracy in LLMs drops significantly on complex tasks, but a new training approach boosts performance from 34.4% to 68.4% on challenging benchmarks.
Prioritizing domains based on their cross-domain transferability can boost multi-domain RLVR performance by up to 10%.