Search papers, labs, and topics across Lattice.
2
1
4
0
A new framework, Open-MOPD, boosts capability integration in multi-teacher distillation from 35.6% to 83.4% by addressing critical optimization imbalances.
Weak models can teach strong ones how to act better, boosting performance without the heavy lifting of direct RL training.