Search papers, labs, and topics across Lattice.
3
0
4
A single distilled model can outperform larger heterogeneous teachers by effectively integrating their strengths without interference.
Bridging disparate model families, Any-OPD achieves a 4.5% increase in performance while reducing model size by 80%.
FlowAWR accelerates convergence in generative flow models by up to 5 times while ensuring superior alignment and quality under complex reward structures.