Search papers, labs, and topics across Lattice.
Affiliation:
2
0
6
7
Learned routers can outperform fixed-model baselines by 14.6%, revealing a new frontier in efficient LLM deployment.
Self-evolving rubric rewards can dramatically enhance audio reasoning in models, outperforming traditional methods by adapting to the model's evolving capabilities.