Search papers, labs, and topics across Lattice.
2
1
5
4
Minor architectural tweaks can lead to a staggering 47% drop in long context performance, challenging assumptions about model design.
By reusing existing data mixture ratios and only recomputing for affected domains, Olmix slashes compute costs by 74% without sacrificing downstream task performance during iterative LM development.