Search papers, labs, and topics across Lattice.
2
0
5
4
K-EXAONE 2.0 achieves over three times the capacity of its predecessor while enhancing multilingual capabilities and long-context reasoning.
Diffusion language models can achieve up to 26x inference speedups with almost no accuracy loss, thanks to a clever entropy-based KV caching strategy that avoids costly full forward passes.