Search papers, labs, and topics across Lattice.
2
4
4
5
Achieving 1,500 tokens per second, DiffusionGemma redefines the speed-capability trade-off in language models, outpacing conventional autoregressive approaches.
Gemma 4's unified architecture and reasoning mode enable it to outperform larger models in human-rated tasks while maintaining high efficiency.