Search papers, labs, and topics across Lattice.
2
99
4
12
ReToken boosts vision-language model performance by over 20% on key benchmarks while fitting training and inference within a single GPU's memory constraints.
Forget task-specific models: Magma, a single foundation model, now outperforms them in both UI navigation and robotic manipulation by bridging verbal and action abilities.