Search papers, labs, and topics across Lattice.
4
0
8
0
Atomic visual perception in MLLMs is largely unsolved, with no model surpassing 60% accuracy on a new benchmark designed to isolate perceptual capabilities.
Kimi K3's innovative architecture achieves a 2.5x scaling efficiency improvement, enabling robust performance across diverse long-horizon tasks.
Mags-RL lets multimodal LLMs see the forest *and* the trees, using reinforcement learning to guide a super-resolution agent that selectively enhances image regions for improved reasoning without extra annotations.
Training on 500K automatically-curated ophthalmology instructions lets a vision-language model leapfrog general medical models in a specialized domain.