Search papers, labs, and topics across Lattice.
4
1
5
6
LLMs may excel at predicting outcomes in sports, but they often converge on incorrect answers, revealing a critical flaw in their forecasting abilities.
LLMs struggle with social forecasting, achieving only 75% accuracy on a benchmark that reveals critical gaps in their understanding of temporal dynamics and probability calibration.
Atomic visual perception in MLLMs is largely unsolved, with no model surpassing 60% accuracy on a new benchmark designed to isolate perceptual capabilities.
Kimi K3's innovative architecture achieves a 2.5x scaling efficiency improvement, enabling robust performance across diverse long-horizon tasks.