Search papers, labs, and topics across Lattice.
2
0
4
16
AdvancedMathBench reveals that even state-of-the-art models struggle with advanced mathematical reasoning, achieving only 75.8% accuracy in proof generation.
Visual Pretraining outperforms text-only methods, revealing that rich visual cues can enhance language model performance in ways previously underestimated.