Search papers, labs, and topics across Lattice.
1
0
3
LLMs still fail basic science: even the best models struggle to answer questions grounded in procedurally-generated, noise-free scientific data, achieving only 45% accuracy.