Search papers, labs, and topics across Lattice.
Affiliation:
2
0
3
Runtime scaffolding beats model scale in scientific discovery: structured hypothesis tracking and trajectory branching push DeepSeek-v4-flash from 20.7% to 73.0% accuracy on blind symbolic regression, matching GPT-5.5 without relying on semantic domain clues.
AI scientific agents frequently solve empirical equations for entirely the wrong reasons, failing to identify the correct generative mechanism in over 64% of the cases where they successfully recover the observable phenomenal law.