Search papers, labs, and topics across Lattice.
2
0
3
0
Agents excel at selecting knowledge sources but struggle with actual task completion, achieving only 56.1-75.3% accuracy in answers despite near-perfect routing.
LLMs can now be benchmarked for their ability to prepare training data, revealing that a new evaluation metric outperforms traditional methods in predicting downstream utility.