Search papers, labs, and topics across Lattice.
2
0
5
0
Autoresearch agents can waste compute resolving the same issues repeatedly, but targeted interventions can dramatically enhance their efficiency and performance.
Model accuracy in reasoning about human rights law varies dramatically, with scores ranging from 0.025 to 0.774, highlighting the urgent need for robust evaluation tools in AI.