Search papers, labs, and topics across Lattice.
1
0
2
3
Model accuracy in reasoning about human rights law varies dramatically, with scores ranging from 0.025 to 0.774, highlighting the urgent need for robust evaluation tools in AI.