Search papers, labs, and topics across Lattice.
The University of Manchester, Institute for Infocomm Research (I虏R), A*STAR
1
0
2
2
LLM-generated rubrics can nearly match human evaluation standards, but they often miss the mark with excessive detail and scoring bias.