Search papers, labs, and topics across Lattice.
1
0
2
It is shown that the standard advice to prefer log-probabilities no longer holds on post-2025 models, where verbalized confidence is the better signal, and recommended broader use of soft scoring in LLM-as-a-Judge is recommended.