Search papers, labs, and topics across Lattice.
Affiliation:
1
0
1
Reasoning-oriented training amplifies self-correction and uncertainty acknowledgment, yet fails to enhance the most predictive behaviors like confidence calibration, revealing a critical gap in model training.