Search papers, labs, and topics across Lattice.
2
3
3
11
Finite-sample guarantees reveal how localized conformal prediction can significantly reduce miscalibration while preserving coverage.
Ditch reward models: Nash Mirror Prox achieves fast, stable convergence to a Nash equilibrium directly from human preferences, sidestepping the limitations of traditional RLHF.