Search papers, labs, and topics across Lattice.
This paper introduces ReliableNet, a novel approach that constrains the probability of making confident yet incorrect predictions (Joint Confident-Wrong or JCW) during training, addressing a critical reliability failure in deep learning models. By formulating this as a chance-constrained empirical risk minimization problem, ReliableNet ensures that the JCW probability remains below a user-defined risk budget across various datasets. The results demonstrate that ReliableNet consistently meets the JCW constraint while outperforming existing methods in terms of empirical JCW, accuracy, and selective ranking under diverse challenges such as demographic and covariate shifts.
ReliableNet is the first method to guarantee that the probability of confident but incorrect predictions stays within a user-defined risk budget across multiple datasets.
A prediction that is both confident and wrong is a critical reliability failure because it can bypass abstention and human review precisely when the model is mistaken. Empirical risk minimization (ERM) controls average loss but not this failure directly, while calibration, uncertainty estimation, conformal risk control, and selective prediction methods target related reliability properties rather than bounding the joint failure event during training. We propose ReliableNet, which constrains the Joint Confident-Wrong (JCW) probability, the probability that a prediction is simultaneously confident and incorrect, below a user-specified risk budget $伪\in(0,1)$. We formulate this as a chance-constrained ERM problem, use a conservative smooth inner approximation whose population feasibility implies the original JCW constraint. Across four tabular and two image datasets, ReliableNet is the only method certified within the JCW budget for every dataset and seed in distribution, when compared against baselines spanning ERM, post-hoc calibration, conformal risk control, and selective prediction. Under demographic, ambiguity, spurious-correlation, novel-class, and covariate shifts, it achieves the lowest empirical JCW among the compared methods while remaining very competitive in accuracy, coverage, calibration, and selective prediction. Risk-coverage results further indicate that ReliableNet achieves better selective ranking than the benchmark methods on most datasets. Overall, ReliableNet provides a principled approach to trustworthy classification.