Search papers, labs, and topics across Lattice.
This paper introduces ShapShift, a Shapley value-based method for attributing prediction shifts in ML models to changes in conditional probabilities of interpretable subgroups defined by decision trees. ShapShift decomposes the overall prediction shift by quantifying the contribution of each subgroup's conditional probability change to the model's output variation. The method is extended to tree ensembles and generalized to model-agnostic settings using surrogate trees trained with a novel objective function, enabling explanation of prediction shifts in complex models like neural networks.
Pinpointing *why* your model's predictions are changing just got easier: ShapShift uses Shapley values and decision trees to explain prediction shifts by attributing them to changes in the conditional probabilities of interpretable data subgroups.
Changes in input distribution can induce shifts in the average predictions of machine learning models. Such prediction shifts may impact downstream business outcomes (e.g. a bank's loan approval rate), so understanding their causes can be crucial. We propose \ours{}: a Shapley value method for attributing prediction shifts to changes in the conditional probabilities of interpretable subgroups of data, where these subgroups are defined by the structure of decision trees. We initially apply this method to single decision trees, providing exact explanations based on conditional probability changes at split nodes. Next, we extend it to tree ensembles by selecting the most explanatory tree and accounting for residual effects. Finally, we propose a model-agnostic variant using surrogate trees grown with a novel objective function, allowing application to models like neural networks. While exact computation can be intensive, approximation techniques enable practical application. We show that \ours{} provides simple, faithful, and near-complete explanations of prediction shifts across model classes, aiding model monitoring in dynamic environments.