Search papers, labs, and topics across Lattice.
This paper investigates how counterfactual explanations can enhance contestability in algorithmic decision-making, thereby restoring agency to decision-subjects. The authors define contestability as the provision of sufficient information for individuals to challenge decisions made by opaque machine learning models and differentiate it from related concepts like justification and recourse. Through a critical examination of counterfactual explanations and their limitations, the study proposes modifications to improve their effectiveness, including a multi-shot querying approach that allows users to test their own counterfactuals.
Counterfactual explanations can empower individuals to contest algorithmic decisions, but only if they are tailored to effectively reveal underlying errors.
The automation of consequential decisions through opaque machine learning models in societal domains impedes our agency. This paper is about how agency can be reinstated by the provision of certain kinds of knowledge. More precisely, we discuss whether a specific type of explanation, counterfactual explanations, facilitates our ability to contest algorithmic decisions. Against this backdrop, our paper makes three contributions: First, we develop an account of contestability, where contestability is defined as the provision of information, sufficient for a decision-subject to use as a basis for demanding that a decision be revoked. We also demarcate contestability from adjacent concepts in the discourse surrounding the right to explanation, such as justification and recourse. Second, we examine to what extent counterfactual explanations are conducive to contestability by considering a variety of failure modes causing problematic algorithmic decisions and scrutinize to what extent counterfactual explanations help us detect the underlying errors. Third, we propose ways in which, with certain modifications, counterfactual explanations can be made more fitting to serve the desired function. In this vein, we sketch the contours of a multi-shot approach to counterfactuals, where decision-subjects can query a model to test their own counterfactuals for a (limited) number of times.