Search papers, labs, and topics across Lattice.
This paper critiques the current evaluation methods for Explainable Artificial Intelligence (XAI), particularly highlighting their inadequacies through the lens of the DetoxAI system, which focuses on bias detection and concept unlearning in image recognition. It presents a human-grounded evaluation of explanation methods for image classification and discusses the adaptation of explanations to dynamic data streams affected by concept drift. The findings underscore the complexities of maintaining effective explanations as data, models, and explanations co-evolve, revealing significant gaps in existing evaluation frameworks.
Current XAI evaluation methods fall short, risking the effectiveness of bias detection and concept unlearning in evolving data environments.
This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through the DetoxAI image recognition system for bias detection and concept unlearning. Then, an example of a human-grounded evaluation of methods for explaining image classification is presented. The paper further explores methods for adapting explanations to evolving data streams with concept drift. Experiences with adapting counterfactuals for this problem are discussed. Finally it is related to the challenges of tracking the co-evolution of data, models, and explanations.\footnote{This paper has been accepted for a publication in J.Nalepa (ed) Explainable AI in Space. Proceedings of EASi 2026 Workshop at IJCAI-ECAI 2026 Bremen, Springer CCIS vol 3107 (2016).}