Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels

Wielopolski, Patryk; Furman, Oleksii; Stefanowski, Jerzy; Zięba, Maciej

Computer Science > Machine Learning

arXiv:2405.17642v1 (cs)

[Submitted on 27 May 2024 (this version), latest version 29 May 2025 (v2)]

Title:Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels

Authors:Patryk Wielopolski, Oleksii Furman, Jerzy Stefanowski, Maciej Zięba

View PDF HTML (experimental)

Abstract:Growing regulatory and societal pressures demand increased transparency in AI, particularly in understanding the decisions made by complex machine learning models. Counterfactual Explanations (CFs) have emerged as a promising technique within Explainable AI (xAI), offering insights into individual model predictions. However, to understand the systemic biases and disparate impacts of AI models, it is crucial to move beyond local CFs and embrace global explanations, which offer a~holistic view across diverse scenarios and populations. Unfortunately, generating Global Counterfactual Explanations (GCEs) faces challenges in computational complexity, defining the scope of "global," and ensuring the explanations are both globally representative and locally plausible. We introduce a novel unified approach for generating Local, Group-wise, and Global Counterfactual Explanations for differentiable classification models via gradient-based optimization to address these challenges. This framework aims to bridge the gap between individual and systemic insights, enabling a deeper understanding of model decisions and their potential impact on diverse populations. Our approach further innovates by incorporating a probabilistic plausibility criterion, enhancing actionability and trustworthiness. By offering a cohesive solution to the optimization and plausibility challenges in GCEs, our work significantly advances the interpretability and accountability of AI models, marking a step forward in the pursuit of transparent AI.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Methodology (stat.ME)
Cite as:	arXiv:2405.17642 [cs.LG]
	(or arXiv:2405.17642v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2405.17642

Submission history

From: Patryk Wielopolski [view email]
[v1] Mon, 27 May 2024 20:32:09 UTC (2,501 KB)
[v2] Thu, 29 May 2025 17:23:38 UTC (764 KB)

Computer Science > Machine Learning

Title:Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators