Can counterfactual explanations of AI systems' predictions skew lay users' causal intuitions about the world? If so, can we correct for that?

Tesic, Marko; Hahn, Ulrike

doi:10.1016/j.patter.2022.100635

Computer Science > Artificial Intelligence

arXiv:2205.06241 (cs)

[Submitted on 12 May 2022 (v1), last revised 12 Dec 2022 (this version, v2)]

Title:Can counterfactual explanations of AI systems' predictions skew lay users' causal intuitions about the world? If so, can we correct for that?

Authors:Marko Tesic, Ulrike Hahn

View PDF

Abstract:Counterfactual (CF) explanations have been employed as one of the modes of explainability in explainable AI-both to increase the transparency of AI systems and to provide recourse. Cognitive science and psychology, however, have pointed out that people regularly use CFs to express causal relationships. Most AI systems are only able to capture associations or correlations in data so interpreting them as casual would not be justified. In this paper, we present two experiment (total N = 364) exploring the effects of CF explanations of AI system's predictions on lay people's causal beliefs about the real world. In Experiment 1 we found that providing CF explanations of an AI system's predictions does indeed (unjustifiably) affect people's causal beliefs regarding factors/features the AI uses and that people are more likely to view them as causal factors in the real world. Inspired by the literature on misinformation and health warning messaging, Experiment 2 tested whether we can correct for the unjustified change in causal beliefs. We found that pointing out that AI systems capture correlations and not necessarily causal relationships can attenuate the effects of CF explanations on people's causal beliefs.

Subjects:	Artificial Intelligence (cs.AI)
Cite as:	arXiv:2205.06241 [cs.AI]
	(or arXiv:2205.06241v2 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2205.06241
Journal reference:	Patterns, 3(12), 2022
Related DOI:	https://doi.org/10.1016/j.patter.2022.100635

Submission history

From: Marko Tesic [view email]
[v1] Thu, 12 May 2022 17:39:54 UTC (1,573 KB)
[v2] Mon, 12 Dec 2022 14:49:22 UTC (1,579 KB)

Computer Science > Artificial Intelligence

Title:Can counterfactual explanations of AI systems' predictions skew lay users' causal intuitions about the world? If so, can we correct for that?

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Can counterfactual explanations of AI systems' predictions skew lay users' causal intuitions about the world? If so, can we correct for that?

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators