Understanding, Idealization, and Explainable AI-Reference-Cited by-同舟云学术

Understanding, Idealization, and Explainable AI

Published:2022-11-03 Issue:4 Volume:19 Page:534-560
ISSN:1742-3600
Container-title:Episteme
language:en
Short-container-title:Episteme

Author:

Fleisher Will^ORCID

Abstract

AbstractMany AI systems that make important decisions are black boxes: how they function is opaque even to their developers. This is due to their high complexity and to the fact that they are trained rather than programmed. Efforts to alleviate the opacity of black box systems are typically discussed in terms of transparency, interpretability, and explainability. However, there is little agreement about what these key concepts mean, which makes it difficult to adjudicate the success or promise of opacity alleviation methods. I argue for a unified account of these key concepts that treats the concept of understanding as fundamental. This allows resources from the philosophy of science and the epistemology of understanding to help guide opacity alleviation efforts. A first significant benefit of this understanding account is that it defuses one of the primary, in-principle objections to post hoc explainable AI (XAI) methods. This “rationalization objection” argues that XAI methods provide mere rationalizations rather than genuine explanations. This is because XAI methods involve using a separate “explanation” system to approximate the original black box system. These explanation systems function in a completely different way than the original system, yet XAI methods make inferences about the original system based on the behavior of the explanation system. I argue that, if we conceive of XAI methods as idealized scientific models, this rationalization worry is dissolved. Idealized scientific models misrepresent their target phenomena, yet are capable of providing significant and genuine understanding of their targets.

Publisher

Cambridge University Press (CUP)

Subject

History and Philosophy of Science

Reference69 articles.

1. Deep learning

2. Understanding from Machine Learning Models

3. Lakkaraju, H. , Adebayo, J. and Singh, S. (2020). ‘Explaining ml Predictions: State of the Art, Challenges, Opportunities.’ In Neurips ’20. https://explainml-tutorial.github.io/neurips20.

Cited by 17 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Do opaque algorithms have functions?;Synthese;2024-08-29

2. Non-Representational Models and Objectual Understanding;Erkenntnis;2024-08-24

3. Mapping the landscape of ethical considerations in explainable AI research;Ethics and Information Technology;2024-06-25

4. What do algorithms explain? The issue of the goals and capabilities of Explainable Artificial Intelligence (XAI);Humanities and Social Sciences Communications;2024-06-14

5. SIDEs: Separating Idealization from Deceptive 'Explanations' in xAI;The 2024 ACM Conference on Fairness, Accountability, and Transparency;2024-06-03