Author:
Yang Scott Cheng-Hsin,Vong Wai Keen,Sojitra Ravi B.,Folke Tomas,Shafto Patrick
Abstract
AbstractState-of-the-art deep-learning systems use decision rules that are challenging for humans to model. Explainable AI (XAI) attempts to improve human understanding but rarely accounts for how people typically reason about unfamiliar agents. We propose explicitly modelling the human explainee via Bayesian teaching, which evaluates explanations by how much they shift explainees’ inferences toward a desired goal. We assess Bayesian teaching in a binary image classification task across a variety of contexts. Absent intervention, participants predict that the AI’s classifications will match their own, but explanations generated by Bayesian teaching improve their ability to predict the AI’s judgements by moving them away from this prior belief. Bayesian teaching further allows each case to be broken down into sub-examples (here saliency maps). These sub-examples complement whole examples by improving error detection for familiar categories, whereas whole examples help predict correct AI judgements of unfamiliar cases.
Funder
Air Force Research Laboratory and DARPA
U.S. Department of Defense
NSF
Publisher
Springer Science and Business Media LLC
Reference59 articles.
1. Doshi-Velez, F., Kortz, M., Budish, R., Bavitz, C., Gershman, S., O’Brien, D. et al. Accountability of ai under the law: The role of explanation. Preprint at http://arXiv.org/1711.01134 (2017).
2. Rajpurkar, P., Irvin, J., Zhu, K., Yang, B., Mehta, H., Duan, T. et al. Chexnet: Radiologist-level pneumonia detection on chest x-rays with deep learning. Preprint at http://arXiv.org/1711.05225 (2017).
3. Esteva, A. et al. Dermatologist-level classification of skin cancer with deep neural networks. Nature 542 (7639), 115 (2017).
4. European Commission. 2018 Reform of EU Data Protection Rules (European Commission, 2018).
5. Coyle, D. & Weller, A. Explaining machine learning reveals policy challenges. Science 368 (6498), 1433–1434 (2020).
Cited by
40 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献