Both eyes open: Vigilant Incentives help auditors improve AI safety-Reference-Cited by-同舟云学术

Both eyes open: Vigilant Incentives help auditors improve AI safety

Published:2024-05-07 Issue:2 Volume:5 Page:025009
ISSN:2632-072X
Container-title:Journal of Physics: Complexity
language:
Short-container-title:J. Phys. Complex.

Author:

Bova Paolo^ORCID,Stefano Alessandro Di^ORCID,Han The Anh^ORCID

Abstract

Abstract Auditors can play a vital role in ensuring that tech companies develop and deploy AI systems safely, taking into account not just immediate, but also systemic harms that may arise from the use of future AI capabilities. However, to support auditors in evaluating the capabilities and consequences of cutting-edge AI systems, governments may need to encourage a range of potential auditors to invest in new auditing tools and approaches. We use evolutionary game theory to model scenarios where the government wishes to incentivise auditing but cannot discriminate between high and low-quality auditing. We warn that it is alarmingly easy to stumble on ‘Adversarial Incentives’, which prevent a sustainable market for auditing AI systems from forming. Adversarial Incentives mainly reward auditors for catching unsafe behaviour. If AI companies learn to tailor their behaviour to the quality of audits, the lack of opportunities to catch unsafe behaviour will discourage auditors from innovating. Instead, we recommend that governments always reward auditors, except when they find evidence that those auditors failed to detect unsafe behaviour they should have. These ‘Vigilant Incentives’ could encourage auditors to find innovative ways to evaluate cutting-edge AI systems. Overall, our analysis provides useful insights for the design and implementation of efficient incentive strategies for encouraging a robust auditing ecosystem.

Publisher

IOP Publishing

Link

https://iopscience.iop.org/article/10.1088/2632-072X/ad424c/pdf

Reference90 articles.

1. Collective Games on Hypergraphs;Alvarez-Rodriguez,2022

2. Concrete problems in AI safety;Amodei,2016

3. Racing to the precipice: a model of artificial intelligence development;Armstrong;AI Soc.,2016

4. The Role of Cooperation in Responsible AI Development;Askell,2019

5. A dynamic quality ladder model with entry and exit: exploring the equilibrium correspondence using the homotopy method;Bar (formerly Borkovsky),2009

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Does government policy matter in the digital transformation of farmers’ cooperatives?—A tripartite evolutionary game analysis;Frontiers in Sustainable Food Systems;2024-06-25

2. Three-party evolutionary game-based analysis and stability enhancement of improved PBFT consensus mechanism;Cluster Computing;2024-06-09