Testing human ability to detect ‘deepfake’ images of human faces-Reference-Cited by-同舟云学术

Testing human ability to detect ‘deepfake’ images of human faces

Published:2023-01-01 Issue:1 Volume:9 Page:
ISSN:2057-2085
Container-title:Journal of Cybersecurity
language:en
Short-container-title:

Author:

Bray Sergi D¹^ORCID,Johnson Shane D¹,Kleinberg Bennett²

Affiliation:

1. Department of Security and Crime Science, University College London (UCL) , London, WC1H 9EZ , UK

2. Department of Methodology, University of Tilburg , 5000 LE Tilburg , The Netherlands

Abstract

Abstract ‘Deepfakes’ are computationally created entities that falsely represent reality. They can take image, video, and audio modalities, and pose a threat to many areas of systems and societies, comprising a topic of interest to various aspects of cybersecurity and cybersafety. In 2020, a workshop consulting AI experts from academia, policing, government, the private sector, and state security agencies ranked deepfakes as the most serious AI threat. These experts noted that since fake material can propagate through many uncontrolled routes, changes in citizen behaviour may be the only effective defence. This study aims to assess human ability to identify image deepfakes of human faces (these being uncurated output from the StyleGAN2 algorithm as trained on the FFHQ dataset) from a pool of non-deepfake images (these being random selection of images from the FFHQ dataset), and to assess the effectiveness of some simple interventions intended to improve detection accuracy. Using an online survey, participants (N = 280) were randomly allocated to one of four groups: a control group, and three assistance interventions. Each participant was shown a sequence of 20 images randomly selected from a pool of 50 deepfake images of human faces and 50 images of real human faces. Participants were asked whether each image was AI-generated or not, to report their confidence, and to describe the reasoning behind each response. Overall detection accuracy was only just above chance and none of the interventions significantly improved this. Of equal concern was the fact that participants’ confidence in their answers was high and unrelated to accuracy. Assessing the results on a per-image basis reveals that participants consistently found certain images easy to label correctly and certain images difficult, but reported similarly high confidence regardless of the image. Thus, although participant accuracy was 62% overall, this accuracy across images ranged quite evenly between 85 and 30%, with an accuracy of below 50% for one in every five images. We interpret the findings as suggesting that there is a need for an urgent call to action to address this threat.

Funder

EPSRC

Publisher

Oxford University Press (OUP)

Subject

Law,Computer Networks and Communications,Political Science and International Relations,Safety, Risk, Reliability and Quality,Social Psychology,Computer Science (miscellaneous)

Link

https://academic.oup.com/cybersecurity/article-pdf/9/1/tyad011/51343718/tyad011.pdf

Reference143 articles.

1. The Deepfake Detection Challenge (DFDC) preview dataset;Dolhansky;arXiv:1910.08854,2019

2. Protecting World Leaders Against Deep Fakes;Agarwal,2019

3. Prescriptivist vs descriptivist—what’s the difference?;Conor;One Minute English,2021

4. Authority and american usage;Foster Wallace,2006

Cited by 16 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Comprehensive multiparametric analysis of human deepfake speech recognition;EURASIP Journal on Image and Video Processing;2024-08-30

2. Identifying and preventing future forms of crimes using situational crime prevention;Security Journal;2024-08-29

3. A systematic review of AI literacy scales;npj Science of Learning;2024-08-06

4. Synthetic And Natural Face Identity Processing Share Common Mechanisms;2024-08-06

5. Clicks and tricks: The dark art of online persuasion;Current Opinion in Psychology;2024-08