Warning: Humans Cannot Reliably Detect Speech Deepfakes

Mai, Kimberly T.; Bray, Sergi D.; Davies, Toby; Griffin, Lewis D.

doi:10.1371/journal.pone.0285333

Full-text links:

Download:

Current browse context:

cs.HC

< prev | next >

new | recent | 2301

Computer Science > Human-Computer Interaction

Title: Warning: Humans Cannot Reliably Detect Speech Deepfakes

Authors: Kimberly T. Mai, Sergi D. Bray, Toby Davies, Lewis D. Griffin

(Submitted on 19 Jan 2023 (v1), last revised 2 Aug 2023 (this version, v2))

Abstract: Speech deepfakes are artificial voices generated by machine learning models. Previous literature has highlighted deepfakes as one of the biggest security threats arising from progress in artificial intelligence due to their potential for misuse. However, studies investigating human detection capabilities are limited. We presented genuine and deepfake audio to n = 529 individuals and asked them to identify the deepfakes. We ran our experiments in English and Mandarin to understand if language affects detection performance and decision-making rationale. We found that detection capability is unreliable. Listeners only correctly spotted the deepfakes 73% of the time, and there was no difference in detectability between the two languages. Increasing listener awareness by providing examples of speech deepfakes only improves results slightly. As speech synthesis algorithms improve and become more realistic, we can expect the detection task to become harder. The difficulty of detecting speech deepfakes confirms their potential for misuse and signals that defenses against this threat are needed.

Subjects:	Human-Computer Interaction (cs.HC); Sound (cs.SD); Audio and Speech Processing (eess.AS)
Journal reference:	PLoS ONE 18(8) (2023): e0285333
DOI:	10.1371/journal.pone.0285333
Cite as:	arXiv:2301.07829 [cs.HC]
	(or arXiv:2301.07829v2 [cs.HC] for this version)

Submission history

From: Kimberly Mai [view email]
[v1] Thu, 19 Jan 2023 00:17:48 GMT (2265kb,D)
[v2] Wed, 2 Aug 2023 10:02:46 GMT (2140kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2301.07829

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Human-Computer Interaction

Title: Warning: Humans Cannot Reliably Detect Speech Deepfakes

Submission history