English

Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories

Computers and Society 2026-05-12 v1

Abstract

For nearly a decade, deepfake detection has been framed as a classification task: given an audio or video clip, decide whether it is real or synthetic. Top detectors often report high accuracy on standard benchmarks; however, performance drops sharply on content from newer or unseen generators. We argue that better classifiers of synthetic media alone will not solve this problem, especially for interactive deepfakes such as impersonation in video and voice calls, where the harm lies not in the artifact (manipulated media signal) but in the act of deception. Deepfake detection therefore requires a complementary analytical layer focused on communicative interaction, not just media realism. We identify five assumptions that artifact-based detection (the forensic analysis of low-level signal traces) relies on and show that all five are eroding as generative models improve, producing what we call the Generalization Illusion. To address this, we draw on three well-established frameworks from philosophy of language and social psychology, namely, Speech Act Theory, Grice's Cooperative Principle, and Cialdini's principles of influence, to examine forensic signals at three levels: the utterance, the conversation, and the listener response. The result is a unified framework that complements existing forensic methods. We close with open problems for future work. https://jesseeho.github.io/deepfake-deception/

Keywords

Cite

@article{arxiv.2605.09007,
  title  = {Detecting Deception, Not Deepfakes: Why Media Forensics Needs Social Theories},
  author = {Jessee Ho and Shweta Khushu and Shaina Raza},
  journal= {arXiv preprint arXiv:2605.09007},
  year   = {2026}
}

Comments

Position paper. 18 pages, 5 figures, 4 tables. Project page included

R2 v1 2026-07-01T13:00:05.296Z