English

A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection

Sound 2024-08-27 v1 Cryptography and Security Audio and Speech Processing

Abstract

Audio spoofing detection has become increasingly important due to the rise in real-world cases. Current spoofing detectors, referred to as spoofing countermeasures (CM), are mainly trained and focused on audio waveforms with a single speaker and short duration. This study explores spoofing detection in more realistic scenarios, where the audio is long in duration and features multiple speakers and complex acoustic conditions. We test the widely-acquired AASIST under this challenging scenario, looking at the impact of multiple variations such as duration, speaker presence, and acoustic complexities on CM performance. Our work reveals key issues with current methods and suggests preliminary ways to improve them. We aim to make spoofing detection more applicable in more in-the-wild scenarios. This research is served as an important step towards developing detection systems that can handle the challenges of audio spoofing in real-world applications.

Keywords

Cite

@article{arxiv.2408.14066,
  title  = {A Preliminary Case Study on Long-Form In-the-Wild Audio Spoofing Detection},
  author = {Xuechen Liu and Xin Wang and Junichi Yamagishi},
  journal= {arXiv preprint arXiv:2408.14066},
  year   = {2024}
}

Comments

Accepted to the 23rd International Conference of the Biometrics Special Interest Group (BIOSIG 2024). Copyright might be transferred, in such case the current version may be replaced

R2 v1 2026-06-28T18:23:39.460Z