中文

ASVspoof 5:面向欺骗、深度伪造与对抗性攻击检测的资源设计、收集与验证

音频与语音处理 2025-04-28 v4

摘要

ASVspoof 5 是系列挑战赛的第五篇,旨在推动语音欺骗与深度伪造攻击的研究,以及检测解决方案的设计。我们介绍了 ASVspoof 5 数据库,其数据以众所周知的方式收集,来自多样化的声学环境(与早期 ASVspoof 数据库相比的数据质量更高),来自约 2000 名说话人(相较于约 100 名)。该数据库包含由 32 种不同算法生成的攻击,同样由众所周知方式生成,并针对不同程度的新型对抗检测模型进行优化。其中包括使用混合旧式与当代文本合成语音与语音转换模型生成的攻击,此外还包含首次纳入的对抗性攻击。ASVspoof 5 协议包括七个说话人无交叉的划分,包括用于训练不同攻击模型的两个划分、用于开发和评估替代检测模型的两个划分,以及三个额外的划分,分别构成 ASVspoof 5 的训练、开发与评估集。还可使用额外收集的来自约 3 万名说话人的辅助数据集训练说话人编码器,以实现攻击算法的实现。此外,本文还描述了使用一套自动说话人验证与欺骗/深度伪造基线检测器对新 ASVspoof 5 数据库进行的实验验证。除协议与工具用于生成伪造/深度伪造语音外,本文所述的资源已由 ASVspoof 5 2024 挑战赛参赛者使用,现向社区免费提供。

关键词

引用

@article{arxiv.2502.08857,
  title  = {ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech},
  author = {Xin Wang and Héctor Delgado and Hemlata Tak and Jee-weon Jung and Hye-jin Shim and Massimiliano Todisco and Ivan Kukanov and Xuechen Liu and Md Sahidullah and Tomi Kinnunen and Nicholas Evans and Kong Aik Lee and Junichi Yamagishi and Myeonghun Jeong and Ge Zhu and Yongyi Zang and You Zhang and Soumi Maiti and Florian Lux and Nicolas Müller and Wangyou Zhang and Chengzhe Sun and Shuwei Hou and Siwei Lyu and Sébastien Le Maguer and Cheng Gong and Hanjie Guo and Liping Chen and Vishwanath Singh},
  journal= {arXiv preprint arXiv:2502.08857},
  year   = {2025}
}

备注

Database link: https://zenodo.org/records/14498691, Database mirror link: https://huggingface.co/datasets/jungjee/asvspoof5, ASVspoof 5 Challenge Workshop Proceeding: https://www.isca-archive.org/asvspoof_2024/index.html