中文

有效检测有效数据投毒攻击的可证明方法

密码学与安全 2025-01-22 v1 计算机视觉与模式识别 机器学习 机器学习

摘要

本文给出数据投毒攻击的精确数学定义,并证明有效地投毒数据集的行为确保该攻击可以被有效检测。我们提供了一种新的统计检测方法——Conformal Separability Test(保 conformal 可分离性检验),以数学保证数据投毒可被识别。我们还提供了实验证据,表明我们可以在真实世界中有效检测投毒尝试。

关键词

引用

@article{arxiv.2501.11795,
  title  = {Provably effective detection of effective data poisoning attacks},
  author = {Jonathan Gallagher and Yasaman Esfandiari and Callen MacPhee and Michael Warren},
  journal= {arXiv preprint arXiv:2501.11795},
  year   = {2025}
}