English

Efficient Fine-Tuning with Domain Adaptation for Privacy-Preserving Vision Transformer

Computer Vision and Pattern Recognition 2024-02-12 v2 Machine Learning

Abstract

We propose a novel method for privacy-preserving deep neural networks (DNNs) with the Vision Transformer (ViT). The method allows us not only to train models and test with visually protected images but to also avoid the performance degradation caused from the use of encrypted images, whereas conventional methods cannot avoid the influence of image encryption. A domain adaptation method is used to efficiently fine-tune ViT with encrypted images. In experiments, the method is demonstrated to outperform conventional methods in an image classification task on the CIFAR-10 and ImageNet datasets in terms of classification accuracy.

Keywords

Cite

@article{arxiv.2401.05126,
  title  = {Efficient Fine-Tuning with Domain Adaptation for Privacy-Preserving Vision Transformer},
  author = {Teru Nagamori and Sayaka Shiota and Hitoshi Kiya},
  journal= {arXiv preprint arXiv:2401.05126},
  year   = {2024}
}

Comments

Accepted by APSIPA Transactions on Signal and Information Processing. arXiv admin note: substantial text overlap with arXiv:2309.02556