English
Related papers

Related papers: Zero-Shot Head Swapping in Real-World Scenarios

200 papers

Visual-inertial simultaneous localization and mapping (SLAM) is a key module of robotics and low-speed autonomous vehicles, which is usually limited by the high computation burden for practical applications. To this end, an innovative…

Robotics · Computer Science 2025-05-28 Bingxiang Kang , Jie Zou , Guofa Li , Pengwei Zhang , Jie Zeng , Kan Wang , Jie Li

Pansharpening aims to generate a high spatial resolution multispectral image (HRMS) by fusing a low spatial resolution multispectral image (LRMS) and a panchromatic image (PAN). The most challenging issue for this task is that only the…

Image and Video Processing · Electrical Eng. & Systems 2024-11-08 Xiangyu Rui , Xiangyong Cao , Yining Li , Deyu Meng

Existing 3D-aware facial generation methods face a dilemma in quality versus editability: they either generate editable results in low resolution or high-quality ones with no editing flexibility. In this work, we propose a new approach that…

Computer Vision and Pattern Recognition · Computer Science 2022-06-01 Jingxiang Sun , Xuan Wang , Yichun Shi , Lizhen Wang , Jue Wang , Yebin Liu

Face morphing is a problem in computer graphics with numerous artistic and forensic applications. It is challenging due to variations in pose, lighting, gender, and ethnicity. This task consists of a warping for feature alignment and a…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Guilherme Schardong , Tiago Novello , Hallison Paz , Iurii Medvedev , Vinícius da Silva , Luiz Velho , Nuno Gonçalves

One-shot talking head video generation uses a source image and driving video to create a synthetic video where the source person's facial movements imitate those of the driving video. However, differences in scale between the source and…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Fa-Ting Hong , Dan Xu

Zero-shot out-of-distribution (OOD) detection is a task that detects OOD images during inference with only in-distribution (ID) class names. Existing methods assume ID images contain a single, centered object, and do not consider the more…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Atsuyuki Miyai , Qing Yu , Go Irie , Kiyoharu Aizawa

DeepFake face swapping presents a significant threat to online security and social media, which can replace the source face in an arbitrary photo/video with the target face of an entirely different person. In order to prevent this fraud,…

Computer Vision and Pattern Recognition · Computer Science 2022-04-27 Junhao Dong , Yuan Wang , Jianhuang Lai , Xiaohua Xie

Maliciously-manipulated images or videos - so-called deep fakes - especially face-swap images and videos have attracted more and more malicious attackers to discredit some key figures. Previous pixel-level artifacts based detection…

Computer Vision and Pattern Recognition · Computer Science 2021-04-29 Weinan Guan , Wei Wang , Jing Dong , Bo Peng , Tieniu Tan

Recent approaches attempt to adapt powerful interactive segmentation models, such as SAM, to interactive matting and fine-tune the models based on synthetic matting datasets. However, models trained on synthetic data fail to generalize to…

Computer Vision and Pattern Recognition · Computer Science 2024-10-10 Ruihao Xia , Yu Liang , Peng-Tao Jiang , Hao Zhang , Qianru Sun , Yang Tang , Bo Li , Pan Zhou

Facial parts swapping aims to selectively transfer regions of interest from the source image onto the target image while maintaining the rest of the target image unchanged. Most studies on face swapping designed specifically for full-face…

Computer Vision and Pattern Recognition · Computer Science 2024-11-12 Zheng Yu , Yaohua Wang , Siying Cui , Aixi Zhang , Wei-Long Zheng , Senzhang Wang

We present an inverse image-formation module that can enhance the robustness of existing visual SLAM pipelines for casually captured scenarios. Casual video captures often suffer from motion blur and varying appearances, which degrade the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Gwangtak Bae , Changwoon Choi , Hyeongjun Heo , Sang Min Kim , Young Min Kim

We present a novel high-resolution face swapping method using the inherent prior knowledge of a pre-trained GAN model. Although previous research can leverage generative priors to produce high-resolution results, their quality can suffer…

Computer Vision and Pattern Recognition · Computer Science 2022-03-31 Yangyang Xu , Bailin Deng , Junle Wang , Yanqing Jing , Jia Pan , Shengfeng He

This study presents a new high-fidelity multi-modal dataset containing 16000+ geometric variants of automotive hoods useful for machine learning (ML) applications such as engineering component design and process optimization, and…

Machine Learning · Computer Science 2025-11-11 Vansh Sharma , Harish Jai Ganesh , Maryam Akram , Wanjiao Liu , Venkat Raman

We propose a novel zero-shot approach to computing correspondences between 3D shapes. Existing approaches mainly focus on isometric and near-isometric shape pairs (e.g., human vs. human), but less attention has been given to strongly…

Computer Vision and Pattern Recognition · Computer Science 2023-09-28 Ahmed Abdelreheem , Abdelrahman Eldesokey , Maks Ovsjanikov , Peter Wonka

Diffusion Models have become very popular for Semantic Image Synthesis (SIS) of human faces. Nevertheless, their training and inference is computationally expensive and their computational requirements are high due to the quadratic…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Filippo Botti , Alex Ergasti , Tomaso Fontanini , Claudio Ferrari , Massimo Bertozzi , Andrea Prati

Modern camera pipelines apply extensive on-device processing, such as exposure adjustment, white balance, and color correction, which, while beneficial individually, often introduce photometric inconsistencies across views. These appearance…

Computer Vision and Pattern Recognition · Computer Science 2025-10-01 Jisu Shin , Richard Shaw , Seunghyun Shin , Zhensong Zhang , Hae-Gon Jeon , Eduardo Perez-Pellitero

Transformers have revolutionized image modeling tasks with adaptations like DeIT, Swin, SVT, Biformer, STVit, and FDVIT. However, these models often face challenges with inductive bias and high quadratic complexity, making them less…

Computer Vision and Pattern Recognition · Computer Science 2024-06-05 Badri N. Patro , Suhas Ranganath , Vinay P. Namboodiri , Vijay S. Agneeswaran

We present ImageBind, an approach to learn a joint embedding across six different modalities - images, text, audio, depth, thermal, and IMU data. We show that all combinations of paired data are not necessary to train such a joint…

Computer Vision and Pattern Recognition · Computer Science 2023-06-01 Rohit Girdhar , Alaaeldin El-Nouby , Zhuang Liu , Mannat Singh , Kalyan Vasudev Alwala , Armand Joulin , Ishan Misra

Facial image manipulation has achieved great progress in recent years. However, previous methods either operate on a predefined set of face attributes or leave users little freedom to interactively manipulate images. To overcome these…

Computer Vision and Pattern Recognition · Computer Science 2020-04-02 Cheng-Han Lee , Ziwei Liu , Lingyun Wu , Ping Luo

Generating realistic images from arbitrary views based on a single source image remains a significant challenge in computer vision, with broad applications ranging from e-commerce to immersive virtual experiences. Recent advancements in…

Computer Vision and Pattern Recognition · Computer Science 2024-10-25 Ido Sobol , Chenfeng Xu , Or Litany