中文
相关论文

相关论文: Zero-Shot Head Swapping in Real-World Scenarios

200 篇论文

Visual-inertial simultaneous localization and mapping (SLAM) is a key module of robotics and low-speed autonomous vehicles, which is usually limited by the high computation burden for practical applications. To this end, an innovative…

机器人学 · 计算机科学 2025-05-28 Bingxiang Kang , Jie Zou , Guofa Li , Pengwei Zhang , Jie Zeng , Kan Wang , Jie Li

Pansharpening aims to generate a high spatial resolution multispectral image (HRMS) by fusing a low spatial resolution multispectral image (LRMS) and a panchromatic image (PAN). The most challenging issue for this task is that only the…

图像与视频处理 · 电气工程与系统科学 2024-11-08 Xiangyu Rui , Xiangyong Cao , Yining Li , Deyu Meng

Existing 3D-aware facial generation methods face a dilemma in quality versus editability: they either generate editable results in low resolution or high-quality ones with no editing flexibility. In this work, we propose a new approach that…

计算机视觉与模式识别 · 计算机科学 2022-06-01 Jingxiang Sun , Xuan Wang , Yichun Shi , Lizhen Wang , Jue Wang , Yebin Liu

Face morphing is a problem in computer graphics with numerous artistic and forensic applications. It is challenging due to variations in pose, lighting, gender, and ethnicity. This task consists of a warping for feature alignment and a…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Guilherme Schardong , Tiago Novello , Hallison Paz , Iurii Medvedev , Vinícius da Silva , Luiz Velho , Nuno Gonçalves

One-shot talking head video generation uses a source image and driving video to create a synthetic video where the source person's facial movements imitate those of the driving video. However, differences in scale between the source and…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Fa-Ting Hong , Dan Xu

Zero-shot out-of-distribution (OOD) detection is a task that detects OOD images during inference with only in-distribution (ID) class names. Existing methods assume ID images contain a single, centered object, and do not consider the more…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Atsuyuki Miyai , Qing Yu , Go Irie , Kiyoharu Aizawa

DeepFake face swapping presents a significant threat to online security and social media, which can replace the source face in an arbitrary photo/video with the target face of an entirely different person. In order to prevent this fraud,…

计算机视觉与模式识别 · 计算机科学 2022-04-27 Junhao Dong , Yuan Wang , Jianhuang Lai , Xiaohua Xie

Maliciously-manipulated images or videos - so-called deep fakes - especially face-swap images and videos have attracted more and more malicious attackers to discredit some key figures. Previous pixel-level artifacts based detection…

计算机视觉与模式识别 · 计算机科学 2021-04-29 Weinan Guan , Wei Wang , Jing Dong , Bo Peng , Tieniu Tan

Recent approaches attempt to adapt powerful interactive segmentation models, such as SAM, to interactive matting and fine-tune the models based on synthetic matting datasets. However, models trained on synthetic data fail to generalize to…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Ruihao Xia , Yu Liang , Peng-Tao Jiang , Hao Zhang , Qianru Sun , Yang Tang , Bo Li , Pan Zhou

Facial parts swapping aims to selectively transfer regions of interest from the source image onto the target image while maintaining the rest of the target image unchanged. Most studies on face swapping designed specifically for full-face…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Zheng Yu , Yaohua Wang , Siying Cui , Aixi Zhang , Wei-Long Zheng , Senzhang Wang

We present an inverse image-formation module that can enhance the robustness of existing visual SLAM pipelines for casually captured scenarios. Casual video captures often suffer from motion blur and varying appearances, which degrade the…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Gwangtak Bae , Changwoon Choi , Hyeongjun Heo , Sang Min Kim , Young Min Kim

We present a novel high-resolution face swapping method using the inherent prior knowledge of a pre-trained GAN model. Although previous research can leverage generative priors to produce high-resolution results, their quality can suffer…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Yangyang Xu , Bailin Deng , Junle Wang , Yanqing Jing , Jia Pan , Shengfeng He

This study presents a new high-fidelity multi-modal dataset containing 16000+ geometric variants of automotive hoods useful for machine learning (ML) applications such as engineering component design and process optimization, and…

机器学习 · 计算机科学 2025-11-11 Vansh Sharma , Harish Jai Ganesh , Maryam Akram , Wanjiao Liu , Venkat Raman

We propose a novel zero-shot approach to computing correspondences between 3D shapes. Existing approaches mainly focus on isometric and near-isometric shape pairs (e.g., human vs. human), but less attention has been given to strongly…

计算机视觉与模式识别 · 计算机科学 2023-09-28 Ahmed Abdelreheem , Abdelrahman Eldesokey , Maks Ovsjanikov , Peter Wonka

Diffusion Models have become very popular for Semantic Image Synthesis (SIS) of human faces. Nevertheless, their training and inference is computationally expensive and their computational requirements are high due to the quadratic…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Filippo Botti , Alex Ergasti , Tomaso Fontanini , Claudio Ferrari , Massimo Bertozzi , Andrea Prati

Modern camera pipelines apply extensive on-device processing, such as exposure adjustment, white balance, and color correction, which, while beneficial individually, often introduce photometric inconsistencies across views. These appearance…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Jisu Shin , Richard Shaw , Seunghyun Shin , Zhensong Zhang , Hae-Gon Jeon , Eduardo Perez-Pellitero

Transformers have revolutionized image modeling tasks with adaptations like DeIT, Swin, SVT, Biformer, STVit, and FDVIT. However, these models often face challenges with inductive bias and high quadratic complexity, making them less…

计算机视觉与模式识别 · 计算机科学 2024-06-05 Badri N. Patro , Suhas Ranganath , Vinay P. Namboodiri , Vijay S. Agneeswaran

We present ImageBind, an approach to learn a joint embedding across six different modalities - images, text, audio, depth, thermal, and IMU data. We show that all combinations of paired data are not necessary to train such a joint…

计算机视觉与模式识别 · 计算机科学 2023-06-01 Rohit Girdhar , Alaaeldin El-Nouby , Zhuang Liu , Mannat Singh , Kalyan Vasudev Alwala , Armand Joulin , Ishan Misra

Facial image manipulation has achieved great progress in recent years. However, previous methods either operate on a predefined set of face attributes or leave users little freedom to interactively manipulate images. To overcome these…

计算机视觉与模式识别 · 计算机科学 2020-04-02 Cheng-Han Lee , Ziwei Liu , Lingyun Wu , Ping Luo

Generating realistic images from arbitrary views based on a single source image remains a significant challenge in computer vision, with broad applications ranging from e-commerce to immersive virtual experiences. Recent advancements in…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Ido Sobol , Chenfeng Xu , Or Litany