中文
相关论文

相关论文: SwapAnything: Enabling Arbitrary Object Swapping i…

200 篇论文

Video face swapping is crucial in film and entertainment production, where achieving high fidelity and temporal consistency over long and complex video sequences remains a significant challenge. Inspired by recent advances in…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Zekai Luo , Zongze Du , Zhouhang Zhu , Hao Zhong , Muzhi Zhu , Wen Wang , Yuling Xi , Chenchen Jing , Hao Chen , Chunhua Shen

We propose a new technique for visual attribute transfer across images that may have very different appearance but have perceptually similar semantic structure. By visual attribute transfer, we mean transfer of visual information (such as…

计算机视觉与模式识别 · 计算机科学 2017-06-07 Jing Liao , Yuan Yao , Lu Yuan , Gang Hua , Sing Bing Kang

Image style transfer occupies an important place in both computer graphics and computer vision. However, most current methods require reference to stylized images and cannot individually stylize specific objects. To overcome this…

计算机视觉与模式识别 · 计算机科学 2023-11-30 Junhao Chen , Peng Rong , Jingbo Sun , Chao Li , Xiang Li , Hongwu Lv

Video face swapping aims to address two primary challenges: effectively transferring the source identity to the target video and accurately preserving the dynamic attributes of the target face, such as head poses, facial expressions,…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Xiangyang Luo , Ye Zhu , Yunfei Liu , Lijian Lin , Cong Wan , Zijian Cai , Shao-Lun Huang , Yu Li

Facial parts swapping aims to selectively transfer regions of interest from the source image onto the target image while maintaining the rest of the target image unchanged. Most studies on face swapping designed specifically for full-face…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Zheng Yu , Yaohua Wang , Siying Cui , Aixi Zhang , Wei-Long Zheng , Senzhang Wang

Large intra-class variation is the result of changes in multiple object characteristics. Images, however, only show the superposition of different variable factors such as appearance or shape. Therefore, learning to disentangle and…

计算机视觉与模式识别 · 计算机科学 2019-06-18 Dominik Lorenz , Leonard Bereska , Timo Milbich , Björn Ommer

Recent advances in diffusion-based text-to-image models have simplified creating high-fidelity images, but preserving the identity (ID) of specific elements, like a personal dog, is still challenging. Object customization, using reference…

计算机视觉与模式识别 · 计算机科学 2024-11-26 Lingjie Kong , Kai Wu , Xiaobin Hu , Wenhui Han , Jinlong Peng , Chengming Xu , Donghao Luo , Mengtian Li , Jiangning Zhang , Chengjie Wang , Yanwei Fu

Most text-to-image customization techniques fine-tune models on a small set of \emph{personal concept} images captured in minimal contexts. This often results in the model becoming overfitted to these training images and unable to…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Taewook Kim , Wei Chen , Qiang Qiu

Humans tend to form quick subjective first impressions of non-physical attributes when seeing someone's face, such as perceived trustworthiness or attractiveness. To understand what variations in a face lead to different subjective…

计算机视觉与模式识别 · 计算机科学 2025-03-19 Chaitanya Roygaga , Joshua Krinsky , Kai Zhang , Kenny Kwok , Aparna Bharati

Incorporating a customized object into image generation presents an attractive feature in text-to-image generation. However, existing optimization-based and encoder-based methods are hindered by drawbacks such as time-consuming…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Ziyang Yuan , Mingdeng Cao , Xintao Wang , Zhongang Qi , Chun Yuan , Ying Shan

Face swapping aims to transfer the identity of a source face onto a target face while preserving target-specific attributes such as pose, expression, lighting, skin tone, and makeup. However, since real ground truth for face swapping is…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Jiwon Kang , Yeji Choi , JoungBin Lee , Wooseok Jang , Jinhyeok Choi , Taekeun Kang , Yongjae Park , Myungin Kim , Seungryong Kim

While Visual Question Answering (VQA) has progressed rapidly, previous works raise concerns about robustness of current VQA models. In this work, we study the robustness of VQA models from a novel perspective: visual context. We suggest…

计算机视觉与模式识别 · 计算机科学 2022-04-06 Vipul Gupta , Zhuowan Li , Adam Kortylewski , Chenyu Zhang , Yingwei Li , Alan Yuille

Style transfer has been an important topic both in computer vision and graphics. Since the seminal work of Gatys et al. first demonstrates the power of stylization through optimization in the deep feature space, quite a few approaches have…

计算机视觉与模式识别 · 计算机科学 2019-12-24 Zhijie Wu , Chunjin Song , Yang Zhou , Minglun Gong , Hui Huang

Face-swapping models have been drawing attention for their compelling generation quality, but their complex architectures and loss functions often require careful tuning for successful training. We propose a new face-swapping model called…

计算机视觉与模式识别 · 计算机科学 2022-05-06 Jiseob Kim , Jihoon Lee , Byoung-Tak Zhang

Concept personalization methods enable large text-to-image models to learn specific subjects (e.g., objects/poses/3D models) and synthesize renditions in new contexts. Given that the image references are highly biased towards visual…

计算机视觉与模式识别 · 计算机科学 2024-04-01 You Wu , Kean Liu , Xiaoyue Mi , Fan Tang , Juan Cao , Jintao Li

Recent advances in text-to-image diffusion models have substantially improved the quality of image customization, enabling the synthesis of highly realistic images. Despite this progress, achieving fast and efficient personalization remains…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Aniket Roy , Maitreya Suin , Rama Chellappa

With the great success of text-conditioned diffusion models in creative text-to-image generation, various text-driven image editing approaches have attracted the attentions of many researchers. However, previous works mainly focus on…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Zhiyuan Ma , Guoli Jia , Bowen Zhou

Attribute image manipulation has been a very active topic since the introduction of Generative Adversarial Networks (GANs). Exploring the disentangled attribute space within a transformation is a very challenging task due to the multiple…

计算机视觉与模式识别 · 计算机科学 2020-10-07 Andrés Romero , Luc Van Gool , Radu Timofte

Recent advances in text-to-image diffusion models spurred research on personalization, i.e., a customized image synthesis, of subjects within reference images. Although existing personalization methods are able to alter the subjects'…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Yongjin Choi , Chanhun Park , Seung Jun Baek

Recent advances in text-to-3D scene generation have demonstrated significant potential to transform content creation across multiple industries. Although the research community has made impressive progress in addressing the challenges of…