English
Related papers

Related papers: TransEditor: Transformer-Based Dual-Space GAN for …

200 papers

Facial attributes in StyleGAN generated images are entangled in the latent space which makes it very difficult to independently control a specific attribute without affecting the others. Supervised attribute editing requires annotated…

Computer Vision and Pattern Recognition · Computer Science 2021-07-20 Kanglin Liu , Gaofeng Cao , Fei Zhou , Bozhi Liu , Jiang Duan , Guoping Qiu

Deep generative models like StyleGAN hold the promise of semantic image editing: modifying images by their content, rather than their pixel values. Unfortunately, working with arbitrary images requires inverting the StyleGAN generator,…

Computer Vision and Pattern Recognition · Computer Science 2022-05-16 Yohan Poirier-Ginter , Alexandre Lessard , Ryan Smith , Jean-François Lalonde

Current Generative Adversarial Networks (GANs) produce photorealistic renderings of portrait images. Embedding real images into the latent space of such models enables high-level image editing. While recent methods provide considerable…

Graphics · Computer Science 2021-09-21 Thomas Leimkühler , George Drettakis

Facial Attribute Manipulation (FAM) aims to aesthetically modify a given face image to render desired attributes, which has received significant attention due to its broad practical applications ranging from digital entertainment to…

Computer Vision and Pattern Recognition · Computer Science 2022-10-25 Yunfan Liu , Qi Li , Qiyao Deng , Zhenan Sun , Ming-Hsuan Yang

Generative Adversarial Networks (GANs) have emerged as powerful tools for high-quality image generation and real image editing by manipulating their latent spaces. Recent advancements in GANs include 3D-aware models such as EG3D, which…

Computer Vision and Pattern Recognition · Computer Science 2025-04-16 Bahri Batuhan Bilecen , Yigit Yalin , Ning Yu , Aysegul Dundar

While the recent advances in research on video reenactment have yielded promising results, the approaches fall short in capturing the fine, detailed, and expressive facial features (e.g., lip-pressing, mouth puckering, mouth gaping, and…

Computer Vision and Pattern Recognition · Computer Science 2023-02-15 Trevine Oorloff , Yaser Yacoob

Image-to-image (i2i) translation is the dense regression problem of learning how to transform an input image into an output using aligned image pairs. Remarkable progress has been made in i2i translation with the advent of Deep…

Computer Vision and Pattern Recognition · Computer Science 2019-08-27 Evangelos Ververas , Stefanos Zafeiriou

Generative adversarial networks (GANs) synthesize realistic images from random latent vectors. Although manipulating the latent vectors controls the synthesized outputs, editing real images with GANs suffers from i) time-consuming…

Computer Vision and Pattern Recognition · Computer Science 2021-06-24 Hyunsu Kim , Yunjey Choi , Junho Kim , Sungjoo Yoo , Youngjung Uh

StyleGAN has demonstrated the ability of GANs to synthesize highly-realistic faces of imaginary people from random noise. One limitation of GAN-based image generation is the difficulty of controlling the features of the generated image, due…

Computer Vision and Pattern Recognition · Computer Science 2025-04-25 Zhuo He , Paul Henderson , Nicolas Pugeault

Generating and manipulating human facial images using high-level attributal controls are important and interesting problems. The models proposed in previous work can solve one of these two problems (generation or manipulation), but not both…

Computer Vision and Pattern Recognition · Computer Science 2017-04-10 Weidong Yin , Yanwei Fu , Leonid Sigal , Xiangyang Xue

3D-aware GANs offer new capabilities for view synthesis while preserving the editing functionalities of their 2D counterparts. GAN inversion is a crucial step that seeks the latent code to reconstruct input images or videos, subsequently…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Yiran Xu , Zhixin Shu , Cameron Smith , Seoung Wug Oh , Jia-Bin Huang

Most existing text-to-image synthesis tasks are static single-turn generation, based on pre-defined textual descriptions of images. To explore more practical and interactive real-life applications, we introduce a new task - Interactive…

Computer Vision and Pattern Recognition · Computer Science 2020-08-07 Yu Cheng , Zhe Gan , Yitong Li , Jingjing Liu , Jianfeng Gao

StyleGAN has achieved great progress in 2D face reconstruction and semantic editing via image inversion and latent editing. While studies over extending 2D StyleGAN to 3D faces have emerged, a corresponding generic 3D GAN inversion…

Computer Vision and Pattern Recognition · Computer Science 2022-12-19 Yushi Lan , Xuyi Meng , Shuai Yang , Chen Change Loy , Bo Dai

GAN inversion aims at inverting given images into corresponding latent codes for Generative Adversarial Networks (GANs), especially StyleGAN where exists a disentangled latent space that allows attribute-based image manipulation at latent…

Computer Vision and Pattern Recognition · Computer Science 2023-08-01 Chenyi Zhuang , Pan Gao , Aljosa Smolic

Image editing has been a long-standing challenge in the research community with its far-reaching impact on numerous applications. Recently, text-driven methods started to deliver promising results in domains like human faces, but their…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Chaerin Kong , Seungyong Lee , Soohyeok Im , Wonsuk Yang

Recent works for face editing usually manipulate the latent space of StyleGAN via the linear semantic directions. However, they usually suffer from the entanglement of facial attributes, need to tune the optimal editing strength, and are…

Computer Vision and Pattern Recognition · Computer Science 2023-07-18 Zhizhong Huang , Siteng Ma , Junping Zhang , Hongming Shan

Kinship face synthesis is a challenging problem due to the scarcity and low quality of the available kinship data. Existing methods often struggle to generate descendants with both high diversity and fidelity while precisely controlling…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Pin-Yen Chiu , Dai-Jie Wu , Po-Hsun Chu , Chia-Hsuan Hsu , Hsiang-Chen Chiu , Chih-Yu Wang , Jun-Cheng Chen

Artistic text style transfer is the task of migrating the style from a source image to the target text to create artistic typography. Recent style transfer methods have considered texture control to enhance usability. However, controlling…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Shuai Yang , Zhangyang Wang , Zhaowen Wang , Ning Xu , Jiaying Liu , Zongming Guo

Although 2D generative models have made great progress in face image generation and animation, they often suffer from undesirable artifacts such as 3D inconsistency when rendering images from different camera viewpoints. This prevents them…

Computer Vision and Pattern Recognition · Computer Science 2022-10-13 Yue Wu , Yu Deng , Jiaolong Yang , Fangyun Wei , Qifeng Chen , Xin Tong

While recent research has progressively overcome the low-resolution constraint of one-shot face video re-enactment with the help of StyleGAN's high-fidelity portrait generation, these approaches rely on at least one of the following:…

Computer Vision and Pattern Recognition · Computer Science 2023-02-16 Trevine Oorloff , Yaser Yacoob