English
Related papers

Related papers: NARRATE: A Normal Assisted Free-View Portrait Styl…

200 papers

Recognizing text in natural images is a challenging task with many unsolved problems. Different from those in documents, words in natural images often possess irregular shapes, which are caused by perspective distortion, curved character…

Computer Vision and Pattern Recognition · Computer Science 2016-04-20 Baoguang Shi , Xinggang Wang , Pengyuan Lyu , Cong Yao , Xiang Bai

Current Generative Adversarial Networks (GANs) produce photorealistic renderings of portrait images. Embedding real images into the latent space of such models enables high-level image editing. While recent methods provide considerable…

Graphics · Computer Science 2021-09-21 Thomas Leimkühler , George Drettakis

Equivariance is a nice property to have as it produces much more parameter efficient neural architectures and preserves the structure of the input through the feature mapping. Even though some combinations of transformations might never…

Computer Vision and Pattern Recognition · Computer Science 2020-02-11 David W. Romero , Mark Hoogendoorn

We present a unified framework for reconstructing animatable 3D human avatars from a single portrait across head, half-body, and full-body inputs. Our method tackles three bottlenecks: pose- and framing-sensitive feature representations,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Jiawei Zhang , Lei Chu , Jiahao Li , Zhenyu Zang , Chong Li , Xiao Li , Xun Cao , Hao Zhu , Yan Lu

Reconstructing an object's geometry and appearance from multiple images, also known as inverse rendering, is a fundamental problem in computer graphics and vision. Inverse rendering is inherently ill-posed because the captured image is an…

Computer Vision and Pattern Recognition · Computer Science 2022-03-28 Akshat Dave , Yongyi Zhao , Ashok Veeraraghavan

Text-to-image diffusion models often face a severe trilemma in human portrait generation: text-image alignment, photorealism, and human-perceived aesthetics inherently inhibit one another. Supervised Fine-Tuning (SFT) is an effective method…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Yunlong Wang , Jinjin Shi , Wenbin Gao , Xuran Xu , Runyu Shi , Ying Huang

Casually-taken portrait photographs often suffer from unflattering lighting and shadowing because of suboptimal conditions in the environment. Aesthetic qualities such as the position and softness of shadows and the lighting ratio between…

Computer Vision and Pattern Recognition · Computer Science 2020-05-21 Xuaner Cecilia Zhang , Jonathan T. Barron , Yun-Ta Tsai , Rohit Pandey , Xiuming Zhang , Ren Ng , David E. Jacobs

Learning editable high-resolution scene representations for dynamic scenes is an open problem with applications across the domains from autonomous driving to creative editing - the most successful approaches today make a trade-off between…

Optical interferometric image reconstruction is a challenging, ill-posed optimization problem which usually relies on heavy regularization for convergence. Conventional algorithms regularize in the pixel domain, without cognizance of…

Recent advancements in text-guided diffusion models have unlocked powerful image manipulation capabilities. However, applying these methods to real images necessitates the inversion of the images into the domain of the pretrained diffusion…

Computer Vision and Pattern Recognition · Computer Science 2024-03-22 Daniel Garibi , Or Patashnik , Andrey Voynov , Hadar Averbuch-Elor , Daniel Cohen-Or

We propose a method for synthesizing photo-realistic digital avatars from only one portrait as the reference. Given a portrait, our method synthesizes a coarse talking head video using driving keypoints features. And with the coarse video,…

Computer Vision and Pattern Recognition · Computer Science 2023-07-20 Shaoxu Li

Embedding 3D morphable basis functions into deep neural networks opens great potential for models with better representation power. However, to faithfully learn those models from an image collection, it requires strong regularization to…

Computer Vision and Pattern Recognition · Computer Science 2019-04-11 Luan Tran , Feng Liu , Xiaoming Liu

We present NeuSE, a novel Neural SE(3)-Equivariant Embedding for objects, and illustrate how it supports object SLAM for consistent spatial understanding with long-term scene changes. NeuSE is a set of latent object embeddings created from…

Robotics · Computer Science 2023-07-11 Jiahui Fu , Yilun Du , Kurran Singh , Joshua B. Tenenbaum , John J. Leonard

In this work, a system for creating a relightable 3D portrait of a human head is presented. Our neural pipeline operates on a sequence of frames captured by a smartphone camera with the flash blinking (flash-no flash sequence). A coarse…

Computer Vision and Pattern Recognition · Computer Science 2020-12-21 Artem Sevastopolsky , Savva Ignatiev , Gonzalo Ferrer , Evgeny Burnaev , Victor Lempitsky

We present a novel framework for free-viewpoint facial performance relighting using diffusion-based image-to-image translation. Leveraging a subject-specific dataset containing diverse facial expressions captured under various lighting…

Computer Vision and Pattern Recognition · Computer Science 2024-10-11 Mingming He , Pascal Clausen , Ahmet Levent Taşel , Li Ma , Oliver Pilarski , Wenqi Xian , Laszlo Rikker , Xueming Yu , Ryan Burgert , Ning Yu , Paul Debevec

Portrait animation from a single source image and a driving video is a long-standing problem. Recent approaches tend to adopt diffusion-based image/video generation models for realistic and expressive animation. However, none of these…

Computer Vision and Pattern Recognition · Computer Science 2025-12-18 Yuxiang Shi , Zhe Li , Yanwen Wang , Hao Zhu , Xun Cao , Ligang Liu

Photorealistic editing of portraits is a challenging task as humans are very sensitive to inconsistencies in faces. We present an approach for high-quality intuitive editing of the camera viewpoint and scene illumination in a portrait…

Diffusion-based video generation techniques have significantly improved zero-shot talking-head avatar generation, enhancing the naturalness of both head motion and facial expressions. However, existing methods suffer from poor…

Graphics · Computer Science 2025-04-24 Lingzhou Mu , Baiji Liu , Ruonan Zhang , Guiming Mo , Jiawei Jin , Kai Zhang , Haozhi Huang

Recent advances in Neural Radiance Fields (NeRFs) have made it possible to reconstruct and reanimate dynamic portrait scenes with control over head-pose, facial expressions and viewing direction. However, training such models assumes…

Computer Vision and Pattern Recognition · Computer Science 2023-09-22 ShahRukh Athar , Zhixin Shu , Zexiang Xu , Fujun Luan , Sai Bi , Kalyan Sunkavalli , Dimitris Samaras

In this paper, we propose a novel framework, Tracking-free Relightable Avatar (TRAvatar), for capturing and reconstructing high-fidelity 3D avatars. Compared to previous methods, TRAvatar works in a more practical and efficient setting.…

Computer Vision and Pattern Recognition · Computer Science 2023-09-11 Haotian Yang , Mingwu Zheng , Wanquan Feng , Haibin Huang , Yu-Kun Lai , Pengfei Wan , Zhongyuan Wang , Chongyang Ma
‹ Prev 1 3 4 5 6 7 10 Next ›