English
Related papers

Related papers: Pixel Codec Avatars

200 papers

While progress in 2D generative models of human appearance has been rapid, many applications require 3D avatars that can be animated and rendered. Unfortunately, most existing methods for learning generative models of 3D humans with diverse…

Computer Vision and Pattern Recognition · Computer Science 2023-05-04 Zijian Dong , Xu Chen , Jinlong Yang , Michael J. Black , Otmar Hilliges , Andreas Geiger

Creating high-fidelity head avatars from multi-view videos is a core issue for many AR/VR applications. However, existing methods usually struggle to obtain high-quality renderings for all different head components simultaneously since they…

Computer Vision and Pattern Recognition · Computer Science 2026-02-20 Cong Wang , Di Kang , He-Yi Sun , Shen-Han Qian , Zi-Xuan Wang , Linchao Bao , Song-Hai Zhang

Text-to-image person re-identification (ReID) aims to retrieve images of a person based on a given textual description. The key challenge is to learn the relations between detailed information from visual and textual modalities. Existing…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Dixuan Lin , Yixing Peng , Jingke Meng , Wei-Shi Zheng

We present HRM$^2$Avatar, a framework for creating high-fidelity avatars from monocular phone scans, which can be rendered and animated in real time on mobile devices. Monocular capture with smartphones provides a low-cost alternative to…

Graphics · Computer Science 2025-10-30 Chao Shi , Shenghao Jia , Jinhui Liu , Yong Zhang , Liangchao Zhu , Zhonglei Yang , Jinze Ma , Chaoyue Niu , Chengfei Lv

While recent 3D head avatar creation methods attempt to animate facial dynamics, they often fail to capture personalized details, limiting realism and expressiveness. To fill this gap, we present DipGuava (Disentangled and Personalized…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Jeonghaeng Lee , Seok Keun Choi , Zhixuan Li , Weisi Lin , Sanghoon Lee

One of the most exciting applications of vision models involve pixel-level reasoning. Despite the abundance of vision foundation models, we still lack representations that effectively embed spatio-temporal properties of visual scenes at the…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Nikita Araslanov , Martin Sundermeyer , Hidenobu Matsuki , David Joseph Tan , Federico Tombari

Near infrared (NIR) to Visible (VIS) face matching is challenging due to the significant domain gaps as well as a lack of sufficient data for cross-modality model training. To overcome this problem, we propose a novel method for paired…

Computer Vision and Pattern Recognition · Computer Science 2022-11-14 Yunqi Miao , Alexandros Lattas , Jiankang Deng , Jungong Han , Stefanos Zafeiriou

Visual foundation models like CLIP excel in learning feature representations from extensive datasets through self-supervised methods, demonstrating remarkable transfer learning and generalization capabilities. A growing number of…

Computer Vision and Pattern Recognition · Computer Science 2023-06-23 Binjie Zhang , Yixiao Ge , Xuyuan Xu , Ying Shan , Mike Zheng Shou

Image inpainting is a fundamental research area between image editing and image generation. Recent state-of-the-art (SOTA) methods have explored novel attention mechanisms, lightweight architectures, and context-aware modeling,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Ziyang Xu , Kangsheng Duan , Xiaolei Shen , Zhifeng Ding , Wenyu Liu , Xiaohu Ruan , Xiaoxin Chen , Xinggang Wang

Neural representations have emerged as a new paradigm for applications in rendering, imaging, geometric modeling, and simulation. Compared to traditional representations such as meshes, point clouds, or volumes they can be flexibly…

Computer Vision and Pattern Recognition · Computer Science 2021-05-07 Julien N. P. Martel , David B. Lindell , Connor Z. Lin , Eric R. Chan , Marco Monteiro , Gordon Wetzstein

We present AvatarPopUp, a method for fast, high quality 3D human avatar generation from different input modalities, such as images and text prompts and with control over the generated pose and shape. The common theme is the use of…

Computer Vision and Pattern Recognition · Computer Science 2024-07-15 Nikos Kolotouros , Thiemo Alldieck , Enric Corona , Eduard Gabriel Bazavan , Cristian Sminchisescu

Reconstructing a dynamic human with loose clothing is an important but difficult task. To address this challenge, we propose a method named DLCA-Recon to create human avatars from monocular videos. The distance from loose clothing to the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-21 Chunjie Luo , Fei Luo , Yusen Wang , Enxu Zhao , Chunxia Xiao

We present DenseRaC, a novel end-to-end framework for jointly estimating 3D human pose and body shape from a monocular RGB image. Our two-step framework takes the body pixel-to-surface correspondence map (i.e., IUV map) as proxy…

Computer Vision and Pattern Recognition · Computer Science 2019-10-10 Yuanlu Xu , Song-Chun Zhu , Tony Tung

Recent advances in full-head reconstruction have been obtained by optimizing a neural field through differentiable surface or volume rendering to represent a single scene. While these techniques achieve an unprecedented accuracy, they take…

Computer Vision and Pattern Recognition · Computer Science 2024-04-08 Antonio Canela , Pol Caselles , Ibrar Malik , Eduard Ramon , Jaime García , Jordi Sánchez-Riera , Gil Triginer , Francesc Moreno-Noguer

We present PrismAvatar: a 3D head avatar model which is designed specifically to enable real-time animation and rendering on resource-constrained edge devices, while still enjoying the benefits of neural volumetric rendering at training…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Prashant Raina , Felix Taubner , Mathieu Tuli , Eu Wern Teh , Kevin Ferreira

We present DreamHuman, a method to generate realistic animatable 3D human avatar models solely from textual descriptions. Recent text-to-3D methods have made considerable strides in generation, but are still lacking in important aspects.…

Computer Vision and Pattern Recognition · Computer Science 2023-06-16 Nikos Kolotouros , Thiemo Alldieck , Andrei Zanfir , Eduard Gabriel Bazavan , Mihai Fieraru , Cristian Sminchisescu

Acquisition and creation of digital human avatars is an important problem with applications to virtual telepresence, gaming, and human modeling. Most contemporary approaches for avatar generation can be viewed either as 3D-based methods,…

Computer Vision and Pattern Recognition · Computer Science 2022-03-30 Amit Raj , Umar Iqbal , Koki Nagano , Sameh Khamis , Pavlo Molchanov , James Hays , Jan Kautz

We present AvatarReX, a new method for learning NeRF-based full-body avatars from video data. The learnt avatar not only provides expressive control of the body, hands and the face together, but also supports real-time animation and…

Computer Vision and Pattern Recognition · Computer Science 2023-05-09 Zerong Zheng , Xiaochen Zhao , Hongwen Zhang , Boning Liu , Yebin Liu

We introduce a highly robust GAN-based framework for digitizing a normalized 3D avatar of a person from a single unconstrained photo. While the input image can be of a smiling person or taken in extreme lighting conditions, our method can…

Computer Vision and Pattern Recognition · Computer Science 2021-06-23 Huiwen Luo , Koki Nagano , Han-Wei Kung , Mclean Goldwhite , Qingguo Xu , Zejian Wang , Lingyu Wei , Liwen Hu , Hao Li

In this work we propose a novel model-based deep convolutional autoencoder that addresses the highly challenging problem of reconstructing a 3D human face from a single in-the-wild color image. To this end, we combine a convolutional…

Computer Vision and Pattern Recognition · Computer Science 2017-12-11 Ayush Tewari , Michael Zollhöfer , Hyeongwoo Kim , Pablo Garrido , Florian Bernard , Patrick Pérez , Christian Theobalt