English
Related papers

Related papers: Pixel Codec Avatars

200 papers

Multimodal alignment between language and vision is the fundamental topic in current vision-language model research. Contrastive Captioners (CoCa), as a representative method, integrates Contrastive Language-Image Pretraining (CLIP) and…

Computer Vision and Pattern Recognition · Computer Science 2024-01-05 Ziping Ma , Furong Xu , Jian Liu , Ming Yang , Qingpei Guo

Creating virtual avatars with realistic rendering is one of the most essential and challenging tasks to provide highly immersive virtual reality (VR) experiences. It requires not only sophisticated deep neural network (DNN) based codec…

Hardware Architecture · Computer Science 2021-03-09 Xiaofan Zhang , Dawei Wang , Pierce Chuang , Shugao Ma , Deming Chen , Yuecheng Li

We present Real2Code, a novel approach to reconstructing articulated objects via code generation. Given visual observations of an object, we first reconstruct its part geometry using an image segmentation model and a shape completion model.…

Computer Vision and Pattern Recognition · Computer Science 2024-06-14 Zhao Mandi , Yijia Weng , Dominik Bauer , Shuran Song

Existing Gaussian avatar methods typically parameterize geometry on a body-template surface, which entangles the avatar's representation space with the template's deformation space and limits the capture of layered, off-body, and non-rigid…

Graphics · Computer Science 2026-05-21 Julian Kaltheuner , Jan Spindler , Sina Kitz , Patrick Stotko , Reinhard Klein

While self-supervised learning has enabled effective representation learning in the absence of labels, for vision, video remains a relatively untapped source of supervision. To address this, we propose Pixel-level Correspondence (PiCo), a…

Computer Vision and Pattern Recognition · Computer Science 2022-07-11 Yash Sharma , Yi Zhu , Chris Russell , Thomas Brox

Neural implicit fields are powerful for representing 3D scenes and generating high-quality novel views, but it remains challenging to use such implicit representations for creating a 3D human avatar with a specific identity and artistic…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Ruixiang Jiang , Can Wang , Jingbo Zhang , Menglei Chai , Mingming He , Dongdong Chen , Jing Liao

Despite recent progress in developing animatable full-body avatars, realistic modeling of clothing - one of the core aspects of human self-expression - remains an open challenge. State-of-the-art physical simulation methods can generate…

We propose a universal image reconstruction method to represent detailed images purely from binary sparse edge and flat color domain. Inspired by the procedures of painting, our framework, based on generative adversarial network, consists…

Computer Vision and Pattern Recognition · Computer Science 2019-03-26 Sheng You , Ning You , Minxue Pan

We present a novel pipeline for learning high-quality triangular human avatars from multi-view videos. Recent methods for avatar learning are typically based on neural radiance fields (NeRF), which is not compatible with traditional…

Computer Vision and Pattern Recognition · Computer Science 2024-07-12 Yushuo Chen , Zerong Zheng , Zhe Li , Chao Xu , Yebin Liu

Modeling and rendering photorealistic avatars is of crucial importance in many applications. Existing methods that build a 3D avatar from visual observations, however, struggle to reconstruct clothed humans. We introduce PhysAvatar, a novel…

The human visual system has a hierarchical structure consisting of layers of processing, such as the retina, V1, V2, etc. Understanding the functional roles of these visual processing layers would help to integrate the psychophysiological…

Computer Vision and Pattern Recognition · Computer Science 2014-12-19 Honghao Shan , Garrison Cottrell

Despite progress in human motion capture, existing multi-view methods often face challenges in estimating the 3D pose and shape of multiple closely interacting people. This difficulty arises from reliance on accurate 2D joint estimations,…

Computer Vision and Pattern Recognition · Computer Science 2024-08-21 Feichi Lu , Zijian Dong , Jie Song , Otmar Hilliges

The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing the users idiosyncrasies and detailed facial expressions. However, current 3D head avatar…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Jalees Nehvi , Timo Bolkart , Thabo Beeler , Justus Thies

Most vision-language systems are static observers: they describe pixels, do not act, and cannot safely improve under shift. This passivity limits generalizable, physically grounded visual intelligence. Learning through action, not static…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Yunpeng Zhou

Building photorealistic, animatable full-body digital humans remains a longstanding challenge in computer graphics and vision. Recent advances in animatable avatar modeling have largely progressed along two directions: improving the…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Heming Zhu , Guoxing Sun , Marc Habermann

We present NBAvatar - a method for realistic rendering of head avatars handling non-rigid deformations caused by hand-face interaction. We introduce a novel representation for animated avatars by combining the training of oriented planar…

Computer Vision and Pattern Recognition · Computer Science 2026-03-13 David Svitov , Mahtab Dahaghin

Pixel-based language models have emerged as a compelling alternative to subword-based language modelling, particularly because they can represent virtually any script. PIXEL, a canonical example of such a model, is a vision transformer that…

Computation and Language · Computer Science 2024-10-17 Kushal Tatariya , Vladimir Araujo , Thomas Bauwens , Miryam de Lhoneux

Recent works based on convolutional encoder-decoder architecture and 3DMM parameterization have shown great potential for canonical view reconstruction from a single input image. Conventional CNN architectures benefit from exploiting the…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Zhiqian Lin , Jiangke Lin , Lincheng Li , Yi Yuan , Zhengxia Zou

Face inpainting is important in various applications, such as photo restoration, image editing, and virtual reality. Despite the significant advances in face generative models, ensuring that a person's unique facial identity is maintained…

Computer Vision and Pattern Recognition · Computer Science 2023-12-07 Jianjin Xu , Saman Motamed , Praneetha Vaddamanu , Chen Henry Wu , Christian Haene , Jean-Charles Bazin , Fernando de la Torre

We propose a novel framework for decomposing arbitrarily posed humans into animatable multi-layered 3D human avatars, separating the body and garments. Conventional single-layer reconstruction methods lock clothing to one identity, while…

Computer Vision and Pattern Recognition · Computer Science 2026-01-12 Yinghan Xu , John Dingliana