English
Related papers

Related papers: MeshLAM: Feed-Forward One-Shot Animatable Textured…

200 papers

Visual SLAM (Simultaneous Localization and Mapping) methods typically rely on handcrafted visual features or raw RGB values for establishing correspondences between images. These features, while suitable for sparse mapping, often lead to…

Computer Vision and Pattern Recognition · Computer Science 2018-11-21 Chamara Saroj Weerasekera , Ravi Garg , Yasir Latif , Ian Reid

Despite recent progress in 3D Gaussian-based head avatar modeling, efficiently generating high fidelity avatars remains a challenge. Current methods typically rely on extensive multi-view capture setups or monocular videos with per-identity…

Computer Vision and Pattern Recognition · Computer Science 2026-02-02 Xinya Ji , Sebastian Weiss , Manuel Kansy , Jacek Naruniec , Xun Cao , Barbara Solenthaler , Derek Bradley

Feedforward reconstruction is crucial for autonomous driving applications, where rapid scene reconstruction enables efficient utilization of large-scale driving datasets in closed-loop simulation and other downstream tasks, eliminating the…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Zhongrui Yu , Zhao Wang , Yijia Xie , Yida Wang , Xueyang Zhang , Yifei Zhan , Kun Zhan

Mesh reconstruction is a cornerstone process across various applications, including in-silico trials, digital twins, surgical planning, and navigation. Recent advancements in deep learning have notably enhanced mesh reconstruction speeds.…

Image and Video Processing · Electrical Eng. & Systems 2025-05-22 Fengting Zhang , Boxu Liang , Qinghao Liu , Min Liu , Xiang Chen , Yaonan Wang

We present an approach for the reconstruction of textured 3D meshes of human heads from one or few views. Since such few-shot reconstruction is underconstrained, it requires prior knowledge which is hard to impose on traditional 3D…

Computer Vision and Pattern Recognition · Computer Science 2023-09-12 Egor Burkov , Ruslan Rakhimov , Aleksandr Safin , Evgeny Burnaev , Victor Lempitsky

In this paper, we introduce the Volumetric Relightable Morphable Model (VRMM), a novel volumetric and parametric facial prior for 3D face modeling. While recent volumetric prior models offer improvements over traditional methods like 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-05-09 Haotian Yang , Mingwu Zheng , Chongyang Ma , Yu-Kun Lai , Pengfei Wan , Haibin Huang

In this paper, we introduce OneReward, a unified reinforcement learning framework that enhances the model's generative capabilities across multiple tasks under different evaluation criteria using only \textit{One Reward} model. By employing…

Computer Vision and Pattern Recognition · Computer Science 2025-08-29 Yuan Gong , Xionghui Wang , Jie Wu , Shiyin Wang , Yitong Wang , Xinglong Wu

Despite much progress, achieving real-time high-fidelity head avatar animation is still difficult and existing methods have to trade-off between speed and quality. 3DMM based methods often fail to model non-facial structures such as…

Graphics · Computer Science 2024-06-25 Zhongyuan Zhao , Zhenyu Bao , Qing Li , Guoping Qiu , Kanglin Liu

The goal of face reenactment is to transfer a target expression and head pose to a source face while preserving the source identity. With the popularity of face-related applications, there has been much research on this topic. However, the…

Computer Vision and Pattern Recognition · Computer Science 2022-05-27 Wonjun Kang , Geonsu Lee , Hyung Il Koo , Nam Ik Cho

We have recently seen great progress in building photorealistic animatable full-body codec avatars, but generating high-fidelity animation of clothing is still difficult. To address these difficulties, we propose a method to build an…

Computer Vision and Pattern Recognition · Computer Science 2021-10-07 Donglai Xiang , Fabian Prada , Timur Bagautdinov , Weipeng Xu , Yuan Dong , He Wen , Jessica Hodgins , Chenglei Wu

Recent advancements in 3D avatar generation excel with multi-view supervision for photorealistic models. However, monocular counterparts lag in quality despite broader applicability. We propose ReCaLaB to close this gap. ReCaLaB is a…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Yuchen Rao , Eduardo Perez Pellitero , Benjamin Busam , Yiren Zhou , Jifei Song

Animatable head avatar generation typically requires extensive data for training. To reduce the data requirements, a natural solution is to leverage existing data-free static avatar generation methods, such as pre-trained diffusion models…

Computer Vision and Pattern Recognition · Computer Science 2025-03-26 Zhenglin Zhou , Fan Ma , Hehe Fan , Tat-Seng Chua

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

Automated construction of surface geometries of cardiac structures from volumetric medical images is important for a number of clinical applications. While deep-learning-based approaches have demonstrated promising reconstruction precision,…

Image and Video Processing · Electrical Eng. & Systems 2021-09-15 Fanwei Kong , Nathan Wilson , Shawn C. Shadden

We present a novel framework for reconstructing animatable human avatars from multiple images, termed CanonicalFusion. Our central concept involves integrating individual reconstruction results into the canonical space. To be specific, we…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Jisu Shin , Junmyeong Lee , Seongmin Lee , Min-Gyu Park , Ju-Mi Kang , Ju Hong Yoon , Hae-Gon Jeon

Recent advances in generative diffusion models have enabled the previously unfeasible capability of generating 3D assets from a single input image or a text prompt. In this work, we aim to enhance the quality and functionality of these…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Xiyi Chen , Marko Mihajlovic , Shaofei Wang , Sergey Prokudin , Siyu Tang

To address the ill-posed problem caused by partial observations in monocular human volumetric capture, we present AvatarCap, a novel framework that introduces animatable avatars into the capture pipeline for high-fidelity reconstruction in…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Zhe Li , Zerong Zheng , Hongwen Zhang , Chaonan Ji , Yebin Liu

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the complex head poses and…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Kangjie Chen , Minghan Qin , Yu Li , Haoqian Wang

Traditional 3D morphable face models (3DMMs) provide fine-grained control over expression but cannot easily capture geometric and appearance details. Neural volumetric representations approach photorealism but are hard to animate and do not…

Computer Vision and Pattern Recognition · Computer Science 2022-11-07 Yufeng Zheng , Victoria Fernández Abrevaya , Marcel C. Bühler , Xu Chen , Michael J. Black , Otmar Hilliges

Speech-driven three-dimensional (3D) facial animation synthesis aims to build a mapping from one-dimensional (1D) speech signals to time-varying 3D facial motion signals. Current methods still face challenges in maintaining lip-sync…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Bin Liu , Zhixiang Xiong , Zhifen He , Bo Li