中文
相关论文

相关论文: AvatarCap: Animatable Avatar Conditioned Monocular…

200 篇论文

Reconstructing an avatar from a portrait image has many applications in multimedia, but remains a challenging research problem. Extracting reflectance maps and geometry from one image is ill-posed: recovering geometry is a one-to-many…

计算机视觉与模式识别 · 计算机科学 2023-12-25 Abdallah Dib , Luiz Gustavo Hafemann , Emeline Got , Trevor Anderson , Amin Fadaeinejad , Rafael M. O. Cruz , Marc-Andre Carbonneau

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

We present a novel method for high detail-preserving human avatar creation from monocular video. A parameterized body model is refined and optimized to maximally resemble subjects from a video showing them from all sides. Our avatars…

计算机视觉与模式识别 · 计算机科学 2018-08-07 Thiemo Alldieck , Marcus Magnor , Weipeng Xu , Christian Theobalt , Gerard Pons-Moll

Reconstructing high-fidelity and animatable 3D head avatars from monocular videos remains a challenging yet essential task. Existing methods based on 3D Gaussian Splatting typically bind Gaussians to mesh triangles and model deformations…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Jiankuo Zhao , Xiangyu Zhu , Zidu Wang , Zhen Lei

We introduce FlexAvatar, a method for creating high-quality and complete 3D head avatars from a single image. A core challenge lies in the limited availability of multi-view data and the tendency of monocular training to yield incomplete 3D…

计算机视觉与模式识别 · 计算机科学 2026-04-09 Tobias Kirschstein , Simon Giebenhain , Matthias Nießner

We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is prevalent in real-world dynamic video capture, especially due to human movements in 3D human…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Xianrui Luo , Juewen Peng , Zhongang Cai , Lei Yang , Fan Yang , Zhiguo Cao , Guosheng Lin

With NeRF widely used for facial reenactment, recent methods can recover photo-realistic 3D head avatar from just a monocular video. Unfortunately, the training process of the NeRF-based methods is quite time-consuming, as MLP used in the…

计算机视觉与模式识别 · 计算机科学 2023-05-04 Yuelang Xu , Lizhen Wang , Xiaochen Zhao , Hongwen Zhang , Yebin Liu

In this paper, we propose ARCH (Animatable Reconstruction of Clothed Humans), a novel end-to-end framework for accurate reconstruction of animation-ready 3D clothed humans from a monocular image. Existing approaches to digitize 3D humans…

图形学 · 计算机科学 2020-04-14 Zeng Huang , Yuanlu Xu , Christoph Lassner , Hao Li , Tony Tung

We present Better Together, a method that simultaneously solves the human pose estimation problem while reconstructing a photorealistic 3D human avatar from multi-view videos. While prior art usually solves these problems separately, we…

计算机视觉与模式识别 · 计算机科学 2025-03-13 Arthur Moreau , Mohammed Brahimi , Richard Shaw , Athanasios Papaioannou , Thomas Tanay , Zhensong Zhang , Eduardo Pérez-Pellitero

We present a convolutional autoencoder that enables high fidelity volumetric reconstructions of human performance to be captured from multi-view video comprising only a small set of camera views. Our method yields similar end-to-end…

计算机视觉与模式识别 · 计算机科学 2018-07-11 Andrew Gilbert , Marco Volino , John Collomosse , Adrian Hilton

This paper proposes GraviCap, i.e., a new approach for joint markerless 3D human motion capture and object trajectory estimation from monocular RGB videos. We focus on scenes with objects partially observed during a free flight. In contrast…

计算机视觉与模式识别 · 计算机科学 2021-08-20 Rishabh Dabral , Soshi Shimada , Arjun Jain , Christian Theobalt , Vladislav Golyanik

Codec Avatars are a recent class of learned, photorealistic face models that accurately represent the geometry and texture of a person in 3D (i.e., for virtual reality), and are almost indistinguishable from video. In this paper we describe…

计算机视觉与模式识别 · 计算机科学 2020-08-13 Alexander Richard , Colin Lea , Shugao Ma , Juergen Gall , Fernando de la Torre , Yaser Sheikh

We present a method to build animatable dog avatars from monocular videos. This is challenging as animals display a range of (unpredictable) non-rigid movements and have a variety of appearance details (e.g., fur, spots, tails). We develop…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Remy Sabathier , Niloy J. Mitra , David Novotny

High-quality, animatable 3D human avatar reconstruction from monocular videos offers significant potential for reducing reliance on complex hardware, making it highly practical for applications in game development, augmented reality, and…

计算机视觉与模式识别 · 计算机科学 2025-05-02 Xia Yuan , Hai Yuan , Wenyi Ge , Ying Fu , Xi Wu , Guanyu Xing

Despite progress in human motion capture, existing multi-view methods often face challenges in estimating the 3D pose and shape of multiple closely interacting people. This difficulty arises from reliance on accurate 2D joint estimations,…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Feichi Lu , Zijian Dong , Jie Song , Otmar Hilliges

We propose a novel approach for reconstructing animatable 3D Gaussian avatars from monocular videos captured by commodity devices like smartphones. Photorealistic 3D head avatar reconstruction from such recordings is challenging due to…

计算机视觉与模式识别 · 计算机科学 2025-04-15 Jiapeng Tang , Davide Davoli , Tobias Kirschstein , Liam Schoneveld , Matthias Niessner

3D human motion capture from monocular RGB images respecting interactions of a subject with complex and possibly deformable environments is a very challenging, ill-posed and under-explored problem. Existing methods address it only weakly…

计算机视觉与模式识别 · 计算机科学 2022-08-18 Zhi Li , Soshi Shimada , Bernt Schiele , Christian Theobalt , Vladislav Golyanik

Modeling animatable human avatars from monocular or multi-view videos has been widely studied, with recent approaches leveraging neural radiance fields (NeRFs) or 3D Gaussian Splatting (3DGS) achieving impressive results in novel-view and…

计算机视觉与模式识别 · 计算机科学 2025-04-03 Yahui Li , Zhi Zeng , Liming Pang , Guixuan Zhang , Shuwu Zhang

Building realistic and animatable avatars still requires minutes of multi-view or monocular self-rotating videos, and most methods lack precise control over gestures and expressions. To push this boundary, we address the challenge of…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Jun Xiang , Yudong Guo , Leipeng Hu , Boyang Guo , Yancheng Yuan , Juyong Zhang

We present Vid2Avatar-Pro, a method to create photorealistic and animatable 3D human avatars from monocular in-the-wild videos. Building a high-quality avatar that supports animation with diverse poses from a monocular video is challenging…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Chen Guo , Junxuan Li , Yash Kant , Yaser Sheikh , Shunsuke Saito , Chen Cao