中文
相关论文

相关论文: DevilSight: Augmenting Monocular Human Avatar Reco…

200 篇论文

We present a method for generating a full 360{\deg} orbit video around a person from a single input image. Existing methods typically adapt image-based diffusion models for multi-view synthesis, but yield inconsistent results across views…

计算机视觉与模式识别 · 计算机科学 2026-03-02 Keito Suzuki , Kunyao Chen , Lei Wang , Bang Du , Runfa Blark Li , Peng Liu , Ning Bi , Truong Nguyen

We present IntrinsicAvatar, a novel approach to recovering the intrinsic properties of clothed human avatars including geometry, albedo, material, and environment lighting from only monocular videos. Recent advancements in human-based…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Shaofei Wang , Božidar Antić , Andreas Geiger , Siyu Tang

We introduce a free-viewpoint rendering method -- HumanNeRF -- that works on a given monocular video of a human performing complex body motions, e.g. a video from YouTube. Our method enables pausing the video at any frame and rendering the…

计算机视觉与模式识别 · 计算机科学 2022-06-16 Chung-Yi Weng , Brian Curless , Pratul P. Srinivasan , Jonathan T. Barron , Ira Kemelmacher-Shlizerman

Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, precise motion analysis in sports, and natural collaborative behavior in human-robot tasks.…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Qi Xia , Peishan Cong , Ziyi Wang , Yujing Sun , Qin Sun , Xinge Zhu , Mao Ye , Ruigang Yang , Yuexin Ma

We propose a novel approach for reconstructing animatable 3D Gaussian avatars from monocular videos captured by commodity devices like smartphones. Photorealistic 3D head avatar reconstruction from such recordings is challenging due to…

计算机视觉与模式识别 · 计算机科学 2025-04-15 Jiapeng Tang , Davide Davoli , Tobias Kirschstein , Liam Schoneveld , Matthias Niessner

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

计算机视觉与模式识别 · 计算机科学 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

Creating a photorealistic scene and human reconstruction from a single monocular in-the-wild video figures prominently in the perception of a human-centric 3D world. Recent neural rendering advances have enabled holistic human-scene…

计算机视觉与模式识别 · 计算机科学 2025-04-21 Zetong Zhang , Manuel Kaufmann , Lixin Xue , Jie Song , Martin R. Oswald

Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motion prediction, virtual reality, and digital twinning. Despite significant interest in…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Yeheng Zong , Pou-Chun Kung , Yike Pan , Seth Isaacson , Yizhou Chen , Ram Vasudevan , Katherine A. Skinner

In this paper, we propose a novel learning approach for feed-forward one-shot 4D head avatar synthesis. Different from existing methods that often learn from reconstructing monocular videos guided by 3DMM, we employ pseudo multi-view videos…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Yu Deng , Duomin Wang , Baoyuan Wang

We introduce a new method that generates photo-realistic humans under novel views and poses given a monocular video as input. Despite the significant progress recently on this topic, with several methods exploring shared canonical neural…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Tiantian Wang , Nikolaos Sarafianos , Ming-Hsuan Yang , Tony Tung

The ability to animate photo-realistic head avatars reconstructed from monocular portrait video sequences represents a crucial step in bridging the gap between the virtual and real worlds. Recent advancements in head avatar techniques,…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Yufan Chen , Lizhen Wang , Qijing Li , Hongjiang Xiao , Shengping Zhang , Hongxun Yao , Yebin Liu

This paper tackles the challenge of creating relightable and animatable neural avatars from sparse-view (or even monocular) videos of dynamic humans under unknown illumination. Compared to studio environments, this setting is more practical…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Zhen Xu , Sida Peng , Chen Geng , Linzhan Mou , Zihan Yan , Jiaming Sun , Hujun Bao , Xiaowei Zhou

In this paper, we present a novel method that facilitates the creation of vivid 3D Gaussian avatars from monocular video inputs (GVA). Our innovation lies in addressing the intricate challenges of delivering high-fidelity human body…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Xinqi Liu , Chenming Wu , Jialun Liu , Xing Liu , Jinbo Wu , Chen Zhao , Haocheng Feng , Errui Ding , Jingdong Wang

Recovering 4D human-object interaction (HOI) from monocular video is a key step toward scalable 3D content creation, embodied AI, and simulation-based learning. Recent methods can reconstruct temporally coherent human and object…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Yubo Zhao , Yujin Chai , Yunao Dong , Chengfeng Zhao , Zijiao Zeng , Yuan Liu , Chi-Keung Tang

We present GaussianAvatar, an efficient approach to creating realistic human avatars with dynamic 3D appearances from a single video. We start by introducing animatable 3D Gaussians to explicitly represent humans in various poses and…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Liangxiao Hu , Hongwen Zhang , Yuxiang Zhang , Boyao Zhou , Boning Liu , Shengping Zhang , Liqiang Nie

Head avatar reenactment focuses on creating animatable personal avatars from monocular videos, serving as a foundational element for applications like social signal understanding, gaming, human-machine interaction, and computer vision.…

计算机视觉与模式识别 · 计算机科学 2026-01-28 Wei Liang , Hui Yu , Derui Ding , Rachael E. Jack , Philippe G. Schyns

The efficient reconstruction of high-quality and intuitively editable human avatars presents a pressing challenge in the field of computer vision. Recent advancements, such as 3DGS, have demonstrated impressive reconstruction efficiency and…

图形学 · 计算机科学 2025-11-25 Mengtian Li , Shengxiang Yao , Yichen Pan , Haiyao Xiao , Zhongmei Li , Zhifeng Xie , Keyu Chen

The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing the users idiosyncrasies and detailed facial expressions. However, current 3D head avatar…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Jalees Nehvi , Timo Bolkart , Thabo Beeler , Justus Thies

Human-motion video generation has been a challenging task, primarily due to the difficulty inherent in learning human body movements. While some approaches have attempted to drive human-centric video generation explicitly through pose…

计算机视觉与模式识别 · 计算机科学 2025-04-02 Boyuan Wang , Xiaofeng Wang , Chaojun Ni , Guosheng Zhao , Zhiqin Yang , Zheng Zhu , Muyang Zhang , Yukun Zhou , Xinze Chen , Guan Huang , Lihong Liu , Xingang Wang

Traditional methods for constructing high-quality, personalized head avatars from monocular videos demand extensive face captures and training time, posing a significant challenge for scalability. This paper introduces a novel approach to…

计算机视觉与模式识别 · 计算机科学 2024-02-20 Zhixuan Yu , Ziqian Bai , Abhimitra Meka , Feitong Tan , Qiangeng Xu , Rohit Pandey , Sean Fanello , Hyun Soo Park , Yinda Zhang