中文
相关论文

相关论文: Generalizable and Animatable 3D Full-Head Gaussian…

200 篇论文

This paper proposes an efficient 3D avatar coding framework that leverages compact human priors and canonical-to-target transformation to enable high-quality 3D human avatar video compression at ultra-low bit rates. The framework begins by…

图像与视频处理 · 电气工程与系统科学 2025-10-14 Shanzhi Yin , Bolin Chen , Xinju Wu , Ru-Ling Liao , Jie Chen , Shiqi Wang , Yan Ye

Reconstructing personalized animatable head avatars has significant implications in the fields of AR/VR. Existing methods for achieving explicit face control of 3D Morphable Models (3DMM) typically rely on multi-view images or videos of a…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Haoyu Ma , Tong Zhang , Shanlin Sun , Xiangyi Yan , Kun Han , Xiaohui Xie

Building realistic and animatable avatars still requires minutes of multi-view or monocular self-rotating videos, and most methods lack precise control over gestures and expressions. To push this boundary, we address the challenge of…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Jun Xiang , Yudong Guo , Leipeng Hu , Boyang Guo , Yancheng Yuan , Juyong Zhang

Real-time rendering of high-fidelity and animatable avatars from monocular videos remains a challenging problem in computer vision and graphics. Over the past few years, the Neural Radiance Field (NeRF) has made significant progress in…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Qipeng Yan , Mingyang Sun , Lihua Zhang

We present UIKA, a feed-forward animatable Gaussian head model from an arbitrary number of pose-free inputs, including a single image, multi-view captures, and smartphone-captured videos. Unlike the traditional avatar method, which requires…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Zijian Wu , Boyao Zhou , Liangxiao Hu , Hongyu Liu , Yuan Sun , Xuan Wang , Xun Cao , Yujun Shen , Hao Zhu

To address the ill-posed problem caused by partial observations in monocular human volumetric capture, we present AvatarCap, a novel framework that introduces animatable avatars into the capture pipeline for high-fidelity reconstruction in…

计算机视觉与模式识别 · 计算机科学 2022-07-13 Zhe Li , Zerong Zheng , Hongwen Zhang , Chaonan Ji , Yebin Liu

Previous animatable 3D-aware GANs for human generation have primarily focused on either the human head or full body. However, head-only videos are relatively uncommon in real life, and full body generation typically does not deal with…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Yue Wu , Sicheng Xu , Jianfeng Xiang , Fangyun Wei , Qifeng Chen , Jiaolong Yang , Xin Tong

Faithful real-time facial animation is essential for avatar-mediated telepresence in Virtual Reality (VR). To emulate authentic communication, avatar animation needs to be efficient and accurate: able to capture both extreme and subtle…

计算机视觉与模式识别 · 计算机科学 2024-07-19 Shaojie Bai , Te-Li Wang , Chenghui Li , Akshay Venkatesh , Tomas Simon , Chen Cao , Gabriel Schwartz , Ryan Wrench , Jason Saragih , Yaser Sheikh , Shih-En Wei

In this paper, we propose a novel approach to reconstruct 3D human body shapes based on a sparse set of RGBD frames using a single RGBD camera. We specifically focus on the realistic settings where human subjects move freely during the…

计算机视觉与模式识别 · 计算机科学 2020-06-16 Xinxin Zuo , Sen Wang , Jiangbin Zheng , Weiwei Yu , Minglun Gong , Ruigang Yang , Li Cheng

3D Gaussian Splatting (3DGS) has enabled photorealistic and real-time rendering of 3D head avatars. Existing 3DGS-based avatars typically rely on tens of thousands of 3D Gaussian points (Gaussians), with the number of Gaussians fixed after…

计算机视觉与模式识别 · 计算机科学 2025-10-08 Peizhi Yan , Rabab Ward , Qiang Tang , Shan Du

We present a unified and generalizable framework for synthesizing view-consistent and temporally coherent avatars from a single image, addressing the challenging task of single-image avatar generation. Existing diffusion-based methods often…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Yixing Lu , Junting Dong , Youngjoong Kwon , Qin Zhao , Bo Dai , Fernando De la Torre

Photorealistic avatars have become essential for immersive applications in virtual reality (VR) and augmented reality (AR), enabling lifelike interactions in areas such as training simulations, telemedicine, and virtual collaboration. These…

图形学 · 计算机科学 2025-04-18 Rendong Zhang , Alexandra Watkins , Nilanjan Sarkar

In this paper, we propose to create animatable avatars for interacting hands with 3D Gaussian Splatting (GS) and single-image inputs. Existing GS-based methods designed for single subjects often yield unsatisfactory results due to limited…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Xuan Huang , Hanhui Li , Wanquan Liu , Xiaodan Liang , Yiqiang Yan , Yuhao Cheng , Chengqiang Gao

3D modeling has long been an important area in computer vision and computer graphics. Recently, thanks to the breakthroughs in neural representations and generative models, we witnessed a rapid development of 3D modeling. 3D human modeling,…

计算机视觉与模式识别 · 计算机科学 2024-06-07 Ruihe Wang , Yukang Cao , Kai Han , Kwan-Yee K. Wong

We present DreamHuman, a method to generate realistic animatable 3D human avatar models solely from textual descriptions. Recent text-to-3D methods have made considerable strides in generation, but are still lacking in important aspects.…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Nikos Kolotouros , Thiemo Alldieck , Andrei Zanfir , Eduard Gabriel Bazavan , Mihai Fieraru , Cristian Sminchisescu

In this paper, we introduce GaussianMotion, a novel human rendering model that generates fully animatable scenes aligned with textual descriptions using Gaussian Splatting. Although existing methods achieve reasonable text-to-3D generation…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Gyumin Shim , Sangmin Lee , Jaegul Choo

We present FaceLift, a novel feed-forward approach for generalizable high-quality 360-degree 3D head reconstruction from a single image. Our pipeline first employs a multi-view latent diffusion model to generate consistent side and back…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Weijie Lyu , Yi Zhou , Ming-Hsuan Yang , Zhixin Shu

Modern 3D-GANs synthesize geometry and texture by training on large-scale datasets with a consistent structure. Training such models on stylized, artistic data, with often unknown, highly variable geometry, and camera information has not…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Rameen Abdal , Hsin-Ying Lee , Peihao Zhu , Menglei Chai , Aliaksandr Siarohin , Peter Wonka , Sergey Tulyakov

We introduce MeshLAM, a feed-forward framework for one-shot animatable mesh head reconstruction that generates high-fidelity, animatable 3D head avatars from a single image. Unlike previous work that relies on time-consuming test-time…

计算机视觉与模式识别 · 计算机科学 2026-04-28 Yisheng He , Steven Hoi

The ability to generate diverse 3D articulated head avatars is vital to a plethora of applications, including augmented reality, cinematography, and education. Recent work on text-guided 3D object generation has shown great promise in…

计算机视觉与模式识别 · 计算机科学 2023-07-12 Alexander W. Bergman , Wang Yifan , Gordon Wetzstein