中文
相关论文

相关论文: ELITE: Efficient Gaussian Head Avatar from a Monoc…

200 篇论文

Recent advances in 3D Gaussian Splatting (3DGS) have enabled fast, photorealistic rendering of dynamic 3D scenes, showing strong potential in immersive communication. However, in digital human encoding and transmission, the compression…

计算机视觉与模式识别 · 计算机科学 2025-10-21 Haocheng Tang , Ruoke Yan , Xinhui Yin , Qi Zhang , Xinfeng Zhang , Siwei Ma , Wen Gao , Chuanmin Jia

3D Gaussian splatting-based talking head synthesis has recently gained attention for its ability to render high-fidelity images with real-time inference speed. However, since it is typically trained on only a short video that lacks the…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Junuk Cha , Seongro Yoon , Valeriya Strizhkova , Francois Bremond , Seungryul Baek

Recent studies have combined 3D Gaussian and 3D Morphable Models (3DMM) to construct high-quality 3D head avatars. In this line of research, existing methods either fail to capture the dynamic textures or incur significant overhead in terms…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Yating Wang , Xuan Wang , Ran Yi , Yanbo Fan , Jichen Hu , Jingcheng Zhu , Lizhuang Ma

Creating realistic avatars from a single RGB image is an attractive yet challenging problem. Due to its ill-posed nature, recent works leverage powerful prior from 2D diffusion models pretrained on large datasets. Although 2D diffusion…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Yuxuan Xue , Xianghui Xie , Riccardo Marin , Gerard Pons-Moll

Reconstructing a high-quality, animatable 3D human avatar with expressive facial and hand motions from a single image has gained significant attention due to its broad application potential. 3D human avatar reconstruction typically requires…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Yang Li , Minghan Qin , Yu Li , Haoqian Wang

We introduce the first method for audio-driven universal photorealistic avatar synthesis, combining a person-agnostic speech model with our novel Universal Head Avatar Prior (UHAP). UHAP is trained on cross-identity multi-view videos. In…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Kartik Teotia , Helge Rhodin , Mohit Mendiratta , Hyeongwoo Kim , Marc Habermann , Christian Theobalt

We present PEGASUS, a method for constructing a personalized generative 3D face avatar from monocular video sources. Our generative 3D avatar enables disentangled controls to selectively alter the facial attributes (e.g., hair or nose)…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Hyunsoo Cha , Byungjun Kim , Hanbyul Joo

We propose StreamME, a method focuses on fast 3D avatar reconstruction. The StreamME synchronously records and reconstructs a head avatar from live video streams without any pre-cached data, enabling seamless integration of the…

图形学 · 计算机科学 2025-07-24 Luchuan Song , Yang Zhou , Zhan Xu , Yi Zhou , Deepali Aneja , Chenliang Xu

We present PERSE, a method for building a personalized 3D generative avatar from a reference portrait. Our avatar enables facial attribute editing in a continuous and disentangled latent space to control each facial attribute, while…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Hyunsoo Cha , Inhee Lee , Hanbyul Joo

Video matting has traditionally been limited by the lack of high-quality ground-truth data. Most existing video matting datasets provide only human-annotated imperfect alpha and foreground annotations, which must be composited to background…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Yongtao Ge , Kangyang Xie , Guangkai Xu , Mingyu Liu , Li Ke , Longtao Huang , Hui Xue , Hao Chen , Chunhua Shen

In this paper, we explore a reconstruction and reenactment separated framework for 3D Gaussians head, which requires only a single portrait image as input to generate controllable avatar. Specifically, we developed a large-scale one-shot…

计算机视觉与模式识别 · 计算机科学 2025-09-18 Zhiling Ye , Cong Zhou , Xiubao Zhang , Haifeng Shen , Weihong Deng , Quan Lu

Event cameras are bio-inspired sensors that output asynchronous and sparse event streams, instead of fixed frames. Benefiting from their distinct advantages, such as high dynamic range and high temporal resolution, event cameras have been…

计算机视觉与模式识别 · 计算机科学 2024-09-23 Zixin Zhang , Kanghao Chen , Lin Wang

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

With the advancement of virtual reality, the demand for 3D human avatars is increasing. The emergence of Gaussian Splatting technology has enabled the rendering of Gaussian avatars with superior visual quality and reduced computational…

图形学 · 计算机科学 2025-12-01 Xiaonuo Dongye , Hanzhi Guo , Le Luo , Haiyan Jiang , Yihua Bao , Jie Guo , Zeyu Tian , Dongdong Weng

We introduce Gaussian Articulated Template Model GART, an explicit, efficient, and expressive representation for non-rigid articulated subject capturing and rendering from monocular videos. GART utilizes a mixture of moving 3D Gaussians to…

计算机视觉与模式识别 · 计算机科学 2023-11-28 Jiahui Lei , Yufu Wang , Georgios Pavlakos , Lingjie Liu , Kostas Daniilidis

Creating high-quality animatable 3D human avatars from a single image remains a significant challenge in computer vision due to the inherent difficulty of reconstructing complete 3D information from a single viewpoint. Current approaches…

计算机视觉与模式识别 · 计算机科学 2025-05-09 Yonwoo Choi

We present a universal prior model for 3D head avatars with explicit hair compositionality. Existing approaches to build generalizable priors for 3D head avatars often adopt a holistic modeling approach, treating the face and hair as an…

计算机视觉与模式识别 · 计算机科学 2025-07-28 Byungjun Kim , Shunsuke Saito , Giljoo Nam , Tomas Simon , Jason Saragih , Hanbyul Joo , Junxuan Li

We present Instant Skinned Gaussian Avatars, a real-time and cross-platform 3D avatar system. Many approaches have been proposed to animate Gaussian Splatting, but they often require camera arrays, long preprocessing times, or high-end…

计算几何 · 计算机科学 2025-10-24 Naruya Kondo , Yuto Asano , Yoichi Ochiai

Many works have succeeded in reconstructing Gaussian human avatars from multi-view videos. However, they either struggle to capture pose-dependent appearance details with a single MLP, or rely on a computationally intensive neural network…

图形学 · 计算机科学 2025-04-29 Youyi Zhan , Tianjia Shao , Yin Yang , Kun Zhou

Reconstructing a complete 3D head from a single portrait remains challenging because existing methods still face a sharp quality-speed trade-off: high-fidelity pipelines often rely on multi-stage processing and per-subject optimization,…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Yujie Gao , Yao Xiao , Xiangnan Zhu , Ya Li , Yiyi Zhang , Liqing Zhang , Jianfu Zhang