中文
相关论文

相关论文: GASP: Gaussian Avatars with Synthetic Priors

200 篇论文

Creating realistic avatars from a single RGB image is an attractive yet challenging problem. Due to its ill-posed nature, recent works leverage powerful prior from 2D diffusion models pretrained on large datasets. Although 2D diffusion…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Yuxuan Xue , Xianghui Xie , Riccardo Marin , Gerard Pons-Moll

Modeling animatable human avatars from monocular or multi-view videos has been widely studied, with recent approaches leveraging neural radiance fields (NeRFs) or 3D Gaussian Splatting (3DGS) achieving impressive results in novel-view and…

计算机视觉与模式识别 · 计算机科学 2025-04-03 Yahui Li , Zhi Zeng , Liming Pang , Guixuan Zhang , Shuwu Zhang

Gaussian-based human avatars have achieved an unprecedented level of visual fidelity. However, existing approaches based on high-capacity neural networks typically require a desktop GPU to achieve real-time performance for a single avatar,…

Constructing vivid 3D head avatars for given subjects and realizing a series of animations on them is valuable yet challenging. This paper presents GaussianHead, which models the actional human head with anisotropic 3D Gaussians. In our…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Jie Wang , Jiu-Cheng Xie , Xianyan Li , Feng Xu , Chi-Man Pun , Hao Gao

Although neural rendering has made significant advances in creating lifelike, animatable full-body and head avatars, incorporating detailed expressions into full-body avatars remains largely unexplored. We present DEGAS, the first 3D…

计算机视觉与模式识别 · 计算机科学 2025-02-11 Zhijing Shao , Duotun Wang , Qing-Yao Tian , Yao-Dong Yang , Hengyu Meng , Zeyu Cai , Bo Dong , Yu Zhang , Kang Zhang , Zeyu Wang

In practical real-time XR and telepresence applications, network and computing resources fluctuate frequently. Therefore, a progressive 3D representation is needed. To this end, we propose ProgressiveAvatars, a progressive avatar…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Kaiwen Song , Jinkai Cui , Juyong Zhang

3D Gaussian splatting (3DGS) is an innovative rendering technique that surpasses the neural radiance field (NeRF) in both rendering speed and visual quality by leveraging an explicit 3D scene representation. Existing 3DGS approaches require…

计算机视觉与模式识别 · 计算机科学 2025-06-13 Lintao Xiang , Hongpei Zheng , Yating Huang , Qijun Yang , Hujun Yin

Sparse Multi-view Images can be Learned to predict explicit radiance fields via Generalizable Gaussian Splatting approaches, which can achieve wider application prospects in real-life when ground-truth camera parameters are not required as…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Yanyan Li , Yixin Fang , Federico Tombari , Gim Hee Lee

The ability to animate photo-realistic head avatars reconstructed from monocular portrait video sequences represents a crucial step in bridging the gap between the virtual and real worlds. Recent advancements in head avatar techniques,…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Yufan Chen , Lizhen Wang , Qijing Li , Hongjiang Xiao , Shengping Zhang , Hongxun Yao , Yebin Liu

Talking Head Generation aims at synthesizing natural-looking talking videos from speech and a single portrait image. Previous 3D talking head generation methods have relied on domain-specific heuristics such as warping-based facial motion…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Tong Shi , Melonie de Almeida , Daniela Ivanova , Nicolas Pugeault , Paul Henderson

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the complex head poses and…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Kangjie Chen , Minghan Qin , Yu Li , Haoqian Wang

We present Gaussian Pixel Codec Avatars (GPiCA), photorealistic head avatars that can be generated from multi-view images and efficiently rendered on mobile devices. GPiCA utilizes a unique hybrid representation that combines a triangle…

计算机视觉与模式识别 · 计算机科学 2025-12-18 Divam Gupta , Anuj Pahuja , Nemanja Bartolovic , Tomas Simon , Forrest Iandola , Giljoo Nam

Avatar modelling has broad applications in human animation and virtual try-ons. Recent advancements in this field have focused on high-quality and comprehensive human reconstruction but often overlook the separation of clothing from the…

计算机视觉与模式识别 · 计算机科学 2024-11-18 Jingxuan Chen

3D Gaussian Splatting (3DGS) has made significant strides in novel view synthesis but is limited by the substantial number of Gaussian primitives required, posing challenges for deployment on lightweight devices. Recent methods address this…

计算机视觉与模式识别 · 计算机科学 2025-04-09 Zhengqing Gao , Dongting Hu , Jia-Wang Bian , Huan Fu , Yan Li , Tongliang Liu , Mingming Gong , Kun Zhang

Implicit Neural Representations (INR) have been successfully employed for Arbitrary-scale Super-Resolution (ASR). However, INR-based models need to query the multi-layer perceptron module numerous times and render a pixel in each query,…

图像与视频处理 · 电气工程与系统科学 2025-07-31 Du Chen , Liyi Chen , Zhengqiang Zhang , Lei Zhang

Despite significant progress in 3D avatar reconstruction, it still faces challenges such as high time complexity, sensitivity to data quality, and low data utilization. We propose FastAvatar, a feedforward 3D avatar framework capable of…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Yue Wu , Xuanhong Chen , Yufan Wu , Wen Li , Yuxi Lu , Kairui Feng

Existing dynamic scene reconstruction methods based on Gaussian Splatting enable real-time rendering and generate realistic images. However, adjusting the camera's focal length or the distance between Gaussian primitives and the camera to…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Zilong Chen , Huan-ang Gao , Delin Qu , Haohan Chi , Hao Tang , Kai Zhang , Hao Zhao

Gaussian splatting (GS) along with its extensions and variants provides outstanding performance in real-time scene rendering while meeting reduced storage demands and computational efficiency. While the selection of 2D images capturing the…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Konstantinos D. Polyzos , Athanasios Bacharis , Saketh Madhuvarasu , Nikos Papanikolopoulos , Tara Javidi

The fidelity of relighting is bounded by both geometry and appearance representations. For geometry, both mesh and volumetric approaches have difficulty modeling intricate structures like 3D hair geometry. For appearance, existing…

图形学 · 计算机科学 2024-05-29 Shunsuke Saito , Gabriel Schwartz , Tomas Simon , Junxuan Li , Giljoo Nam

We present a novel, zero-shot pipeline for creating hyperrealistic, identity-preserving 3D avatars from a few unstructured phone images. Existing methods face several challenges: single-view approaches suffer from geometric inconsistencies…