中文
相关论文

相关论文: Monocular and Generalizable Gaussian Talking Head …

200 篇论文

Photorealistic avatars have become essential for immersive applications in virtual reality (VR) and augmented reality (AR), enabling lifelike interactions in areas such as training simulations, telemedicine, and virtual collaboration. These…

图形学 · 计算机科学 2025-04-18 Rendong Zhang , Alexandra Watkins , Nilanjan Sarkar

Reconstructing photorealistic and topology-aware human avatars from monocular videos remains a significant challenge in the fields of computer vision and graphics. While existing 3D human avatar modeling approaches can effectively capture…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Yuze Su , Hongsong Wang , Jie Gui , Liang Wang

Reconstructing photo-realistic and topology-aware animatable human avatars from monocular videos remains challenging in computer vision and graphics. Recently, methods using 3D Gaussians to represent the human body have emerged, offering…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Haoyu Zhao , Chen Yang , Hao Wang , Xingyue Zhao , Wei Shen

Gaussian splatting and single-view depth estimation are typically studied in isolation. In this paper, we present DepthSplat to connect Gaussian splatting and depth estimation and study their interactions. More specifically, we first…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Haofei Xu , Songyou Peng , Fangjinhua Wang , Hermann Blum , Daniel Barath , Andreas Geiger , Marc Pollefeys

In this paper, we propose to create animatable avatars for interacting hands with 3D Gaussian Splatting (GS) and single-image inputs. Existing GS-based methods designed for single subjects often yield unsatisfactory results due to limited…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Xuan Huang , Hanhui Li , Wanquan Liu , Xiaodan Liang , Yiqiang Yan , Yuhao Cheng , Chengqiang Gao

Surface reconstruction has been widely studied in computer vision and graphics. However, existing surface reconstruction works struggle to recover accurate scene geometry when the input views are extremely sparse. To address this issue, we…

图形学 · 计算机科学 2025-11-26 Hanzhi Chang , Ruijie Zhu , Wenjie Chang , Mulin Yu , Yanzhe Liang , Jiahao Lu , Zhuoyuan Li , Tianzhu Zhang

Recent advances in neural radiance fields enable novel view synthesis of photo-realistic images in dynamic settings, which can be applied to scenarios with human animation. Commonly used implicit backbones to establish accurate models,…

计算机视觉与模式识别 · 计算机科学 2023-12-27 HyunJun Jung , Nikolas Brasch , Jifei Song , Eduardo Perez-Pellitero , Yiren Zhou , Zhihao Li , Nassir Navab , Benjamin Busam

Radiance fields have demonstrated impressive performance in synthesizing lifelike 3D talking heads. However, due to the difficulty in fitting steep appearance changes, the prevailing paradigm that presents facial motions by directly…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Jiahe Li , Jiawei Zhang , Xiao Bai , Jin Zheng , Xin Ning , Jun Zhou , Lin Gu

While recent 3D head avatar creation methods attempt to animate facial dynamics, they often fail to capture personalized details, limiting realism and expressiveness. To fill this gap, we present DipGuava (Disentangled and Personalized…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Jeonghaeng Lee , Seok Keun Choi , Zhixuan Li , Weisi Lin , Sanghoon Lee

We present MoGA, a novel method to reconstruct high-fidelity 3D Gaussian avatars from a single-view image. The main challenge lies in inferring unseen appearance and geometric details while ensuring 3D consistency and realism. Most previous…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Zijian Dong , Longteng Duan , Jie Song , Michael J. Black , Andreas Geiger

Recently,3DGaussianSplattinghasshowngreatpotentialin visual Simultaneous Localization And Mapping (SLAM). Existing methods have achieved encouraging results on RGB-D SLAM, but studies of the monocular case are still scarce. Moreover, they…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Tian Lan , Qinwei Lin , Haoqian Wang

Speech-driven 3D facial animation has improved a lot recently while most related works only utilize acoustic modality and neglect the influence of visual and textual cues, leading to unsatisfactory results in terms of precision and…

计算机视觉与模式识别 · 计算机科学 2023-12-06 Tianshun Han , Shengnan Gui , Yiqing Huang , Baihui Li , Lijian Liu , Benjia Zhou , Ning Jiang , Quan Lu , Ruicong Zhi , Yanyan Liang , Du Zhang , Jun Wan

We present the first application of 3D Gaussian Splatting in monocular SLAM, the most fundamental but the hardest setup for Visual SLAM. Our method, which runs live at 3fps, utilises Gaussians as the only 3D representation, unifying the…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Hidenobu Matsuki , Riku Murai , Paul H. J. Kelly , Andrew J. Davison

We present MVSGaussian, a new generalizable 3D Gaussian representation approach derived from Multi-View Stereo (MVS) that can efficiently reconstruct unseen scenes. Specifically, 1) we leverage MVS to encode geometry-aware Gaussian…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Tianqi Liu , Guangcong Wang , Shoukang Hu , Liao Shen , Xinyi Ye , Yuhang Zang , Zhiguo Cao , Wei Li , Ziwei Liu

We present GStalker, a 3D audio-driven talking face generation model with Gaussian Splatting for both fast training (40 minutes) and real-time rendering (125 FPS) with a 3$\sim$5 minute video for training material, in comparison with…

计算机视觉与模式识别 · 计算机科学 2024-05-01 Bo Chen , Shoukang Hu , Qi Chen , Chenpeng Du , Ran Yi , Yanmin Qian , Xie Chen

3D Gaussian Splatting (3DGS) achieves remarkable results in the field of surface reconstruction. However, when Gaussian normal vectors are aligned within the single-view projection plane, while the geometry appears reasonable in the current…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Bo Jia , Yanan Guo , Ying Chang , Benkui Zhang , Ying Xie , Kangning Du , Lin Cao

3D Gaussian Splatting (3DGS) has emerged as a powerful technique for generating photorealistic renderings of a scene in real-time. However, the volumetric nature of 3DGS limits its ability to accurately capture surface geometry. To address…

计算机视觉与模式识别 · 计算机科学 2026-05-04 Prajwal Gupta C. R. , Divyam Sheth , Jinjoo Ha , Mirela Ostrek , Justus Thies

A photorealistic and immersive human avatar experience demands capturing fine, person-specific details such as cloth and hair dynamics, subtle facial expressions, and characteristic motion patterns. Achieving this requires large,…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Michael Steiner , Zhang Chen , Alexander Richard , Vasu Agrawal , Markus Steinberger , Michael Zollhöfer

Recently, 3D Gaussian Splatting has emerged as a prominent research direction owing to its ultrarapid training speed and high-fidelity rendering capabilities. However, the unstructured and irregular nature of Gaussian point clouds poses…

计算机视觉与模式识别 · 计算机科学 2026-02-16 Xiao Ren , Yu Liu , Ning An , Jian Cheng , Xin Qiao , He Kong

A photorealistic and controllable 3D caricaturization framework for faces is introduced. We start with an intrinsic Gaussian curvature-based surface exaggeration technique, which, when coupled with texture, tends to produce over-smoothed…

图形学 · 计算机科学 2026-01-08 Eldad Matmon , Amit Bracha , Noam Rotstein , Ron Kimmel