中文
相关论文

相关论文: HRAvatar: High-Quality and Relightable Gaussian He…

200 篇论文

Virtual production (VP) use LED walls to provide both background imagery and image-based lighting. While this enables on-set compositing, it couples lighting to background and scene appearance, limiting flexibility for downstream editing.…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Adrian Azzarelli , Nantheera Anantrasirichai , James Pollock , David R. Bull

While high fidelity and efficiency are central to the creation of digital head avatars, recent methods relying on 2D or 3D generative models often experience limitations such as shape distortion, expression inaccuracy, and identity…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Xiaochen Zhao , Jingxiang Sun , Lizhen Wang , Jinli Suo , Yebin Liu

Speech-driven talking heads have recently emerged and enable interactive avatars. However, real-world applications are limited, as current methods achieve high visual fidelity but slow or fast yet temporally unstable. Diffusion methods…

计算机视觉与模式识别 · 计算机科学 2025-12-12 Madhav Agarwal , Mingtian Zhang , Laura Sevilla-Lara , Steven McDonagh

Novel-view synthesis and 3D reconstruction from sparse posed images are central to robotics and AR/VR. Yet, feed-forward 3D Gaussian reconstruction fails under lowlight due to noise, color shifts, and unreliable correspondence. We propose…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Fuzhen Jiang , Zengtian Xie , Zhuoran Li

We present LiftAvatar, a new paradigm that completes sparse monocular observations in kinematic space (e.g., facial expressions and head pose) and uses the completed signals to drive high-fidelity avatar animation. LiftAvatar is a…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Hualiang Wei , Shunran Jia , Jialun Liu , Wenhui Li

While there has been significant progress in the field of 3D avatar creation from visual observations, modeling physically plausible dynamics of humans with loose garments remains a challenging problem. Although a few existing works address…

图形学 · 计算机科学 2025-10-03 Changmin Lee , Jihyun Lee , Tae-Kyun Kim

Reconstructing photo-realistic and topology-aware animatable human avatars from monocular videos remains challenging in computer vision and graphics. Recently, methods using 3D Gaussians to represent the human body have emerged, offering…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Haoyu Zhao , Chen Yang , Hao Wang , Xingyue Zhao , Wei Shen

Personalized 3D avatars require an animatable representation of digital humans. Doing so instantly from monocular videos offers scalability to broad class of users and wide-scale applications. In this paper, we present a fast, simple, yet…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Pramish Paudel , Anubhav Khanal , Ajad Chhatkuli , Danda Pani Paudel , Jyoti Tandukar

Reconstructing high-fidelity, animatable 3D head avatars from effortlessly captured monocular videos is a pivotal yet formidable challenge. Although significant progress has been made in rendering performance and manipulation capabilities,…

计算机视觉与模式识别 · 计算机科学 2025-03-24 Jiawei Zhang , Zijian Wu , Zhiyang Liang , Yicheng Gong , Dongfang Hu , Yao Yao , Xun Cao , Hao Zhu

Creating high-fidelity, animatable 3D avatars from a single image remains a formidable challenge. We identified three desirable attributes of avatar generation: 1) the method should be feed-forward, 2) model a 360{\deg} full-head, and 3)…

图形学 · 计算机科学 2026-02-13 Zehao Xia , Yiqun Wang , Zhengda Lu , Kai Liu , Jun Xiao , Peter Wonka

A photorealistic and immersive human avatar experience demands capturing fine, person-specific details such as cloth and hair dynamics, subtle facial expressions, and characteristic motion patterns. Achieving this requires large,…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Michael Steiner , Zhang Chen , Alexander Richard , Vasu Agrawal , Markus Steinberger , Michael Zollhöfer

A photorealistic and controllable 3D caricaturization framework for faces is introduced. We start with an intrinsic Gaussian curvature-based surface exaggeration technique, which, when coupled with texture, tends to produce over-smoothed…

图形学 · 计算机科学 2026-01-08 Eldad Matmon , Amit Bracha , Noam Rotstein , Ron Kimmel

3D video avatars can empower virtual communications by providing compression, privacy, entertainment, and a sense of presence in AR/VR. Best 3D photo-realistic AR/VR avatars driven by video, that can minimize uncanny effects, rely on…

计算机视觉与模式识别 · 计算机科学 2021-03-31 Lele Chen , Chen Cao , Fernando De la Torre , Jason Saragih , Chenliang Xu , Yaser Sheikh

We propose VASA-3D, an audio-driven, single-shot 3D head avatar generator. This research tackles two major challenges: capturing the subtle expression details present in real human faces, and reconstructing an intricate 3D head avatar from…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Sicheng Xu , Guojun Chen , Jiaolong Yang , Yizhong Zhang , Yu Deng , Steve Lin , Baining Guo

We present Vid2Avatar, a method to learn human avatars from monocular in-the-wild videos. Reconstructing humans that move naturally from monocular in-the-wild videos is difficult. Solving it requires accurately separating humans from…

计算机视觉与模式识别 · 计算机科学 2023-02-23 Chen Guo , Tianjian Jiang , Xu Chen , Jie Song , Otmar Hilliges

Inspired by the effectiveness of 3D Gaussian Splatting (3DGS) in reconstructing detailed 3D scenes within multi-view setups and the emergence of large 2D human foundation models, we introduce Arc2Avatar, the first SDS-based method utilizing…

计算机视觉与模式识别 · 计算机科学 2025-01-14 Dimitrios Gerogiannis , Foivos Paraperas Papantoniou , Rolandos Alexandros Potamias , Alexandros Lattas , Stefanos Zafeiriou

Our goal is to efficiently learn personalized animatable 3D head avatars from videos that are geometrically accurate, realistic, relightable, and compatible with current rendering systems. While 3D meshes enable efficient processing and are…

计算机视觉与模式识别 · 计算机科学 2023-10-30 Shrisha Bharadwaj , Yufeng Zheng , Otmar Hilliges , Michael J. Black , Victoria Fernandez-Abrevaya

We present a spatial and angular Gaussian based representation and a triple splatting process, for real-time, high-quality novel lighting-and-view synthesis from multi-view point-lit input images. To describe complex appearance, we employ a…

计算机视觉与模式识别 · 计算机科学 2024-10-16 Zoubin Bi , Yixin Zeng , Chong Zeng , Fan Pei , Xiang Feng , Kun Zhou , Hongzhi Wu

Radiance field methods represent the state of the art in reconstructing complex scenes from multi-view photos. However, these reconstructions often suffer from one or both of the following limitations: First, they typically represent scenes…

计算机视觉与模式识别 · 计算机科学 2024-11-25 Chao Wang , Krzysztof Wolski , Bernhard Kerbl , Ana Serrano , Mojtaba Bemana , Hans-Peter Seidel , Karol Myszkowski , Thomas Leimkühler

Creating high-fidelity, animatable 3D talking heads is crucial for immersive applications, yet often hindered by the prevalence of low-quality image or video sources, which yield poor 3D reconstructions. In this paper, we introduce…

计算机视觉与模式识别 · 计算机科学 2026-02-09 Ding-Jiun Huang , Yuanhao Wang , Shao-Ji Yuan , Albert Mosella-Montoro , Francisco Vicente Carrasco , Cheng Zhang , Fernando De la Torre