中文
相关论文

相关论文: LIFe-GoM: Generalizable Human Rendering with Learn…

200 篇论文

Photorealistic and controllable human avatars have gained popularity in the research community thanks to rapid advances in neural rendering, providing fast and realistic synthesis tools. However, a limitation of current solutions is the…

计算机视觉与模式识别 · 计算机科学 2025-09-03 Mohamed Ilyes Lakhal , Richard Bowden

Real-time, high-fidelity 3D human reconstruction from RGB images is essential for interactive applications such as virtual reality and gaming, yet remains challenging due to the complex non-rigid deformations of dynamic human bodies.…

多媒体 · 计算机科学 2026-04-24 Yang Liu , Zhiyong Zhang

We have recently seen tremendous progress in photo-real human modeling and rendering. Yet, efficiently rendering realistic human performance and integrating it into the rasterization pipeline remains challenging. In this paper, we present…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Yuheng Jiang , Zhehao Shen , Penghao Wang , Zhuo Su , Yu Hong , Yingliang Zhang , Jingyi Yu , Lan Xu

Recent advancements in 3D Gaussian Splatting (3DGS) have unlocked significant potential for modeling 3D head avatars, providing greater flexibility than mesh-based methods and more efficient rendering compared to NeRF-based approaches.…

计算机视觉与模式识别 · 计算机科学 2024-11-07 Peizhi Yan , Rabab Ward , Qiang Tang , Shan Du

In this paper, we present a method to reconstruct the world and multiple dynamic humans in 3D from a monocular video input. As a key idea, we represent both the world and multiple humans via the recently emerging 3D Gaussian Splatting…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Inhee Lee , Byungjun Kim , Hanbyul Joo

Existing full-body Gaussian avatar methods primarily optimize global reconstruction quality and often fail to preserve fine-grained facial geometry and expression details. This challenge arises from limited facial representational capacity…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Willem Menu , Erkut Akdag , Pedro Quesado , Yasaman Kashefbahrami , Egor Bondarev

We present GaussianAvatar, an efficient approach to creating realistic human avatars with dynamic 3D appearances from a single video. We start by introducing animatable 3D Gaussians to explicitly represent humans in various poses and…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Liangxiao Hu , Hongwen Zhang , Yuxiang Zhang , Boyao Zhou , Boning Liu , Shengping Zhang , Liqiang Nie

3D Gaussian Splatting (3DGS) provides an efficient method for high-quality scene reconstruction using anisotropic Gaussians. Recently, 3DGS-based methods have significantly improved the rendering quality of human avatars while enabling…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Hongzhe Liao , Chuhua Xian , Hongmin Cai , Haiyang Liu , Fa-Ting Hong

Creating high-fidelity 3D human head avatars is crucial for applications in VR/AR, digital human, and film production. Recent advances have leveraged morphable face models to generate animated head avatars from easily accessible data,…

计算机视觉与模式识别 · 计算机科学 2024-10-24 Yuelang Xu , Zhaoqi Su , Qingyao Wu , Yebin Liu

In this paper, we introduce GaussianMotion, a novel human rendering model that generates fully animatable scenes aligned with textual descriptions using Gaussian Splatting. Although existing methods achieve reasonable text-to-3D generation…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Gyumin Shim , Sangmin Lee , Jaegul Choo

Recent advancements in radiance field rendering show promising results in 3D scene representation, where Gaussian splatting-based techniques emerge as state-of-the-art due to their quality and efficiency. Gaussian splatting is widely used…

计算机视觉与模式识别 · 计算机科学 2024-11-06 Arnab Dey , Cheng-You Lu , Andrew I. Comport , Srinath Sridhar , Chin-Teng Lin , Jean Martinet

Real-time rendering of high-fidelity and animatable avatars from monocular videos remains a challenging problem in computer vision and graphics. Over the past few years, the Neural Radiance Field (NeRF) has made significant progress in…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Qipeng Yan , Mingyang Sun , Lihua Zhang

Reconstructing photo-realistic drivable human avatars from multi-view image sequences has been a popular and challenging topic in the field of computer vision and graphics. While existing NeRF-based methods can achieve high-quality novel…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Yujiao Jiang , Qingmin Liao , Xiaoyu Li , Li Ma , Qi Zhang , Chaopeng Zhang , Zongqing Lu , Ying Shan

We present a method for recovering the shape and radiance of a scene consisting of multiple people given solely a few images. Multi-human scenes are complex due to additional occlusion and clutter. For single-human settings, existing…

计算机视觉与模式识别 · 计算机科学 2025-02-12 Qian li , Victoria Fernàndez Abrevaya , Franck Multon , Adnane Boukhayma

We present a novel framework for animating humans in 3D scenes using 3D Gaussian Splatting (3DGS), a neural scene representation that has recently achieved state-of-the-art photorealistic results for novel-view synthesis but remains…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Aymen Mir , Jian Wang , Riza Alp Guler , Chuan Guo , Gerard Pons-Moll , Bing Zhou

Radiance field-based methods have recently been used to reconstruct human avatars, showing that we can significantly downscale the systems needed for creating animated human avatars. Although this progress has been initiated by neural…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Nikolaos Zioulis , Nikolaos Kotarelas , Georgios Albanis , Spyridon Thermos , Anargyros Chatzitofis

Reconstructing real-world objects from multi-view images is essential for applications in 3D editing, AR/VR, and digital content creation. Existing methods typically prioritize either geometric accuracy (Multi-View Stereo) or photorealistic…

计算机视觉与模式识别 · 计算机科学 2026-03-05 Zhejia Cai , Puhua Jiang , Shiwei Mao , Hongkun Cao , Ruqi Huang

We present L4GM, the first 4D Large Reconstruction Model that produces animated objects from a single-view video input -- in a single feed-forward pass that takes only a second. Key to our success is a novel dataset of multiview videos…

计算机视觉与模式识别 · 计算机科学 2024-06-18 Jiawei Ren , Kevin Xie , Ashkan Mirzaei , Hanxue Liang , Xiaohui Zeng , Karsten Kreis , Ziwei Liu , Antonio Torralba , Sanja Fidler , Seung Wook Kim , Huan Ling

We propose GS-LRM, a scalable large reconstruction model that can predict high-quality 3D Gaussian primitives from 2-4 posed sparse images in 0.23 seconds on single A100 GPU. Our model features a very simple transformer-based architecture;…

计算机视觉与模式识别 · 计算机科学 2024-05-01 Kai Zhang , Sai Bi , Hao Tan , Yuanbo Xiangli , Nanxuan Zhao , Kalyan Sunkavalli , Zexiang Xu

Faithfully reconstructing textured shapes and physical properties from videos presents an intriguing yet challenging problem. Significant efforts have been dedicated to advancing such a system identification problem in this area. Previous…

图形学 · 计算机科学 2025-06-10 Chuhao Chen , Zhiyang Dou , Chen Wang , Yiming Huang , Anjun Chen , Qiao Feng , Jiatao Gu , Lingjie Liu