中文
相关论文

相关论文: Drivable 3D Gaussian Avatars

200 篇论文

By equipping the most recent 3D Gaussian Splatting representation with head 3D morphable models (3DMM), existing methods manage to create head avatars with high fidelity. However, most existing methods only reconstruct a head without the…

计算机视觉与模式识别 · 计算机科学 2024-05-22 Tianhao Wu , Jing Yang , Zhilin Guo , Jingyi Wan , Fangcheng Zhong , Cengiz Oztireli

Recent advances in diffusion models have made significant progress in digital human generation. However, most existing models still struggle to maintain 3D consistency, temporal coherence, and motion accuracy. A key reason for these…

图形学 · 计算机科学 2025-03-21 Xuan Gao , Jingtao Zhou , Dongyu Liu , Yuqi Zhou , Juyong Zhang

Compression techniques for 3D Gaussian Splatting (3DGS) have recently achieved considerable success in minimizing storage overhead for 3D Gaussians while preserving high rendering quality. Despite the impressive storage reduction, the lack…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Seungjoo Shin , Jaesik Park , Sunghyun Cho

Reconstructing high-fidelity and animatable 3D head avatars from monocular videos remains a challenging yet essential task. Existing methods based on 3D Gaussian Splatting typically bind Gaussians to mesh triangles and model deformations…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Jiankuo Zhao , Xiangyu Zhu , Zidu Wang , Zhen Lei

Existing Gaussian avatar methods typically parameterize geometry on a body-template surface, which entangles the avatar's representation space with the template's deformation space and limits the capture of layered, off-body, and non-rigid…

图形学 · 计算机科学 2026-05-21 Julian Kaltheuner , Jan Spindler , Sina Kitz , Patrick Stotko , Reinhard Klein

We present a method for consistent lighting and shadows when animated 3D Gaussian Splatting (3DGS) avatars interact with 3DGS scenes or with dynamic objects inserted into otherwise static scenes. Our key contribution is Deep Gaussian Shadow…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Aymen Mir , Riza Alp Guler , Jian Wang , Gerard Pons-Moll , Bing Zhou

Existing single-image 3D human avatar methods primarily rely on rigid joint transformations, limiting their ability to model realistic cloth dynamics. We present DynaAvatar, a zero-shot framework that reconstructs animatable 3D human…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Joohyun Kwon , Geonhee Sim , Gyeongsik Moon

Recently, Gaussian splatting has emerged as a robust technique for representing 3D scenes, enabling real-time rasterization and high-fidelity rendering. However, Gaussians' inherent radial symmetry and smoothness constraints limit their…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Yi-Hua Huang , Ming-Xian Lin , Yang-Tian Sun , Ziyi Yang , Xiaoyang Lyu , Yan-Pei Cao , Xiaojuan Qi

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the complex head poses and…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Kangjie Chen , Minghan Qin , Yu Li , Haoqian Wang

3D semantic occupancy prediction is essential for achieving safe, reliable autonomous driving and robotic navigation. Compared to camera-only perception systems, multi-modal pipelines, especially LiDAR-camera fusion methods, can produce…

计算机视觉与模式识别 · 计算机科学 2026-02-17 Lingjun Zhao , Sizhe Wei , James Hays , Lu Gan

We present GSDeformer, a method that enables cage-based deformation on 3D Gaussian Splatting (3DGS). Our approach bridges cage-based deformation and 3DGS by using a proxy point-cloud representation. This point cloud is generated from 3D…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Jiajun Huang , Shuolin Xu , Hongchuan Yu , Tong-Yee Lee

3D-aware generative adversarial networks (GANs) synthesize high-fidelity and multi-view-consistent facial images using only collections of single-view 2D imagery. Towards fine-grained control over facial attributes, recent efforts…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Jingxiang Sun , Xuan Wang , Lizhen Wang , Xiaoyu Li , Yong Zhang , Hongwen Zhang , Yebin Liu

Conformal Geometric Algebra (CGA) is a framework that allows the representation of objects, such as points, planes and spheres, and deformations, such as translations, rotations and dilations as uniform vectors, called multivectors. In this…

图形学 · 计算机科学 2021-05-20 Manos Kamarianakis , George Papagiannakis

Modeling animatable human avatars from videos is a long-standing and challenging problem. While conventional methods require per-instance optimization, recent feed-forward methods have been proposed to generate 3D Gaussians with a learnable…

计算机视觉与模式识别 · 计算机科学 2025-07-28 Yifan Liu , Shengjun Zhang , Chensheng Dai , Yang Chen , Hao Liu , Chen Li , Yueqi Duan

Vision-language-action (VLA) policies have advanced language-conditioned robotic manipulation by transferring semantic priors from pretrained vision-language models to action generation. However, standard action-imitation learning often…

机器人学 · 计算机科学 2026-05-29 Zijian Zhang , Yuqing Jiang , Qian Cheng , Xiaofan Li , Si Liu , Ding Zhao , Ping Luo , Weitao Zhou , Haibao Yu

Two major approaches exist for creating animatable human avatars. The first, a 3D-based approach, optimizes a NeRF- or 3DGS-based avatar from videos of a single person, achieving personalization through a disentangled identity…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Geonhee Sim , Gyeongsik Moon

In medical image visualization, path tracing of volumetric medical data like CT scans produces lifelike three-dimensional visualizations. Immersive VR displays can further enhance the understanding of complex anatomies. Going beyond the…

图形学 · 计算机科学 2026-01-30 Constantin Kleinbeck , Hannah Schieber , Klaus Engel , Ralf Gutjahr , Daniel Roth

Creating relightable and animatable avatars from multi-view or monocular videos is a challenging task for digital human creation and virtual reality applications. Previous methods rely on neural radiance fields or ray tracing, resulting in…

计算机视觉与模式识别 · 计算机科学 2025-05-21 Youyi Zhan , Tianjia Shao , He Wang , Yin Yang , Kun Zhou

Virtual try-on systems allow users to interactively try different products within VR scenarios. However, most existing VTON methods operate only on predefined eyewear templates and lack support for fine-grained, user-driven customization.…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Rui-Yang Ju , Jen-Shiun Chiang

The fidelity of relighting is bounded by both geometry and appearance representations. For geometry, both mesh and volumetric approaches have difficulty modeling intricate structures like 3D hair geometry. For appearance, existing…

图形学 · 计算机科学 2024-05-29 Shunsuke Saito , Gabriel Schwartz , Tomas Simon , Junxuan Li , Giljoo Nam