English
Related papers

Related papers: Drivable 3D Gaussian Avatars

200 papers

By equipping the most recent 3D Gaussian Splatting representation with head 3D morphable models (3DMM), existing methods manage to create head avatars with high fidelity. However, most existing methods only reconstruct a head without the…

Computer Vision and Pattern Recognition · Computer Science 2024-05-22 Tianhao Wu , Jing Yang , Zhilin Guo , Jingyi Wan , Fangcheng Zhong , Cengiz Oztireli

Recent advances in diffusion models have made significant progress in digital human generation. However, most existing models still struggle to maintain 3D consistency, temporal coherence, and motion accuracy. A key reason for these…

Graphics · Computer Science 2025-03-21 Xuan Gao , Jingtao Zhou , Dongyu Liu , Yuqi Zhou , Juyong Zhang

Compression techniques for 3D Gaussian Splatting (3DGS) have recently achieved considerable success in minimizing storage overhead for 3D Gaussians while preserving high rendering quality. Despite the impressive storage reduction, the lack…

Computer Vision and Pattern Recognition · Computer Science 2025-10-17 Seungjoo Shin , Jaesik Park , Sunghyun Cho

Reconstructing high-fidelity and animatable 3D head avatars from monocular videos remains a challenging yet essential task. Existing methods based on 3D Gaussian Splatting typically bind Gaussians to mesh triangles and model deformations…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Jiankuo Zhao , Xiangyu Zhu , Zidu Wang , Zhen Lei

Existing Gaussian avatar methods typically parameterize geometry on a body-template surface, which entangles the avatar's representation space with the template's deformation space and limits the capture of layered, off-body, and non-rigid…

Graphics · Computer Science 2026-05-21 Julian Kaltheuner , Jan Spindler , Sina Kitz , Patrick Stotko , Reinhard Klein

We present a method for consistent lighting and shadows when animated 3D Gaussian Splatting (3DGS) avatars interact with 3DGS scenes or with dynamic objects inserted into otherwise static scenes. Our key contribution is Deep Gaussian Shadow…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Aymen Mir , Riza Alp Guler , Jian Wang , Gerard Pons-Moll , Bing Zhou

Existing single-image 3D human avatar methods primarily rely on rigid joint transformations, limiting their ability to model realistic cloth dynamics. We present DynaAvatar, a zero-shot framework that reconstructs animatable 3D human…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Joohyun Kwon , Geonhee Sim , Gyeongsik Moon

Recently, Gaussian splatting has emerged as a robust technique for representing 3D scenes, enabling real-time rasterization and high-fidelity rendering. However, Gaussians' inherent radial symmetry and smoothness constraints limit their…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Yi-Hua Huang , Ming-Xian Lin , Yang-Tian Sun , Ziyi Yang , Xiaoyang Lyu , Yan-Pei Cao , Xiaojuan Qi

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the complex head poses and…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Kangjie Chen , Minghan Qin , Yu Li , Haoqian Wang

3D semantic occupancy prediction is essential for achieving safe, reliable autonomous driving and robotic navigation. Compared to camera-only perception systems, multi-modal pipelines, especially LiDAR-camera fusion methods, can produce…

Computer Vision and Pattern Recognition · Computer Science 2026-02-17 Lingjun Zhao , Sizhe Wei , James Hays , Lu Gan

We present GSDeformer, a method that enables cage-based deformation on 3D Gaussian Splatting (3DGS). Our approach bridges cage-based deformation and 3DGS by using a proxy point-cloud representation. This point cloud is generated from 3D…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Jiajun Huang , Shuolin Xu , Hongchuan Yu , Tong-Yee Lee

3D-aware generative adversarial networks (GANs) synthesize high-fidelity and multi-view-consistent facial images using only collections of single-view 2D imagery. Towards fine-grained control over facial attributes, recent efforts…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Jingxiang Sun , Xuan Wang , Lizhen Wang , Xiaoyu Li , Yong Zhang , Hongwen Zhang , Yebin Liu

Conformal Geometric Algebra (CGA) is a framework that allows the representation of objects, such as points, planes and spheres, and deformations, such as translations, rotations and dilations as uniform vectors, called multivectors. In this…

Graphics · Computer Science 2021-05-20 Manos Kamarianakis , George Papagiannakis

Modeling animatable human avatars from videos is a long-standing and challenging problem. While conventional methods require per-instance optimization, recent feed-forward methods have been proposed to generate 3D Gaussians with a learnable…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Yifan Liu , Shengjun Zhang , Chensheng Dai , Yang Chen , Hao Liu , Chen Li , Yueqi Duan

Vision-language-action (VLA) policies have advanced language-conditioned robotic manipulation by transferring semantic priors from pretrained vision-language models to action generation. However, standard action-imitation learning often…

Robotics · Computer Science 2026-05-29 Zijian Zhang , Yuqing Jiang , Qian Cheng , Xiaofan Li , Si Liu , Ding Zhao , Ping Luo , Weitao Zhou , Haibao Yu

Two major approaches exist for creating animatable human avatars. The first, a 3D-based approach, optimizes a NeRF- or 3DGS-based avatar from videos of a single person, achieving personalization through a disentangled identity…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Geonhee Sim , Gyeongsik Moon

In medical image visualization, path tracing of volumetric medical data like CT scans produces lifelike three-dimensional visualizations. Immersive VR displays can further enhance the understanding of complex anatomies. Going beyond the…

Graphics · Computer Science 2026-01-30 Constantin Kleinbeck , Hannah Schieber , Klaus Engel , Ralf Gutjahr , Daniel Roth

Creating relightable and animatable avatars from multi-view or monocular videos is a challenging task for digital human creation and virtual reality applications. Previous methods rely on neural radiance fields or ray tracing, resulting in…

Computer Vision and Pattern Recognition · Computer Science 2025-05-21 Youyi Zhan , Tianjia Shao , He Wang , Yin Yang , Kun Zhou

Virtual try-on systems allow users to interactively try different products within VR scenarios. However, most existing VTON methods operate only on predefined eyewear templates and lack support for fine-grained, user-driven customization.…

Computer Vision and Pattern Recognition · Computer Science 2026-01-27 Rui-Yang Ju , Jen-Shiun Chiang

The fidelity of relighting is bounded by both geometry and appearance representations. For geometry, both mesh and volumetric approaches have difficulty modeling intricate structures like 3D hair geometry. For appearance, existing…

Graphics · Computer Science 2024-05-29 Shunsuke Saito , Gabriel Schwartz , Tomas Simon , Junxuan Li , Giljoo Nam