English
Related papers

Related papers: GART: Gaussian Articulated Template Models

200 papers

We present an approach for high-quality dynamic Gaussian Splatting from monocular videos. To this end, we in this work go one step further beyond previous methods to explicitly model continuous position and orientation deformation of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Xuankai Zhang , Junjin Xiao , Shangwei Huang , Wei-shi Zheng , Qing Zhang

We introduce a novel 3D generation method for versatile and high-quality 3D asset creation. The cornerstone is a unified Structured LATent (SLAT) representation which allows decoding to different output formats, such as Radiance Fields, 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-06-02 Jianfeng Xiang , Zelong Lv , Sicheng Xu , Yu Deng , Ruicheng Wang , Bowen Zhang , Dong Chen , Xin Tong , Jiaolong Yang

In this paper, we introduce GaussianMotion, a novel human rendering model that generates fully animatable scenes aligned with textual descriptions using Gaussian Splatting. Although existing methods achieve reasonable text-to-3D generation…

Computer Vision and Pattern Recognition · Computer Science 2025-02-18 Gyumin Shim , Sangmin Lee , Jaegul Choo

We present a novel method for 6-DoF object tracking and high-quality 3D reconstruction from monocular RGBD video. Existing methods, while achieving impressive results, often struggle with complex objects, particularly those exhibiting…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Takuya Ikeda , Sergey Zakharov , Muhammad Zubair Irshad , Istvan Balazs Opra , Shun Iwase , Dian Chen , Mark Tjersland , Robert Lee , Alexandre Dilly , Rares Ambrus , Koichi Nishiwaki

Achieving dexterous robotic grasping with multi-fingered hands remains a significant challenge. While existing methods rely on complete 3D scans to predict grasp poses, these approaches face limitations due to the difficulty of acquiring…

Reconstructing clean, distractor-free 3D scenes from real-world captures remains a significant challenge, particularly in highly dynamic and cluttered settings such as egocentric videos. To tackle this problem, we introduce DeGauss, a…

Computer Vision and Pattern Recognition · Computer Science 2025-07-25 Rui Wang , Quentin Lohmeyer , Mirko Meboldt , Siyu Tang

Modeling animatable human avatars from RGB videos is a long-standing and challenging problem. Recent works usually adopt MLP-based neural radiance fields (NeRF) to represent 3D humans, but it remains difficult for pure MLPs to regress…

Computer Vision and Pattern Recognition · Computer Science 2024-05-28 Zhe Li , Yipengjing Sun , Zerong Zheng , Lizhen Wang , Shengping Zhang , Yebin Liu

Reconstructing dynamic 3D scenes from monocular video remains fundamentally challenging due to the need to jointly infer motion, structure, and appearance from limited observations. Existing dynamic scene reconstruction methods based on…

Computer Vision and Pattern Recognition · Computer Science 2025-08-07 Jiahui Li , Shengeng Tang , Jingxuan He , Gang Huang , Zhangye Wang , Yantao Pan , Lechao Cheng

Free-moving object reconstruction from monocular video remains challenging, particularly without reliable pose or depth cues and under arbitrary object motion. We introduce OnlineSplatter, a novel online feed-forward framework generating…

Computer Vision and Pattern Recognition · Computer Science 2025-10-24 Mark He Huang , Lin Geng Foo , Christian Theobalt , Ying Sun , De Wen Soh

Reconstructing dynamic 3D scenes from monocular input is fundamentally under-constrained, with ambiguities arising from occlusion and extreme novel views. While dynamic Gaussian Splatting offers an efficient representation, vanilla models…

Computer Vision and Pattern Recognition · Computer Science 2026-03-02 Fengzhi Guo , Chih-Chuan Hsu , Sihao Ding , Cheng Zhang

In this work, we introduce Monocular and Generalizable Gaussian Talking Head Animation (MGGTalk), which requires monocular datasets and generalizes to unseen identities without personalized re-training. Compared with previous 3D Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Shengjie Gong , Haojie Li , Jiapeng Tang , Dongming Hu , Shuangping Huang , Hao Chen , Tianshui Chen , Zhuoman Liu

Implicit Neural Representation (INR) has demonstrated remarkable advances in the field of image representation but demands substantial GPU resources. GaussianImage recently pioneered the use of Gaussian Splatting to mitigate this cost,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Zhaojie Zeng , Yuesong Wang , Chao Yang , Tao Guan , Lili Ju

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

High-quality and controllable digital twins of surgical instruments are critical for Real2Sim in robot-assisted surgery, as they enable realistic simulation, synthetic data generation, and perception learning under novel poses. We present…

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

Reconstructing articulated 3D objects from a single image requires jointly inferring object geometry, part structure, and motion parameters from limited visual evidence. A key difficulty lies in the entanglement between motion cues and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Haitian Li , Haozhe Xie , Junxiang Xu , Beichen Wen , Fangzhou Hong , Ziwei Liu

Photorealistic and controllable human avatars have gained popularity in the research community thanks to rapid advances in neural rendering, providing fast and realistic synthesis tools. However, a limitation of current solutions is the…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Mohamed Ilyes Lakhal , Richard Bowden

Representing and rendering dynamic scenes has been an important but challenging task. Especially, to accurately model complex motions, high efficiency is usually hard to guarantee. To achieve real-time dynamic scene rendering while also…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Guanjun Wu , Taoran Yi , Jiemin Fang , Lingxi Xie , Xiaopeng Zhang , Wei Wei , Wenyu Liu , Qi Tian , Xinggang Wang

Real2Sim is becoming increasingly important with the rapid development of surgical artificial intelligence (AI) and autonomy. In this work, we propose a novel Real2Sim methodology, Instrument-Splatting, that leverages 3D Gaussian Splatting…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Shuojue Yang , Zijian Wu , Mingxuan Hong , Qian Li , Daiyun Shen , Septimiu E. Salcudean , Yueming Jin

3D Gaussian Splatting (3DGS) provides an efficient method for high-quality scene reconstruction using anisotropic Gaussians. Recently, 3DGS-based methods have significantly improved the rendering quality of human avatars while enabling…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Hongzhe Liao , Chuhua Xian , Hongmin Cai , Haiyang Liu , Fa-Ting Hong