English
Related papers

Related papers: GSTalker: Real-time Audio-Driven Talking Face Gene…

200 papers

Modern Gaussian Splatting methods have proven highly effective for real-time photorealistic rendering of 3D scenes. However, integrating semantic information into this representation remains a significant challenge, especially in…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Roman Titkov , Egor Zubkov , Dmitry Yudin , Jaafar Mahmoud , Malik Mohrat , Gennady Sidorov

Dynamic extensions of 3D Gaussian Splatting (3DGS) achieve high-quality reconstructions through neural motion fields, but per-Gaussian neural inference makes these models computationally expensive. Building on DeformableGS, we introduce…

Graphics · Computer Science 2026-03-31 Allen Tu , Haiyang Ying , Alex Hanson , Yonghan Lee , Tom Goldstein , Matthias Zwicker

Recently, 3D Gaussian Splatting (3DGS) has emerged as an efficient approach for accurately representing scenes. However, despite its superior novel view synthesis capabilities, extracting the geometry of the scene directly from the Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Yaniv Wolf , Amit Bracha , Ron Kimmel

Sparse Multi-view Images can be Learned to predict explicit radiance fields via Generalizable Gaussian Splatting approaches, which can achieve wider application prospects in real-life when ground-truth camera parameters are not required as…

Computer Vision and Pattern Recognition · Computer Science 2024-11-28 Yanyan Li , Yixin Fang , Federico Tombari , Gim Hee Lee

Speech-driven 3D facial animation has been widely explored, with applications in gaming, character animation, virtual reality, and telepresence systems. State-of-the-art methods deform the face topology of the target actor to sync the input…

Computer Vision and Pattern Recognition · Computer Science 2023-01-03 Balamurugan Thambiraja , Ikhsanul Habibie , Sadegh Aliakbarian , Darren Cosker , Christian Theobalt , Justus Thies

Modeling animatable human avatars from monocular or multi-view videos has been widely studied, with recent approaches leveraging neural radiance fields (NeRFs) or 3D Gaussian Splatting (3DGS) achieving impressive results in novel-view and…

Computer Vision and Pattern Recognition · Computer Science 2025-04-03 Yahui Li , Zhi Zeng , Liming Pang , Guixuan Zhang , Shuwu Zhang

3D Gaussian Splatting (GS) enables highly photorealistic scene reconstruction from posed image sequences but struggles with viewpoint extrapolation due to its anisotropic nature, leading to overfitting and poor generalization, particularly…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Shuohan Tao , Boyao Zhou , Hanzhang Tu , Yuwang Wang , Yebin Liu

We present a novel animatable 3D Gaussian model for rendering high-fidelity free-view human motions in real time. Compared to existing NeRF-based methods, the model owns better capability in synthesizing high-frequency details without the…

Computer Vision and Pattern Recognition · Computer Science 2023-11-28 Keyang Ye , Tianjia Shao , Kun Zhou

Gaussian splatting enables fast novel view synthesis in static 3D environments. However, reconstructing real-world environments remains challenging as distractors or occluders break the multi-view consistency assumption required for…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Yihao Wang , Marcus Klasson , Matias Turkulainen , Shuzhe Wang , Juho Kannala , Arno Solin

Text-to-3D synthesis has recently seen intriguing advances by combining the text-to-image priors with 3D representation methods, e.g., 3D Gaussian Splatting (3D GS), via Score Distillation Sampling (SDS). However, a hurdle of existing…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Lutao Jiang , Xu Zheng , Yuanhuiyi Lyu , Jiazhou Zhou , Lin Wang

Recent 4D reconstruction methods have yielded impressive results but rely on sharp videos as supervision. However, motion blur often occurs in videos due to camera shake and object movement, while existing methods render blurry results when…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Renlong Wu , Zhilu Zhang , Mingyang Chen , Zifei Yan , Wangmeng Zuo

Reconstructing photorealistic and topology-aware human avatars from monocular videos remains a significant challenge in the fields of computer vision and graphics. While existing 3D human avatar modeling approaches can effectively capture…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Yuze Su , Hongsong Wang , Jie Gui , Liang Wang

Implicit Neural Representation for Videos (NeRV) has introduced a novel paradigm for video representation and compression, outperforming traditional codecs. As model size grows, however, slow encoding and decoding speed and high memory…

Computer Vision and Pattern Recognition · Computer Science 2025-03-07 Inseo Lee , Youngyoon Choi , Joonseok Lee

Efficient and high-fidelity reconstruction of deformable surgical scenes is a critical yet challenging task. Building on recent advancements in 3D Gaussian splatting, current methods have seen significant improvements in both reconstruction…

Computer Vision and Pattern Recognition · Computer Science 2025-01-03 Jiwei Shan , Zeyu Cai , Cheng-Tai Hsieh , Shing Shin Cheng , Hesheng Wang

We propose VASA-3D, an audio-driven, single-shot 3D head avatar generator. This research tackles two major challenges: capturing the subtle expression details present in real human faces, and reconstructing an intricate 3D head avatar from…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Sicheng Xu , Guojun Chen , Jiaolong Yang , Yizhong Zhang , Yu Deng , Steve Lin , Baining Guo

Rendering dynamic scenes from monocular videos is a crucial yet challenging task. The recent deformable Gaussian Splatting has emerged as a robust solution to represent real-world dynamic scenes. However, it often leads to heavily redundant…

Computer Vision and Pattern Recognition · Computer Science 2025-02-28 Hanyang Kong , Xingyi Yang , Xinchao Wang

We introduce Dr. Splat, a novel approach for open-vocabulary 3D scene understanding leveraging 3D Gaussian Splatting. Unlike existing language-embedded 3DGS methods, which rely on a rendering process, our method directly associates…

Computer Vision and Pattern Recognition · Computer Science 2025-02-25 Kim Jun-Seong , GeonU Kim , Kim Yu-Ji , Yu-Chiang Frank Wang , Jaesung Choe , Tae-Hyun Oh

Interactive segmentation of 3D Gaussians opens a great opportunity for real-time manipulation of 3D scenes thanks to the real-time rendering capability of 3D Gaussian Splatting. However, the current methods suffer from time-consuming…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Seokhun Choi , Hyeonseop Song , Jaechul Kim , Taehyeong Kim , Hoseok Do

We introduce the \method, an ultra-efficient approach for monocular 3D object reconstruction. Splatter Image is based on Gaussian Splatting, which allows fast and high-quality reconstruction of 3D scenes from multiple images. We apply…

Computer Vision and Pattern Recognition · Computer Science 2024-04-17 Stanislaw Szymanowicz , Christian Rupprecht , Andrea Vedaldi

Dynamic scene reconstruction has garnered significant attention in recent years due to its capabilities in high-quality and real-time rendering. Among various methodologies, constructing a 4D spatial-temporal representation, such as 4D-GS,…

Computer Vision and Pattern Recognition · Computer Science 2024-08-27 Weiwei Cai , Weicai Ye , Peng Ye , Tong He , Tao Chen
‹ Prev 1 4 5 6 7 8 10 Next ›