English
Related papers

Related papers: GaussianCross: Cross-modal Self-supervised 3D Repr…

200 papers

Recent feed-forward Gaussian reconstruction models adopt a pixel-aligned formulation that maps each 2D pixel to a 3D Gaussian, entangling Gaussian representations tightly with the input images. In this paper, we propose AnchorSplat, a novel…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Xiaoxue Zhang , Xiaoxu Zheng , Yixuan Yin , Tiao Zhao , Kaihua Tang , Michael Bi Mi , Zhan Xu , Dave Zhenyu Chen

We propose a novel 3D deepfake generation framework based on 3D Gaussian Splatting that enables realistic, identity-preserving face swapping and reenactment in a fully controllable 3D space. Compared to conventional 2D deepfake approaches…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Wending Liu , Siyun Liang , Huy H. Nguyen , Isao Echizen

In this paper, we explore the existing challenges in 3D artistic scene generation by introducing ART3D, a novel framework that combines diffusion models and 3D Gaussian splatting techniques. Our method effectively bridges the gap between…

Computer Vision and Pattern Recognition · Computer Science 2024-05-20 Pengzhi Li , Chengshuai Tang , Qinxuan Huang , Zhiheng Li

Self-modeling enables robots to build task-agnostic models of their morphology and kinematics based on data that can be automatically collected, with minimal human intervention and prior information, thereby enhancing machine intelligence.…

Robotics · Computer Science 2025-11-17 Kejun Hu , Peng Yu , Ning Tan

3D Gaussian Splatting has recently gained traction for its efficient training and real-time rendering. While its vanilla representation is mainly designed for view synthesis, recent works extended it to scene understanding with language…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Siyun Liang , Sen Wang , Kunyi Li , Michael Niemeyer , Stefano Gasperini , Hendrik P. A. Lensch , Nassir Navab , Federico Tombari

While 3DGS has emerged as a high-fidelity scene representation, encoding rich, general-purpose features directly from its primitives remains under-explored. We address this gap by introducing Chorus, a multi-teacher pretraining framework…

Computer Vision and Pattern Recognition · Computer Science 2026-05-06 Yue Li , Qi Ma , Runyi Yang , Mengjiao Ma , Bin Ren , Nikola Popovic , Nicu Sebe , Theo Gevers , Luc Van Gool , Danda Pani Paudel , Martin R. Oswald

Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content, posing significant challenges for efficient compression and real-time decoding on end-user devices. Inspired by 3D Gaussian Splatting,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Linfei Li , Lin Zhang , Zhong Wang , Ying Shen

Recently, Gaussian Splatting, a method that represents a 3D scene as a collection of Gaussian distributions, has gained significant attention in addressing the task of novel view synthesis. In this paper, we highlight a fundamental…

Computer Vision and Pattern Recognition · Computer Science 2024-10-31 Haoxuan Qu , Zhuoling Li , Hossein Rahmani , Yujun Cai , Jun Liu

Reconstructing and rendering 3D objects from highly sparse views is of critical importance for promoting applications of 3D vision techniques and improving user experience. However, images from sparse views only contain very limited 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-11-14 Chen Yang , Sikuang Li , Jiemin Fang , Ruofan Liang , Lingxi Xie , Xiaopeng Zhang , Wei Shen , Qi Tian

Abstract representations of 3D scenes play a crucial role in computer vision, enabling a wide range of applications such as mapping, localization, surface reconstruction, and even advanced tasks like SLAM and rendering. Among these…

Computer Vision and Pattern Recognition · Computer Science 2024-12-16 Chenggang Yang , Yuang Shi

Image-based 3D reconstruction is a challenging task that involves inferring the 3D shape of an object or scene from a set of input images. Learning-based methods have gained attention for their ability to directly estimate 3D shapes. This…

Computer Vision and Pattern Recognition · Computer Science 2024-09-02 Anurag Dalal , Daniel Hagen , Kjell G. Robbersmyr , Kristian Muri Knausgård

The advent of neural 3D Gaussians has recently brought about a revolution in the field of neural rendering, facilitating the generation of high-quality renderings at real-time speeds. However, the explicit and discrete representation…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Yingwenqi Jiang , Jiadong Tu , Yuan Liu , Xifeng Gao , Xiaoxiao Long , Wenping Wang , Yuexin Ma

Existing neural implicit surface reconstruction methods have achieved impressive performance in multi-view 3D reconstruction by leveraging explicit geometry priors such as depth maps or point clouds as regularization. However, the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Hanlin Chen , Chen Li , Yunsong Wang , Gim Hee Lee

3D Gaussian Splatting is renowned for its high-fidelity reconstructions and real-time novel view synthesis, yet its lack of semantic understanding limits object-level perception. In this work, we propose ObjectGS, an object-aware framework…

Graphics · Computer Science 2025-07-22 Ruijie Zhu , Mulin Yu , Linning Xu , Lihan Jiang , Yixuan Li , Tianzhu Zhang , Jiangmiao Pang , Bo Dai

Photorealistic 3D reconstruction of street scenes is a critical technique for developing real-world simulators for autonomous driving. Despite the efficacy of Neural Radiance Fields (NeRF) for driving scenes, 3D Gaussian Splatting (3DGS)…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Nan Huang , Xiaobao Wei , Wenzhao Zheng , Pengju An , Ming Lu , Wei Zhan , Masayoshi Tomizuka , Kurt Keutzer , Shanghang Zhang

Rendering novel view images in dynamic scenes is a crucial yet challenging task. Current methods mainly utilize NeRF-based methods to represent the static scene and an additional time-variant MLP to model scene deformations, resulting in…

Computer Vision and Pattern Recognition · Computer Science 2024-06-07 Diwen Wan , Ruijie Lu , Gang Zeng

While visual-language models have profoundly linked features between texts and images, the incorporation of 3D modality data, such as point clouds and 3D Gaussians, further enables pretraining for 3D-related tasks, e.g., cross-modal…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 Jiarun Liu , Qifeng Chen , Yiru Zhao , Minghua Liu , Baorui Ma , Sheng Yang

Since its introduction, 3D Gaussian Splatting (3DGS) has rapidly transformed the landscape of 3D scene representations, inspiring an extensive body of associated research. Follow-up work includes analyses and contributions that enhance the…

Computer Vision and Pattern Recognition · Computer Science 2025-10-31 Bernhard Kerbl

The automatic reconstruction of 3D computer-aided design (CAD) models from CAD sketches has recently gained significant attention in the computer vision community. Most existing methods, however, rely on vector CAD sketches and 3D ground…

Computer Vision and Pattern Recognition · Computer Science 2025-03-10 Zheng Zhou , Zhe Li , Bo Yu , Lina Hu , Liang Dong , Zijian Yang , Xiaoli Liu , Ning Xu , Ziwei Wang , Yonghao Dang , Jianqin Yin

3D Gaussian splatting (3DGS) has demonstrated impressive performance in synthesizing high-fidelity novel views. Nonetheless, its effectiveness critically depends on the quality of the initialized point cloud. Specifically, achieving uniform…

Computer Vision and Pattern Recognition · Computer Science 2025-10-13 Yikang Zhang , Rui Fan