中文
相关论文

相关论文: LoopGaussian: Creating 3D Cinemagraph with Multi-v…

200 篇论文

3D Gaussian splatting (3DGS) is an emerging technique for photorealistic 3D scene rendering. However, rendering city-scale 3DGS scenes on resource-constrained mobile devices in real-time remains a significant challenge due to two…

图形学 · 计算机科学 2025-09-29 Zheng Liu , He Zhu , Xinyang Li , Yirun Wang , Yujiao Shi , Yiming Gan , Wei Li , Jingwen Leng , Minyi Guo , Yu Feng

3D Gaussian Splatting (3DGS) has recently enabled real-time rendering of unbounded 3D scenes for novel view synthesis. However, this technique requires dense training views to accurately reconstruct 3D geometry. A limited number of input…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Haolin Xiong , Sairisheek Muttukuru , Rishi Upadhyay , Pradyumna Chari , Achuta Kadambi

Interactive segmentation of 3D Gaussians opens a great opportunity for real-time manipulation of 3D scenes thanks to the real-time rendering capability of 3D Gaussian Splatting. However, the current methods suffer from time-consuming…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Seokhun Choi , Hyeonseop Song , Jaechul Kim , Taehyeong Kim , Hoseok Do

Urban scene reconstruction is crucial for real-world autonomous driving simulators. Although existing methods have achieved photorealistic reconstruction, they mostly focus on pinhole cameras and neglect fisheye cameras. In fact, how to…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Yuan Ren , Guile Wu , Runhao Li , Zheyuan Yang , Yibo Liu , Xingxin Chen , Tongtong Cao , Bingbing Liu

While visual-language models have profoundly linked features between texts and images, the incorporation of 3D modality data, such as point clouds and 3D Gaussians, further enables pretraining for 3D-related tasks, e.g., cross-modal…

计算机视觉与模式识别 · 计算机科学 2026-01-28 Jiarun Liu , Qifeng Chen , Yiru Zhao , Minghua Liu , Baorui Ma , Sheng Yang

Single-image 3D scene reconstruction presents significant challenges due to its inherently ill-posed nature and limited input constraints. Recent advances have explored two promising directions: multiview generative models that train on 3D…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Junlin Hao , Peiheng Wang , Haoyang Wang , Xinggong Zhang , Zongming Guo

Due to the complex and highly dynamic motions in the real world, synthesizing dynamic videos from multi-view inputs for arbitrary viewpoints is challenging. Previous works based on neural radiance field or 3D Gaussian splatting are limited…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Jiahao Wu , Rui Peng , Jianbo Jiao , Jiayu Yang , Luyang Tang , Kaiqiang Xiong , Jie Liang , Jinbo Yan , Runling Liu , Ronggang Wang

The creation of high-quality 3D assets is paramount for applications in digital heritage preservation, entertainment, and robotics. Traditionally, this process necessitates skilled professionals and specialized software for the modeling,…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Shen Chen , Jiale Zhou , Zhongyu Jiang , Tianfang Zhang , Zongkai Wu , Jenq-Neng Hwang , Lei Li

3D Gaussian Splatting (3DGS) has emerged as a cutting-edge technique for real-time radiance field rendering, offering state-of-the-art performance in terms of both quality and speed. 3DGS models a scene as a collection of three-dimensional…

计算机视觉与模式识别 · 计算机科学 2025-03-06 Milena T. Bagdasarian , Paul Knoll , Yi-Hsin Li , Florian Barthel , Anna Hilsmann , Peter Eisert , Wieland Morgenstern

Empowering 3D Gaussian Splatting with generalization ability is appealing. However, existing generalizable 3D Gaussian Splatting methods are largely confined to narrow-range interpolation between stereo images due to their heavy backbones,…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Yunsong Wang , Tianxin Huang , Hanlin Chen , Gim Hee Lee

Recently, 3D Gaussian Splatting (3DGS), an explicit scene representation technique, has shown significant promise for dynamic novel-view synthesis from monocular video input. However, purely data-driven 3DGS often struggles to capture the…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Haoqin Hong , Ding Fan , Fubin Dou , Zhi-Li Zhou , Haoran Sun , Congcong Zhu , Jingrun Chen

Existing NeRF-based methods for large scene reconstruction often have limitations in visual quality and rendering speed. While the recent 3D Gaussian Splatting works well on small-scale and object-centric scenes, scaling it up to large…

计算机视觉与模式识别 · 计算机科学 2024-02-28 Jiaqi Lin , Zhihao Li , Xiao Tang , Jianzhuang Liu , Shiyong Liu , Jiayue Liu , Yangdi Lu , Xiaofei Wu , Songcen Xu , Youliang Yan , Wenming Yang

Novel view synthesis (NVS) in low-light scenes remains a significant challenge due to degraded inputs characterized by severe noise, low dynamic range (LDR) and unreliable initialization. While recent NeRF-based approaches have shown…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Hao Sun , Fenggen Yu , Huiyao Xu , Tao Zhang , Changqing Zou

Text-to-3D scene generation holds immense potential for the gaming, film, and architecture sectors. Despite significant progress, existing methods struggle with maintaining high quality, consistency, and editing flexibility. In this paper,…

计算机视觉与模式识别 · 计算机科学 2024-07-22 Haoran Li , Haolin Shi , Wenli Zhang , Wenjun Wu , Yong Liao , Lin Wang , Lik-hang Lee , Pengyuan Zhou

Reconstructing complete and interactive 3D scenes remains a fundamental challenge in computer vision and robotics, particularly due to persistent object occlusions and limited sensor coverage. Multiview observations from a single scene scan…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Wenhao Hu , Zesheng Li , Haonan Zhou , Liu Liu , Xuexiang Wen , Zhizhong Su , Xi Li , Gaoang Wang

We introduce GeoGS3D, a novel two-stage framework for reconstructing detailed 3D objects from single-view images. Inspired by the success of pre-trained 2D diffusion models, our method incorporates an orthogonal plane decomposition…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Qijun Feng , Zhen Xing , Zuxuan Wu , Yu-Gang Jiang

We present Gaussian See, Gaussian Do, a novel approach for semantic 3D motion transfer from multiview video. Our method enables rig-free, cross-category motion transfer between objects with semantically meaningful correspondence. Building…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Yarin Bekor , Gal Michael Harari , Or Perel , Or Litany

Open-vocabulary 3D scene understanding presents a significant challenge in computer vision, with wide-ranging applications in embodied agents and augmented reality systems. Existing methods adopt neurel rendering methods as 3D…

计算机视觉与模式识别 · 计算机科学 2024-08-26 Jun Guo , Xiaojian Ma , Yue Fan , Huaping Liu , Qing Li

We present CrowdSplat, a novel approach that leverages 3D Gaussian Splatting for real-time, high-quality crowd rendering. Our method utilizes 3D Gaussian functions to represent animated human characters in diverse poses and outfits, which…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Xiaohan Sun , Yinghan Xu , John Dingliana , Carol O'Sullivan

Omnidirectional (or 360-degree) images are increasingly being used for 3D applications since they allow the rendering of an entire scene with a single image. Existing works based on neural radiance fields demonstrate successful 3D…

计算机视觉与模式识别 · 计算机科学 2024-10-29 Suyoung Lee , Jaeyoung Chung , Jaeyoo Huh , Kyoung Mu Lee