中文
相关论文

相关论文: MIGS: Multi-Identity Gaussian Splatting via Tensor…

200 篇论文

Controllable video generation (CVG) has advanced rapidly, yet current systems falter when more than one actor must move, interact, and exchange positions under noisy control signals. We address this gap with DanceTogether, the first…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Junhao Chen , Mingjin Chen , Jianjin Xu , Xiang Li , Junting Dong , Mingze Sun , Puhua Jiang , Hongxiang Li , Yuhang Yang , Hao Zhao , Xiaoxiao Long , Ruqi Huang

We propose a method to enhance 3D Gaussian Splatting (3DGS)~\cite{Kerbl2023}, addressing challenges in initialization, optimization, and density control. Gaussian Splatting is an alternative for rendering realistic images while supporting…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Xingjun Wang , Lianlei Shan

Reliable multimodal sensor fusion algorithms require accurate spatiotemporal calibration. Recently, targetless calibration techniques based on implicit neural representations have proven to provide precise and robust results. Nevertheless,…

计算机视觉与模式识别 · 计算机科学 2024-09-19 Quentin Herau , Moussab Bennehar , Arthur Moreau , Nathan Piasco , Luis Roldao , Dzmitry Tsishkou , Cyrille Migniot , Pascal Vasseur , Cédric Demonceaux

Recent advances in 3D Gaussian Splatting have shown remarkable potential for novel view synthesis. However, most existing large-scale scene reconstruction methods rely on the divide-and-conquer paradigm, which often leads to the loss of…

计算机视觉与模式识别 · 计算机科学 2025-05-30 Chuandong Liu , Huijiao Wang , Lei Yu , Gui-Song Xia

We present MoBGS, a novel motion deblurring 3D Gaussian Splatting (3DGS) framework capable of reconstructing sharp and high-quality novel spatio-temporal views from blurry monocular videos in an end-to-end manner. Existing dynamic novel…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Minh-Quan Viet Bui , Jongmin Park , Juan Luis Gonzalez Bello , Jaeho Moon , Jihyong Oh , Munchurl Kim

We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is prevalent in real-world dynamic video capture, especially due to human movements in 3D human…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Xianrui Luo , Juewen Peng , Zhongang Cai , Lei Yang , Fan Yang , Zhiguo Cao , Guosheng Lin

Many works have succeeded in reconstructing Gaussian human avatars from multi-view videos. However, they either struggle to capture pose-dependent appearance details with a single MLP, or rely on a computationally intensive neural network…

图形学 · 计算机科学 2025-04-29 Youyi Zhan , Tianjia Shao , Yin Yang , Kun Zhou

The online reconstruction of dynamic scenes from multi-view streaming videos faces significant challenges in training, rendering and storage efficiency. Harnessing superior learning speed and real-time rendering capabilities, 3D Gaussian…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Qiankun Gao , Jiarui Meng , Chengxiang Wen , Jie Chen , Jian Zhang

The dominant 3D Gaussian splatting (3DGS) acceleration methods fail to properly regulate the number of Gaussians during training, causing redundant computational time overhead. In this paper, we propose FastGS, a novel, simple, and general…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Shiwei Ren , Tianci Wen , Yongchun Fang , Biao Lu

Gaussian Splatting (GS) is a novel, state-of-the-art technique for rendering points in a 3D scene by approximating their contribution to image pixels through Gaussian distributions, warranting fast training and real-time rendering. The main…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Joanna Waczyńska , Piotr Borycki , Sławomir Tadeja , Jacek Tabor , Przemysław Spurek

Reconstructing photo-realistic and topology-aware animatable human avatars from monocular videos remains challenging in computer vision and graphics. Recently, methods using 3D Gaussians to represent the human body have emerged, offering…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Haoyu Zhao , Chen Yang , Hao Wang , Xingyue Zhao , Wei Shen

Constructing vivid 3D head avatars for given subjects and realizing a series of animations on them is valuable yet challenging. This paper presents GaussianHead, which models the actional human head with anisotropic 3D Gaussians. In our…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Jie Wang , Jiu-Cheng Xie , Xianyan Li , Feng Xu , Chi-Man Pun , Hao Gao

Reconstructing and predicting dynamic 3D scenes from multi-view videos is a foundational task for robotics, AR/VR, and digital twins. Recent physics-informed Gaussian Splatting methods achieve impressive future frame extrapolation but lack…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Denis Gridusov , Maxim Popov , Sergey Kolyubin

High-quality, animatable 3D human avatar reconstruction from monocular videos offers significant potential for reducing reliance on complex hardware, making it highly practical for applications in game development, augmented reality, and…

计算机视觉与模式识别 · 计算机科学 2025-05-02 Xia Yuan , Hai Yuan , Wenyi Ge , Ying Fu , Xi Wu , Guanyu Xing

Implicit Neural Representations (INRs) employ neural networks to approximate discrete data as continuous functions. In the context of video data, such models can be utilized to transform the coordinates of pixel locations along with frame…

计算机视觉与模式识别 · 计算机科学 2026-02-19 Weronika Smolak-Dyżewska , Dawid Malarz , Kornel Howil , Jan Kaczmarczyk , Marcin Mazur , Przemysław Spurek

We leverage increasingly popular three-dimensional neural representations in order to construct a unified and consistent explanation of a collection of uncalibrated images of the human face. Our approach utilizes Gaussian Splatting, since…

计算机视觉与模式识别 · 计算机科学 2025-12-19 Haodi He , Jihun Yu , Ronald Fedkiw

Efficient neural representations for dynamic video scenes are critical for applications ranging from video compression to interactive simulations. Yet, existing methods often face challenges related to high memory usage, lengthy training…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Andrew Bond , Jui-Hsien Wang , Long Mai , Erkut Erdem , Aykut Erdem

3D Gaussian Splatting (3DGS) has gained significant attention for its high-quality rendering capabilities, ultra-fast training, and inference speeds. However, when we apply 3DGS to surface reconstruction tasks, especially in environments…

计算机视觉与模式识别 · 计算机科学 2025-03-14 Chenfeng Hou , Qi Xun Yeo , Mengqi Guo , Yongxin Su , Yanyan Li , Gim Hee Lee

3D Gaussians, as a low-level scene representation, typically involve thousands to millions of Gaussians. This makes it difficult to control the scene in ways that reflect the underlying dynamic structure, where the number of independent…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Yunchao Zhang , Guandao Yang , Leonidas Guibas , Yanchao Yang

We present PEGASUS, a method for constructing a personalized generative 3D face avatar from monocular video sources. Our generative 3D avatar enables disentangled controls to selectively alter the facial attributes (e.g., hair or nose)…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Hyunsoo Cha , Byungjun Kim , Hanbyul Joo