中文
相关论文

相关论文: Every Camera Effect, Every Time, All at Once: 4D G…

200 篇论文

Reconstructing dynamic 3D scenes from monocular videos is a fundamental yet highly challenging task, as real-world motions often involve both long-term smooth transformations and short-term complex deformations. Existing methods either…

计算机视觉与模式识别 · 计算机科学 2026-05-25 Chenyu Wu , Wanhua Li , Zhu-Tian Chen , Hanspeter Pfister

Understanding and predicting dynamics of the physical world can enhance a robot's ability to plan and interact effectively in complex environments. While recent video generation models have shown strong potential in modeling dynamic scenes,…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Zeyi Liu , Shuang Li , Eric Cousineau , Siyuan Feng , Benjamin Burchfiel , Shuran Song

Visuomotor policies learned from teleoperated demonstrations face challenges such as lengthy data collection, high costs, and limited data diversity. Existing approaches address these issues by augmenting image observations in RGB space or…

机器人学 · 计算机科学 2025-04-18 Sizhe Yang , Wenye Yu , Jia Zeng , Jun Lv , Kerui Ren , Cewu Lu , Dahua Lin , Jiangmiao Pang

Existing 4D Gaussian Splatting (4DGS) methods struggle to accurately reconstruct dynamic scenes, often failing to resolve ambiguous pixel correspondences and inadequate densification in dynamic regions. We address these issues by…

计算机视觉与模式识别 · 计算机科学 2025-11-21 Taeho Kang , Jaeyeon Park , Kyungjin Lee , Youngki Lee

Existing video generation models excel at producing photo-realistic videos from text or images, but often lack physical plausibility and 3D controllability. To overcome these limitations, we introduce PhysCtrl, a novel framework for…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Chen Wang , Chuhao Chen , Yiming Huang , Zhiyang Dou , Yuan Liu , Jiatao Gu , Lingjie Liu

Recent advances in driving-scene generation and reconstruction have demonstrated significant potential for enhancing autonomous driving systems by producing scalable and controllable training data. Existing generation methods primarily…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Ziyue Zhu , Zhanqian Wu , Zhenxin Zhu , Lijun Zhou , Haiyang Sun , Bing Wan , Kun Ma , Guang Chen , Hangjun Ye , Jin Xie , jian Yang

In this work, we introduce a novel approach for creating controllable dynamics in 3D-generated Gaussians using casually captured reference videos. Our method transfers the motion of objects from reference videos to a variety of generated 3D…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Zhoujie Fu , Jiacheng Wei , Wenhao Shen , Chaoyue Song , Xiaofeng Yang , Fayao Liu , Xulei Yang , Guosheng Lin

We present GP-4DGS, a novel framework that integrates Gaussian Processes (GPs) into 4D Gaussian Splatting (4DGS) for principled probabilistic modeling of dynamic scenes. While existing 4DGS methods focus on deterministic reconstruction,…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Mijeong Kim , Jungtaek Kim , Bohyung Han

Deformable 3D Gaussian Splatting (3D-GS) is limited by missing intermediate motion information due to the low temporal resolution of RGB cameras. To address this, we introduce the first approach combining event cameras, which capture…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Wenhao Xu , Wenming Weng , Yueyi Zhang , Ruikang Xu , Zhiwei Xiong

360-degree visual content is widely shared on platforms such as YouTube and plays a central role in virtual reality, robotics, and autonomous navigation. However, consumer-grade dual-fisheye systems consistently yield imperfect panoramas…

计算机视觉与模式识别 · 计算机科学 2025-08-28 Changha Shin , Woong Oh Cho , Seon Joo Kim

Gaussian Splatting has been considered as a novel way for view synthesis of dynamic scenes, which shows great potential in AIoT applications such as digital twins. However, recent dynamic Gaussian Splatting methods significantly degrade…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Yiwei Li , Jiannong Cao , Penghui Ruan , Divya Saxena , Songye Zhu , Yinfeng Cao

3D Gaussian Splatting has recently emerged as a powerful representation that can synthesize remarkable novel views using consistent multi-view images as input. However, we notice that images captured in dark environments where the scenes…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Sheng Ye , Zhen-Hui Dong , Yubin Hu , Yu-Hui Wen , Yong-Jin Liu

3D Gaussian Splatting (3DGS) allows flexible adjustments to scene representation, enabling continuous optimization of scene quality during dense visual simultaneous localization and mapping (SLAM) in static environments. However, 3DGS faces…

机器人学 · 计算机科学 2024-11-26 Long Wen , Shixin Li , Yu Zhang , Yuhong Huang , Jianjie Lin , Fengjunjie Pan , Zhenshan Bing , Alois Knoll

Transforming casually captured, monocular videos into fully immersive dynamic experiences is a highly ill-posed task, and comes with significant challenges, e.g., reconstructing unseen regions, and dealing with the ambiguity in monocular…

图形学 · 计算机科学 2026-04-08 Denis Rozumny , Jonathon Luiten , Numair Khan , Johannes Schönberger , Peter Kontschieder

Dynamic Gaussian Splatting approaches have achieved remarkable performance for 4D scene reconstruction. However, these approaches rely on dense-frame video sequences for photorealistic reconstruction. In real-world scenarios, due to…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Changyue Shi , Chuxiao Yang , Xinyuan Hu , Minghao Chen , Wenwen Pan , Yan Yang , Jiajun Ding , Zhou Yu , Jun Yu

3D Gaussian Splatting (3DGS) enables real-time novel view synthesis with high visual quality. However, existing methods struggle with semi-transparent specular surfaces that exhibit both complex reflections and clear transmission, often…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Ji Shi , Xianghua Ying , Bowei Xing , Ruohao Guo , Wenzhen Yue

We present DC-Gaussian, a new method for generating novel views from in-vehicle dash cam videos. While neural rendering techniques have made significant strides in driving scenarios, existing methods are primarily designed for videos…

计算机视觉与模式识别 · 计算机科学 2024-11-07 Linhan Wang , Kai Cheng , Shuo Lei , Shengkun Wang , Wei Yin , Chenyang Lei , Xiaoxiao Long , Chang-Tien Lu

Recent 4D Gaussian Splatting (4DGS) methods achieve impressive dynamic scene reconstruction but often rely on piecewise linear velocity approximations and short temporal windows. This disjointed modeling leads to severe temporal…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Suwoong Yeom , Joonsik Nam , Seunggyu Choi , Lucas Yunkyu Lee , Sangmin Kim , Jaesik Park , Joonsoo Kim , Kugjin Yun , Kyeongbo Kong , Sukju Kang

Dynamic 3D scene representation and novel view synthesis are crucial for enabling immersive experiences required by AR/VR and metaverse applications. It is a challenging task due to the complexity of unconstrained real-world scenes and…

计算机视觉与模式识别 · 计算机科学 2025-08-07 Zeyu Yang , Zijie Pan , Xiatian Zhu , Li Zhang , Jianfeng Feng , Yu-Gang Jiang , Philip H. S. Torr

Predicting physical dynamics from raw visual data remains a major challenge in AI. While recent video generation models have achieved impressive visual quality, they still cannot consistently generate physically plausible videos due to a…

计算机视觉与模式识别 · 计算机科学 2026-02-13 Shiqian Li , Ruihong Shen , Junfeng Ni , Chang Pan , Chi Zhang , Yixin Zhu