中文
相关论文

相关论文: Deformable Gaussian Occupancy: Decoupling Rigid an…

200 篇论文

Recent advances in 3D Gaussian diffusion models suffer from time-intensive denoising and post-denoising processing due to the massive number of Gaussian primitives, resulting in slow generation and limited scalability along sampling…

计算机视觉与模式识别 · 计算机科学 2025-11-21 Zeyuan Yin , Xiaoming Liu

Diffusion models have emerged as powerful generative priors for solving PDE-constrained inverse problems. Compared to end-to-end approaches relying on massive paired datasets, explicitly decoupling the prior distribution of physical…

数值分析 · 数学 2026-04-23 Haibo Liu , Guang Lin

We present GDFusion, a temporal fusion method for vision-based 3D semantic occupancy prediction (VisionOcc). GDFusion opens up the underexplored aspects of temporal fusion within the VisionOcc framework, focusing on both temporal cues and…

计算机视觉与模式识别 · 计算机科学 2025-04-21 Dubing Chen , Huan Zheng , Jin Fang , Xingping Dong , Xianfei Li , Wenlong Liao , Tao He , Pai Peng , Jianbing Shen

Reconstructing dynamic 3D scenes from monocular video has broad applications in AR/VR, robotics, and autonomous navigation, but often fails due to severe motion blur caused by camera and object motion. Existing methods commonly follow a…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Zhijing Wu , Longguang Wang

Feedforward 3D Gaussian Splatting (3DGS) often struggles in trajectory-based sparse-view driving scenes. Existing Gaussian repair methods mainly target optimization-based 3DGS, while diffusion-based repair is typically restricted to…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Rui Song , Tianhui Cai , Markus Gross , Xingcheng Zhou , Zewei Zhou , Zhiyu Huang , Olaf Wysocki , Jiaqi Ma

This study presents a rigid-deformation decomposition framework for vehicle collision dynamics that mitigates the spectral bias of implicit neural representations, that is, coordinate-based neural networks that directly map spatio-temporal…

计算工程、金融与科学 · 计算机科学 2025-12-18 Sanghyuk Kim , Minsik Seo , Sunwoong Yang , Namwoo Kang

3D Gaussian Splatting (3DGS) has attracted significant attention for its high-quality novel view rendering, inspiring research to address real-world challenges. While conventional methods depend on sharp images for accurate scene…

计算机视觉与模式识别 · 计算机科学 2025-05-16 Jungho Lee , Suhwan Cho , Taeoh Kim , Ho-Deok Jang , Minhyeok Lee , Geonho Cha , Dongyoon Wee , Dogyoon Lee , Sangyoun Lee

3D semantic occupancy prediction is crucial for autonomous driving perception, offering comprehensive geometric scene understanding and semantic recognition. However, existing methods struggle with geometric misalignment in view…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Xubo Zhu , Haoyang Zhang , Fei He , Rui Wu , Yanhu Shan , Wen Yang , Huai Yu

3D Gaussian Splatting (3DGS) has become a competitive approach for novel view synthesis (NVS) due to its advanced rendering efficiency through 3D Gaussian projection and blending. However, Gaussians are treated equally weighted for…

计算机视觉与模式识别 · 计算机科学 2025-08-08 Zhihao Guo , Peng Wang , Zidong Chen , Xiangyu Kong , Yan Lyu , Guanyu Gao , Liangxiu Han

Minimally invasive surgery (MIS) requires high-fidelity, real-time visual feedback of dynamic and low-texture surgical scenes. To address these requirements, we introduce FeatureEndo-4DGS (FE-4DGS), the first real time pipeline leveraging…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Kai Li , Junhao Wang , William Han , Ding Zhao

Reconstructing object deformation from a single image remains a significant challenge in computer vision and graphics. Existing methods typically rely on multi-view video to recover deformation, limiting their applicability under…

图形学 · 计算机科学 2025-09-29 Jinhyeok Kim , Jaehun Bang , Seunghyun Seo , Kyungdon Joo

Recent advances in 3D scene representations have enabled high-fidelity novel view synthesis, yet adapting to discrete scene changes and constructing interactive 3D environments remain open challenges in vision and robotics. Existing…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Wenhao Hu , Haonan Zhou , Zesheng Li , Liu Liu , Jiacheng Dong , Zhizhong Su , Gaoang Wang

Inferring the drivable area in a scene is crucial for ensuring a vehicle avoids obstacles and facilitates safe autonomous driving. In this paper, we concentrate on detecting the instantaneous free space surrounding the ego vehicle,…

机器人学 · 计算机科学 2024-07-02 Gao Xiangyu , Ding Sihao , Dasari Harshavardhan Reddy

Recent advances in 3D Gaussian Splatting (3DGS) have achieved state-of-the-art results for novel view synthesis. However, efficiently capturing high-fidelity reconstructions of specific objects within complex scenes remains a significant…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Haiyi Li , Qi Chen , Denis Kalkofen , Hsiang-Ting Chen

For intelligent transportation systems and autonomous vehicles to operate safely and efficiently, they must reliably predict the future motion and trajectory of surrounding agents within complex traffic environments. At the same time, the…

机器学习 · 计算机科学 2025-08-05 Mitch Kosieradzki , Seongjin Choi

Aligning egocentric video with wearable sensors have shown promise for human action recognition, but face practical limitations in user discomfort, privacy concerns, and scalability. We explore exocentric video with ambient sensors as a…

计算机视觉与模式识别 · 计算机科学 2025-12-24 Junho Yoon , Jaemo Jung , Hyunju Kim , Dongman Lee

Recent progress in pre-trained diffusion models and 3D generation have spurred interest in 4D content creation. However, achieving high-fidelity 4D generation with spatial-temporal consistency remains a challenge. In this work, we propose…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Yifei Zeng , Yanqin Jiang , Siyu Zhu , Yuanxun Lu , Youtian Lin , Hao Zhu , Weiming Hu , Xun Cao , Yao Yao

Dynamic scene reconstruction poses a persistent challenge in 3D vision. Deformable 3D Gaussian Splatting has emerged as an effective method for this task, offering real-time rendering and high visual fidelity. This approach decomposes a…

计算机视觉与模式识别 · 计算机科学 2025-10-22 Bing He , Yunuo Chen , Guo Lu , Qi Wang , Qunshan Gu , Rong Xie , Li Song , Wenjun Zhang

Diffusion models usher a new era of video editing, flexibly manipulating the video contents with text prompts. Despite the widespread application demand in editing human-centered videos, these models face significant challenges in handling…

计算机视觉与模式识别 · 计算机科学 2024-08-15 Xiaojing Zhong , Xinyi Huang , Xiaofeng Yang , Guosheng Lin , Qingyao Wu

Dynamic Gaussian splatting has led to impressive scene reconstruction and image synthesis advances in novel views. Existing methods, however, heavily rely on pre-computed poses and Gaussian initialization by Structure from Motion (SfM)…

计算机视觉与模式识别 · 计算机科学 2024-06-27 Hao Li , Jingfeng Li , Dingwen Zhang , Chenming Wu , Jieqi Shi , Chen Zhao , Haocheng Feng , Errui Ding , Jingdong Wang , Junwei Han