中文
相关论文

相关论文: Space-Time Forecasting of Dynamic Scenes with Moti…

200 篇论文

Accurate and robust LiDAR 3D object detection is essential for comprehensive scene understanding in autonomous driving. Despite its importance, LiDAR detection performance is limited by inherent constraints of point cloud data, particularly…

计算机视觉与模式识别 · 计算机科学 2024-09-09 Rui Yu , Runkai Zhao , Cong Nie , Heng Wang , HuaiCheng Yan , Meng Wang

We present an approach for high-quality dynamic Gaussian Splatting from monocular videos. To this end, we in this work go one step further beyond previous methods to explicitly model continuous position and orientation deformation of…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Xuankai Zhang , Junjin Xiao , Shangwei Huang , Wei-shi Zheng , Qing Zhang

While diffusion models can successfully generate data and make predictions, they are predominantly designed for static images. We propose an approach for efficiently training diffusion models for probabilistic spatiotemporal forecasting,…

机器学习 · 计算机科学 2023-10-12 Salva Rühling Cachay , Bo Zhao , Hailey Joren , Rose Yu

Monocular scene flow estimation aims to recover dense 3D motion from image sequences, yet most existing methods are limited to two-frame inputs, restricting temporal modeling and robustness to occlusions. We propose RAFT-MSF++, a…

计算机视觉与模式识别 · 计算机科学 2026-04-22 Xunpei Sun , Zuoxun Hou , Yi Chang , Gang Chen , Wei-Shi Zheng

3D occupancy prediction is critical for comprehensive scene understanding in vision-centric autonomous driving. Recent advances have explored utilizing 3D semantic Gaussians to model occupancy while reducing computational overhead, but they…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Xiaoyang Yan , Muleilan Pei , Shaojie Shen

Dynamic scene rendering has taken a leap forward with the rise of 4D Gaussian Splatting, but there's still one elusive challenge: how to make 3D Gaussians move through time as naturally as they would in the real world, all while keeping the…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Junli Deng , Yihao Luo

Visual SLAM algorithms achieve significant improvements through the exploration of 3D Gaussian Splatting (3DGS) representations, particularly in generating high-fidelity dense maps. However, they depend on a static environment assumption…

机器人学 · 计算机科学 2026-04-15 Yi Liu , Haoxuan Xu , Hongbo Duan , Keyu Fan , Zhengyang Zhang , Peiyu Zhuang , Pengting Luo , Houde Liu

Multi-camera systems are increasingly vital in the environmental perception of autonomous vehicles and robotics. Their physical configuration offers inherent fixed relative pose constraints that benefit Structure-from-Motion (SfM). However,…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Peilin Tao , Hainan Cui , Diantao Tu , Shuhan Shen

3D Gaussian splatting (3DGS) has recently emerged as an alternative representation that leverages a 3D Gaussian-based representation and introduces an approximated volumetric rendering, achieving very fast rendering speed and promising…

计算机视觉与模式识别 · 计算机科学 2024-08-08 Joo Chan Lee , Daniel Rho , Xiangyu Sun , Jong Hwan Ko , Eunbyung Park

This paper presents a unified framework that allows high-quality dynamic Gaussian Splatting from both defocused and motion-blurred monocular videos. Due to the significant difference between the formation processes of defocus blur and…

计算机视觉与模式识别 · 计算机科学 2025-11-03 Xuankai Zhang , Junjin Xiao , Qing Zhang

Performing language-conditioned robotic manipulation tasks in unstructured environments is highly demanded for general intelligent robots. Conventional robotic manipulation methods usually learn semantic representation of the observation…

机器人学 · 计算机科学 2024-07-19 Guanxing Lu , Shiyi Zhang , Ziwei Wang , Changliu Liu , Jiwen Lu , Yansong Tang

Forecasting human activities observed in videos is a long-standing challenge in computer vision, which leads to various real-world applications such as mobile robots, autonomous driving, and assistive systems. In this work, we present a new…

计算机视觉与模式识别 · 计算机科学 2019-11-25 Hiroaki Minoura , Ryo Yonetani , Mai Nishimura , Yoshitaka Ushiku

Accurate, long-term forecasting of pedestrian trajectories in highly dynamic and interactive scenes is a long-standing challenge. Recent advances in using data-driven approaches have achieved significant improvements in terms of prediction…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Rui Zhou , Hongyu Zhou , Huidong Gao , Masayoshi Tomizuka , Jiachen Li , Zhuo Xu

Mobile manipulators in households must both navigate and manipulate. This requires a compact, semantically rich scene representation that captures where objects are, how they function, and which parts are actionable. Scene graphs are a…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Yuanchen Ju , Yongyuan Liang , Yen-Jen Wang , Nandiraju Gireesh , Yuanliang Ju , Seungjae Lee , Qiao Gu , Elvis Hsieh , Furong Huang , Koushil Sreenath

Visualization of large-scale time-dependent simulation data is crucial for domain scientists to analyze complex phenomena, but it demands significant I/O bandwidth, storage, and computational resources. To enable effective visualization on…

图形学 · 计算机科学 2025-07-18 Siyuan Yao , Chaoli Wang

Humans excel at forecasting the future dynamics of a scene given just a single image. Video generation models that can mimic this ability are an essential component for intelligent systems. Recent approaches have improved temporal coherence…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Melonie de Almeida , Daniela Ivanova , Tong Shi , John H. Williamson , Paul Henderson

Gaussian Splatting and its dynamic extensions are effective for reconstructing 3D scenes from 2D images when there is significant camera movement to facilitate motion parallax and when scene objects remain relatively static. However, in…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Liuyue Xie , Joel Julin , Koichiro Niinuma , Laszlo A. Jeni

Gaussian splatting demonstrates proficiency for 3D scene modeling but suffers from substantial data volume due to inherent primitive redundancy. To enable future photorealistic 3D immersive visual communication applications, significant…

图形学 · 计算机科学 2025-04-18 Xiangrui Liu , Xinju Wu , Shiqi Wang , Zhu Li , Sam Kwong

Implicit neural representation has paved the way for new approaches to dynamic scene reconstruction and rendering. Nonetheless, cutting-edge dynamic neural rendering methods rely heavily on these implicit representations, which frequently…

计算机视觉与模式识别 · 计算机科学 2023-11-22 Ziyi Yang , Xinyu Gao , Wen Zhou , Shaohui Jiao , Yuqing Zhang , Xiaogang Jin

This paper tackles the challenge of recovering 4D dynamic scenes from videos captured by as few as four portable cameras. Learning to model scene dynamics for temporally consistent novel-view rendering is a foundational task in computer…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Junsheng Zhou , Zhifan Yang , Liang Han , Wenyuan Zhang , Kanle Shi , Shenkun Xu , Yu-Shen Liu