English
Related papers

Related papers: StreetForward: Perceiving Dynamic Street with Feed…

200 papers

Synthesizing novel views for urban environments is crucial for tasks like autonomous driving and virtual tours. Compared to object-level or indoor situations, outdoor settings present unique challenges, such as inconsistency across frames…

Computer Vision and Pattern Recognition · Computer Science 2023-11-30 Mreenav Shyam Deka , Lu Sang , Daniel Cremers

Ego-centric driving videos available online provide an abundant source of visual data for autonomous driving, yet their lack of annotations makes it difficult to learn representations that capture both semantic structure and 3D geometry.…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Matthew Strong , Wei-Jer Chang , Quentin Herau , Jiezhi Yang , Yihan Hu , Chensheng Peng , Wei Zhan

Video Anomaly Detection (VAD) automatically identifies anomalous events from video, mitigating the need for human operators in large-scale surveillance deployments. However, two fundamental obstacles hinder real-world adoption: domain…

Computer Vision and Pattern Recognition · Computer Science 2025-05-26 Hyogun Lee , Haksub Kim , Ig-Jae Kim , Yonghun Choi

Legged robots with egocentric forward-facing depth cameras can couple exteroception and proprioception to achieve robust forward agility on complex terrain. When these robots walk backward, the forward-only field of view provides no…

Robotics · Computer Science 2026-03-04 Shixin Luo , Songbo Li , Yuan Hao , Yaqi Wang , Jun Zheng , Jun Wu , Qiuguo Zhu

This paper introduces a general approach to dynamic scene reconstruction from multiple moving cameras without prior knowledge or limiting constraints on the scene structure, appearance, or illumination. Existing techniques for dynamic scene…

Computer Vision and Pattern Recognition · Computer Science 2015-10-01 Armin Mustafa , Hansung Kim , Jean-Yves Guillemaut , Adrian Hilton

High-quality 3D scene reconstruction has recently advanced toward generalizable feed-forward architectures, enabling the generation of complex environments in a single forward pass. However, despite their strong performance in static scene…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Kaixin Zhu , Yiwen Tang , Yifan Yang , Renrui Zhang , Bohan Zeng , Ziyu Guo , Ruichuan An , Zhou Liu , Qizhi Chen , Delin Qu , Jaehong Yoon , Wentao Zhang

Feed-forward 3D reconstruction from sparse, low-resolution (LR) images is a crucial capability for real-world applications, such as autonomous driving and embodied AI. However, existing methods often fail to recover fine texture details.…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Xinyuan Hu , Changyue Shi , Chuxiao Yang , Minghao Chen , Jiajun Ding , Tao Wei , Chen Wei , Zhou Yu , Min Tan

LiDAR scene flow is the task of estimating per-point 3D motion between consecutive point clouds. Recent methods achieve centimeter-level accuracy on popular autonomous vehicle (AV) datasets, but are typically only trained and evaluated on a…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Siyi Li , Qingwen Zhang , Ishan Khatri , Kyle Vedder , Eric Eaton , Deva Ramanan , Neehar Peri

Panoramic imagery offers a full 360{\deg} field of view and is increasingly common in consumer devices. However, it introduces non-pinhole distortions that challenge joint pose estimation and 3D reconstruction. Existing feed-forward models,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Yijing Guo , Mengjun Chao , Luo Wang , Tianyang Zhao , Haizhao Dai , Yingliang Zhang , Jingyi Yu , Yujiao Shi

We introduce a fully automatic pipeline for dynamic scene reconstruction from casually captured monocular RGB videos. Rather than designing a new scene representation, we enhance the priors that drive Dynamic Gaussian Splatting. Video…

Computer Vision and Pattern Recognition · Computer Science 2025-12-15 Meng-Li Shih , Ying-Huan Chen , Yu-Lun Liu , Brian Curless

We introduce Diff4Splat, a feed-forward method that synthesizes controllable and explicit 4D scenes from a single image. Our approach unifies the generative priors of video diffusion models with geometry and motion constraints learned from…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Panwang Pan , Chenguo Lin , Jingjing Zhao , Chenxin Li , Yuchen Lin , Haopeng Li , Honglei Yan , Kairun Wen , Yunlong Lin , Yixuan Yuan , Yadong Mu

Recovering sharp video sequence from a motion-blurred image is highly ill-posed due to the significant loss of motion information in the blurring process. For event-based cameras, however, fast motion can be captured as events at high time…

Computer Vision and Pattern Recognition · Computer Science 2020-04-14 Zhe Jiang , Yu Zhang , Dongqing Zou , Jimmy Ren , Jiancheng Lv , Yebin Liu

We address the challenging problem of dense dynamic scene reconstruction and camera pose estimation from multiple freely moving cameras -- a setting that arises naturally when multiple observers capture a shared event. Prior approaches…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Shuo Sun , Unal Artan , Malcolm Mielle , Achim J. Lilienthaland , Martin Magnusson

We present DriveGen3D, a novel framework for generating high-quality and highly controllable dynamic 3D driving scenes that addresses critical limitations in existing methodologies. Current approaches to driving scene synthesis either…

The static world assumption is standard in most simultaneous localisation and mapping (SLAM) algorithms. Increased deployment of autonomous systems to unstructured dynamic environments is driving a need to identify moving objects and…

Robotics · Computer Science 2020-02-25 Mina Henein , Jun Zhang , Robert Mahony , Viorela Ila

Effective environment modeling is the foundation for autonomous driving, underpinning tasks from perception to planning. However, current paradigms often inadequately consider the feedback of ego motion to the observation, which leads to an…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Mingzhe Guo , Yixiang Yang , Chuanrong Han , Rufeng Zhang , Shirui Li , Ji Wan , Zhipeng Zhang

Recent advances in vision foundation models have revolutionized geometry reconstruction and semantic understanding. Yet, most of the existing approaches treat these capabilities in isolation, leading to redundant pipelines and compounded…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Chaoyi Zhou , Run Wang , Feng Luo , Mert D. Pesé , Zhiwen Fan , Yiqi Zhong , Siyu Huang

Self-driving vehicles rely on multimodal motion forecasts to effectively interact with their environment and plan safe maneuvers. We introduce SceneMotion, an attention-based model for forecasting scene-wide motion modes of multiple traffic…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Royden Wagner , Ömer Sahin Tas , Marlon Steiner , Fabian Konstantinidis , Hendrik Königshof , Marvin Klemp , Carlos Fernandez , Christoph Stiller

We present VGGT-SLAM 2.0, a real-time RGB feed-forward SLAM system which substantially improves upon VGGT-SLAM for incrementally aligning submaps created from VGGT. Firstly, we remove high-dimensional 15-degree-of-freedom drift and planar…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Dominic Maggio , Luca Carlone

Current methods for dense 3D point tracking in dynamic scenes typically rely on pairwise processing, require known camera poses, or assume temporal ordering of input frames, thereby constraining their flexibility and applicability.…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Vivek Alumootil , Tuan-Anh Vu