English

DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes

Computer Vision and Pattern Recognition 2025-10-30 v1 Machine Learning Robotics

Abstract

Real-time, high-fidelity reconstruction of dynamic driving scenes is challenged by complex dynamics and sparse views, with prior methods struggling to balance quality and efficiency. We propose DrivingScene, an online, feed-forward framework that reconstructs 4D dynamic scenes from only two consecutive surround-view images. Our key innovation is a lightweight residual flow network that predicts the non-rigid motion of dynamic objects per camera on top of a learned static scene prior, explicitly modeling dynamics via scene flow. We also introduce a coarse-to-fine training paradigm that circumvents the instabilities common to end-to-end approaches. Experiments on nuScenes dataset show our image-only method simultaneously generates high-quality depth, scene flow, and 3D Gaussian point clouds online, significantly outperforming state-of-the-art methods in both dynamic reconstruction and novel view synthesis.

Keywords

Cite

@article{arxiv.2510.24734,
  title  = {DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes},
  author = {Qirui Hou and Wenzhang Sun and Chang Zeng and Chunfeng Wang and Hao Li and Jianxun Cui},
  journal= {arXiv preprint arXiv:2510.24734},
  year   = {2025}
}

Comments

Autonomous Driving, Novel view Synthesis, Multi task Learning

R2 v1 2026-07-01T07:10:09.633Z