English
Related papers

Related papers: Lift, Splat, Shoot: Encoding Images From Arbitrary…

200 papers

We present a novel multi-modal extrinsic calibration framework designed to simultaneously estimate the relative poses between event cameras, LiDARs, and RGB cameras, with particular focus on the challenging event camera calibration. Core of…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Andrea Bertogalli , Giacomo Boracchi , Luca Magri

High-Fidelity 3D scene reconstruction plays a crucial role in autonomous driving by enabling novel data generation from existing datasets. This allows simulating safety-critical scenarios and augmenting training datasets without incurring…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Pou-Chun Kung , Skanda Harisha , Ram Vasudevan , Aline Eid , Katherine A. Skinner

Surround depth estimation provides a cost-effective alternative to LiDAR for 3D perception in autonomous driving. While recent self-supervised methods explore multi-camera settings to improve scale awareness and scene coverage, they are…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Weimin Liu , Jiyuan Qiu , Wenjun Wang , Joshua H. Meng

Multi-camera systems have been shown to improve the accuracy and robustness of SLAM estimates, yet state-of-the-art SLAM systems predominantly support monocular or stereo setups. This paper presents a generic sparse visual SLAM framework…

Online scene perception and topology reasoning are critical for autonomous vehicles to understand their driving environments, particularly for mapless driving systems that endeavor to reduce reliance on costly High-Definition (HD) maps.…

Robotics · Computer Science 2025-06-27 Muleilan Pei , Jiayao Shan , Peiliang Li , Jieqi Shi , Jing Huo , Yang Gao , Shaojie Shen

We propose a framework for aligning and fusing multiple images into a single view using neural image representations (NIRs), also known as implicit or coordinate-based neural representations. Our framework targets burst images that exhibit…

Computer Vision and Pattern Recognition · Computer Science 2022-07-22 Seonghyeon Nam , Marcus A. Brubaker , Michael S. Brown

We propose VideoRFSplat, a direct text-to-3D model leveraging a video generation model to generate realistic 3D Gaussian Splatting (3DGS) for unbounded real-world scenes. To generate diverse camera poses and unbounded spatial extent of…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Hyojun Go , Byeongjun Park , Hyelin Nam , Byung-Hoon Kim , Hyungjin Chung , Changick Kim

With the development of autonomous driving technology, sensor calibration has become a key technology to achieve accurate perception fusion and localization. Accurate calibration of the sensors ensures that each sensor can function properly…

Robotics · Computer Science 2023-05-29 Jixiang Li , Jiahao Pi , Guohang Yan , Yikang Li

Empowering 3D Gaussian Splatting with generalization ability is appealing. However, existing generalizable 3D Gaussian Splatting methods are largely confined to narrow-range interpolation between stereo images due to their heavy backbones,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Yunsong Wang , Tianxin Huang , Hanlin Chen , Gim Hee Lee

Sim2Real transfer has gained popularity because it helps transfer from inexpensive simulators to real world. This paper presents a novel system that fuses components in a traditional World Model into a robust system, trained entirely within…

Robotics · Computer Science 2024-03-26 Kiran Lekkala , Chen Liu , Laurent Itti

3D Gaussian Splatting (3DGS) has emerged as a powerful technique for real-time LiDAR and camera synthesis in autonomous driving simulation. However, simulating LiDAR with 3DGS remains challenging for extrapolated views beyond the training…

Robotics · Computer Science 2026-03-17 Yiming Huang , Xin Kang , Sipeng Zhang , Hongliang Ren , Weihua Zhang , Junjie Lai

Efficient reasoning about the semantic, spatial, and temporal structure of a scene is a crucial prerequisite for autonomous driving. We present NEural ATtention fields (NEAT), a novel representation that enables such reasoning for…

Computer Vision and Pattern Recognition · Computer Science 2021-09-10 Kashyap Chitta , Aditya Prakash , Andreas Geiger

Ensuring the safety of autonomous robots, such as self-driving vehicles, requires extensive testing across diverse driving scenarios. Simulation is a key ingredient for conducting such testing in a cost-effective and scalable way. Neural…

Computer Vision and Pattern Recognition · Computer Science 2025-03-14 Georg Hess , Carl Lindström , Maryam Fatemi , Christoffer Petersson , Lennart Svensson

Recent works in autonomous driving have widely adopted the bird's-eye-view (BEV) semantic map as an intermediate representation of the world. Online prediction of these BEV maps involves non-trivial operations such as multi-camera data…

Computer Vision and Pattern Recognition · Computer Science 2023-03-01 Florent Bartoccioni , Éloi Zablocki , Andrei Bursuc , Patrick Pérez , Matthieu Cord , Karteek Alahari

Scene transfer for vision-based mobile robotics applications is a highly relevant and challenging problem. The utility of a robot greatly depends on its ability to perform a task in the real world, outside of a well-controlled lab…

Robotics · Computer Science 2024-03-01 Jiaxu Xing , Leonard Bauersfeld , Yunlong Song , Chunwei Xing , Davide Scaramuzza

In the field of autonomous driving, a variety of sensor data types exist, each representing different modalities of the same scene. Therefore, it is feasible to utilize data from other sensors to facilitate image compression. However, few…

Computer Vision and Pattern Recognition · Computer Science 2024-12-23 Yiheng Jiang , Haotian Zhang , Li Li , Dong Liu , Zhu Li

This work tackles scene understanding for outdoor robotic navigation, solely relying on images captured by an on-board camera. Conventional visual scene understanding interprets the environment based on specific descriptive categories.…

Robotics · Computer Science 2022-02-07 Galadrielle Humblot-Renaux , Letizia Marchegiani , Thomas B. Moeslund , Rikke Gade

With the proliferation of small aerial vehicles, acquiring close up aerial imagery for high quality reconstruction of complex scenes is gaining importance. We present an adaptive view planning method to collect such images in an automated…

Computer Vision and Pattern Recognition · Computer Science 2019-11-20 Cheng Peng , Volkan Isler

Holistic 3D scene understanding, which jointly models geometry, appearance, and semantics, is crucial for applications like augmented reality and robotic interaction. Existing feed-forward 3D scene understanding methods (e.g., LSM) are…

Computer Vision and Pattern Recognition · Computer Science 2025-06-16 Qijing Li , Jingxiang Sun , Liang An , Zhaoqi Su , Hongwen Zhang , Yebin Liu

Reconstructing 3D scenes from sparse viewpoints is a long-standing challenge with wide applications. Recent advances in feed-forward 3D Gaussian sparse-view reconstruction methods provide an efficient solution for real-time novel view…

Computer Vision and Pattern Recognition · Computer Science 2025-06-05 Yang Xiao , Guoan Xu , Qiang Wu , Wenjing Jia