English
Related papers

Related papers: Lift, Splat, Shoot: Encoding Images From Arbitrary…

200 papers

In agricultural automation, inherent occlusion presents a major challenge for robotic harvesting. We propose a novel imitation learning-based viewpoint planning approach to actively adjust camera viewpoint and capture unobstructed images of…

Robotics · Computer Science 2025-03-14 Lun Li , Hamidreza Kasaei

Prevailing 2D-centric visuomotor policies exhibit a pronounced deficiency in novel view generalization, as their reliance on static observations hinders consistent action mapping across unseen views. In response, we introduce GenSplat, a…

Robotics · Computer Science 2026-04-01 Sen Wang , Huaiyi Dong , Jingyi Tian , Jiayi Li , Zhuo Yang , Tongtong Cao , Anlin Chen , Shuang Wu , Le Wang , Sanping Zhou

Robust and realistic rendering for large-scale road scenes is essential in autonomous driving simulation. Recently, 3D Gaussian Splatting (3D-GS) has made groundbreaking progress in neural rendering, but the general fidelity of large-scale…

Computer Vision and Pattern Recognition · Computer Science 2024-08-28 Saining Zhang , Baijun Ye , Xiaoxue Chen , Yuantao Chen , Zongzheng Zhang , Cheng Peng , Yongliang Shi , Hao Zhao

We present a method to map 2D image observations of a scene to a persistent 3D scene representation, enabling novel view synthesis and disentangled representation of the movable and immovable components of the scene. Motivated by the…

Computer Vision and Pattern Recognition · Computer Science 2023-04-11 Prafull Sharma , Ayush Tewari , Yilun Du , Sergey Zakharov , Rares Ambrus , Adrien Gaidon , William T. Freeman , Fredo Durand , Joshua B. Tenenbaum , Vincent Sitzmann

Semantic segmentation in bird's eye view (BEV) is an important task for autonomous driving. Though this task has attracted a large amount of research efforts, it is still challenging to flexibly cope with arbitrary (single or multiple)…

Computer Vision and Pattern Recognition · Computer Science 2022-08-19 Lang Peng , Zhirong Chen , Zhangjie Fu , Pengpeng Liang , Erkang Cheng

Aerial cinematography is revolutionizing industries that require live and dynamic camera viewpoints such as entertainment, sports, and security. However, safely piloting a drone while filming a moving target in the presence of obstacles is…

Prior point cloud provides 3D environmental context, which enhances the capabilities of monocular camera in downstream vision tasks, such as 3D object detection, via data fusion. However, the absence of accurate and automated registration…

Robotics · Computer Science 2024-04-09 Yu Sheng , Lu Zhang , Xingchen Li , Yifan Duan , Yanyong Zhang , Yu Zhang , Jianmin Ji

We propose FreeSim, a camera simulation method for autonomous driving. FreeSim emphasizes high-quality rendering from viewpoints beyond the recorded ego trajectories. In such viewpoints, previous methods have unacceptable degradation…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 Lue Fan , Hao Zhang , Qitai Wang , Hongsheng Li , Zhaoxiang Zhang

The well-established modular autonomous driving system is decoupled into different standalone tasks, e.g. perception, prediction and planning, suffering from information loss and error accumulation across modules. In contrast, end-to-end…

Computer Vision and Pattern Recognition · Computer Science 2024-06-03 Wenchao Sun , Xuewu Lin , Yining Shi , Chuang Zhang , Haoran Wu , Sifa Zheng

Spatial scene-understanding, including dense depth and ego-motion estimation, is an important problem in computer vision for autonomous vehicles and advanced driver assistance systems. Thus, it is beneficial to design perception modules…

Computer Vision and Pattern Recognition · Computer Science 2023-02-03 Hemang Chawla , Matti Jukola , Shabbir Marzban , Elahe Arani , Bahram Zonooz

Single image 3D photography enables viewers to view a still image from novel viewpoints. Recent approaches combine monocular depth networks with inpainting networks to achieve compelling results. A drawback of these techniques is the use of…

Computer Vision and Pattern Recognition · Computer Science 2021-09-03 Varun Jampani , Huiwen Chang , Kyle Sargent , Abhishek Kar , Richard Tucker , Michael Krainin , Dominik Kaeser , William T. Freeman , David Salesin , Brian Curless , Ce Liu

Feed-forward paradigms for 3D reconstruction have become a focus of recent research, which learn implicit, fixed view transformations to generate a single scene representation. However, their application to complex driving scenes reveals…

Computer Vision and Pattern Recognition · Computer Science 2025-11-27 Haochen Yu , Qiankun Liu , Hongyuan Liu , Jianfei Jiang , Juntao Lyu , Jiansheng Chen , Huimin Ma

Bird's-eye-view (BEV) grid is a common representation for the perception of road components, e.g., drivable area, in autonomous driving. Most existing approaches rely on cameras only to perform segmentation in BEV space, which is…

Computer Vision and Pattern Recognition · Computer Science 2022-11-01 Shubhankar Borse , Marvin Klingner , Varun Ravi Kumar , Hong Cai , Abdulaziz Almuzairee , Senthil Yogamani , Fatih Porikli

Learning powerful representations in bird's-eye-view (BEV) for perception tasks is trending and drawing extensive attention both from industry and academia. Conventional approaches for most autonomous driving algorithms perform detection,…

Feed-forward 3D reconstruction for autonomous driving has advanced rapidly, yet existing methods struggle with the joint challenges of sparse, non-overlapping camera views and complex scene dynamics. We present UniSplat, a general…

Computer Vision and Pattern Recognition · Computer Science 2025-11-07 Chen Shi , Shaoshuai Shi , Xiaoyang Lyu , Chunyang Liu , Kehua Sheng , Bo Zhang , Li Jiang

LiDAR sensors provide rich 3D information about their surrounding{s} and are becoming increasingly important for autonomous vehicles tasks such as {localization}, semantic segmentation, object detection, and tracking. {Simulation}…

Robotics · Computer Science 2022-12-27 Jean Pierre Richa , Jean-Emmanuel Deschaud , François Goulette , Nicolas Dalmasso

Reconstructing dynamic driving scenes is essential for developing autonomous systems through sensor-realistic simulation. Although recent methods achieve high-fidelity reconstructions, they either rely on costly human annotations for object…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Carl Lindström , Mahan Rafidashti , Maryam Fatemi , Lars Hammarstrand , Martin R. Oswald , Lennart Svensson

The current autonomous driving architecture places a heavy burden in signal processing for the graphics processing units (GPUs) in the car. This directly translates into battery drain and lower energy efficiency, crucial factors in electric…

Artificial Intelligence · Computer Science 2018-11-02 Nalin Jayaweera , Nandana Rajatheva , Matti Latva-aho

In this work, we present an effective multi-view approach to closed-loop end-to-end learning of precise manipulation tasks that are 3D in nature. Our method learns to accomplish these tasks using multiple statically placed but uncalibrated…

Robotics · Computer Science 2021-04-02 Iretiayo Akinola , Jacob Varley , Dmitry Kalashnikov

While most recent autonomous driving system focuses on developing perception methods on ego-vehicle sensors, people tend to overlook an alternative approach to leverage intelligent roadside cameras to extend the perception ability beyond…

Computer Vision and Pattern Recognition · Computer Science 2023-04-12 Lei Yang , Kaicheng Yu , Tao Tang , Jun Li , Kun Yuan , Li Wang , Xinyu Zhang , Peng Chen
‹ Prev 1 3 4 5 6 7 10 Next ›