English
Related papers

Related papers: FisheyeDepth: A Real Scale Self-Supervised Depth E…

200 papers

Modern robotic manipulation primarily relies on visual observations in a 2D color space for skill learning but suffers from poor generalization. In contrast, humans, living in a 3D world, depend more on physical properties-such as distance,…

Moving Object Detection (MOD) is an important task for achieving robust autonomous driving. An autonomous vehicle has to estimate collision risk with other interacting objects in the environment and calculate an optional trajectory.…

Computer Vision and Pattern Recognition · Computer Science 2019-09-02 Marie Yahiaoui , Hazem Rashed , Letizia Mariotti , Ganesh Sistu , Ian Clancy , Lucie Yahiaoui , Varun Ravi Kumar , Senthil Yogamani

Depth information is useful for many applications. Active depth sensors are appealing because they obtain dense and accurate depth maps. However, due to issues that range from power constraints to multi-sensor interference, these sensors…

Image and Video Processing · Electrical Eng. & Systems 2020-02-04 James Noraky , Vivienne Sze

The development in the field of autonomous driving goes hand in hand with ever new developments in the field of image processing and machine learning methods. In order to fully exploit the advantages of deep learning, it is necessary to…

Computer Vision and Pattern Recognition · Computer Science 2020-11-12 Tobias Scheck , Adarsh Mallandur , Christian Wiede , Gangolf Hirtz

Accurate 3D geometric perception is an important prerequisite for a wide range of spatial AI systems. While state-of-the-art methods depend on large-scale training data, acquiring consistent and precise 3D annotations from in-the-wild…

In this paper we present a novel self-supervised method to anticipate the depth estimate for a future, unobserved real-world urban scene. This work is the first to explore self-supervised learning for estimation of monocular depth of future…

Computer Vision and Pattern Recognition · Computer Science 2022-07-12 Sauradip Nag , Nisarg Shah , Anran Qi , Raghavendra Ramachandra

The perception of transparent objects for grasp and manipulation remains a major challenge, because existing robotic grasp methods which heavily rely on depth maps are not suitable for transparent objects due to their unique visual…

Computer Vision and Pattern Recognition · Computer Science 2024-05-27 Yifan Zhou , Wanli Peng , Zhongyu Yang , He Liu , Yi Sun

Various depth estimation models are now widely used on many mobile and IoT devices for image segmentation, bokeh effect rendering, object tracking and many other mobile tasks. Thus, it is very crucial to have efficient and accurate depth…

Autonomous vehicles and robots need to operate over a wide variety of scenarios in order to complete tasks efficiently and safely. Multi-camera self-supervised monocular depth estimation from videos is a promising way to reason about the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-08 Takayuki Kanai , Igor Vasiljevic , Vitor Guizilini , Adrien Gaidon , Rares Ambrus

Bokeh rendering and depth estimation share a fundamental optical connection, yet existing methods fail to fully exploit this reciprocity. Conventional bokeh pipelines rely heavily on noisy depth maps that inevitably introduce visual…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Hangwei Zhang , Armando Fortes , Tianyi Wei , Xingang Pan

This paper presents a novel self-supervised two-frame multi-camera metric depth estimation network, termed M${^2}$Depth, which is designed to predict reliable scale-aware surrounding depth in autonomous driving. Unlike the previous works…

Computer Vision and Pattern Recognition · Computer Science 2024-05-06 Yingshuang Zou , Yikang Ding , Xi Qiu , Haoqian Wang , Haotian Zhang

The development of large-scale 3D scene reconstruction and novel view synthesis methods mostly rely on datasets comprising perspective images with narrow fields of view (FoV). While effective for small-scale scenes, these datasets require…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Ulas Gunes , Matias Turkulainen , Xuqian Ren , Arno Solin , Juho Kannala , Esa Rahtu

While most recent autonomous driving system focuses on developing perception methods on ego-vehicle sensors, people tend to overlook an alternative approach to leverage intelligent roadside cameras to extend the perception ability beyond…

Computer Vision and Pattern Recognition · Computer Science 2023-04-12 Lei Yang , Kaicheng Yu , Tao Tang , Jun Li , Kun Yuan , Li Wang , Xinyu Zhang , Peng Chen

In a human-robot collaborative task where a robot helps its partner by finding described objects, the depth dimension plays a critical role in successful task completion. Existing studies have mostly focused on comprehending the object…

Robotics · Computer Science 2021-07-13 Fethiye Irmak Dogan , Iolanda Leite

Per-pixel ground-truth depth data is challenging to acquire at scale. To overcome this limitation, self-supervised learning has emerged as a promising alternative for training models to perform monocular depth estimation. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Clément Godard , Oisin Mac Aodha , Michael Firman , Gabriel Brostow

Humans naturally perceive a 3D scene in front of them through accumulation of information obtained from multiple interconnected projections of the scene and by interpreting their correspondence. This phenomenon has inspired artificial…

Computer Vision and Pattern Recognition · Computer Science 2018-11-20 Amirreza Farnoosh , Sarah Ostadabbas

Following the successful application of deep convolutional neural networks to 2d human pose estimation, the next logical problem to solve is 3d human pose estimation from monocular images. While previous solutions have shown some success,…

Computer Vision and Pattern Recognition · Computer Science 2021-03-04 Alec Diaz-Arias , Mitchell Messmore , Dmitriy Shin , Stephen Baek

Self-supervised monocular depth estimation has seen significant progress in recent years, especially in outdoor environments. However, depth prediction results are not satisfying in indoor scenes where most of the existing data are captured…

Computer Vision and Pattern Recognition · Computer Science 2022-07-20 Runze Li , Pan Ji , Yi Xu , Bir Bhanu

Capturing large fields of view with only one camera is an important aspect in surveillance and automotive applications, but the wide-angle fisheye imagery thus obtained exhibits very special characteristics that may not be very well suited…

Image and Video Processing · Electrical Eng. & Systems 2022-12-01 Andrea Eichenseer , Michel Bätz , Jürgen Seiler , André Kaup

Recently, 3D Gaussian Splatting (3DGS) has garnered attention for its high fidelity and real-time rendering. However, adapting 3DGS to different camera models, particularly fisheye lenses, poses challenges due to the unique 3D to 2D…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Zimu Liao , Siyan Chen , Rong Fu , Yi Wang , Zhongling Su , Hao Luo , Li Ma , Linning Xu , Bo Dai , Hengjie Li , Zhilin Pei , Xingcheng Zhang