English
Related papers

Related papers: MonoLoco: Monocular 3D Pedestrian Localization and…

200 papers

We present a robust real-time LiDAR 3D object detector that leverages heteroscedastic aleatoric uncertainties to significantly improve its detection performance. A multi-loss function is designed to incorporate uncertainty estimations…

Robotics · Computer Science 2019-05-07 Di Feng , Lars Rosenbaum , Fabian Timm , Klaus Dietmayer

Accurate 3D human pose estimation from single images is possible with sophisticated deep-net architectures that have been trained on very large datasets. However, this still leaves open the problem of capturing motions for which no such…

Computer Vision and Pattern Recognition · Computer Science 2018-03-28 Helge Rhodin , Jörg Spörri , Isinsu Katircioglu , Victor Constantin , Frédéric Meyer , Erich Müller , Mathieu Salzmann , Pascal Fua

Monocular 3D object localization in driving scenes is a crucial task, but challenging due to its ill-posed nature. Estimating 3D coordinates for each pixel on the object surface holds great potential as it provides dense 2D-3D geometric…

Computer Vision and Pattern Recognition · Computer Science 2023-05-30 Zhixiang Min , Bingbing Zhuang , Samuel Schulter , Buyu Liu , Enrique Dunn , Manmohan Chandraker

Understanding humans from photographs has always been a fundamental goal of computer vision. In this thesis we have developed a hierarchy of tools that cover a wide range of topics with the objective of understanding humans from monocular…

Computer Vision and Pattern Recognition · Computer Science 2016-04-28 Edgar Simo-Serra

3D object detection with a single image is an essential and challenging task for autonomous driving. Recently, keypoint-based monocular 3D object detection has made tremendous progress and achieved great speed-accuracy trade-off. However,…

Computer Vision and Pattern Recognition · Computer Science 2021-06-15 Lei Yang , Xinyu Zhang , Li Wang , Minghan Zhu , Jun Li

Its numerous applications make multi-human 3D pose estimation a remarkably impactful area of research. Nevertheless, assuming a multiple-view system composed of several regular RGB cameras, 3D multi-pose estimation presents several…

Computer Vision and Pattern Recognition · Computer Science 2024-04-10 Daniel Rodriguez-Criado , Pilar Bachiller , George Vogiatzis , Luis J. Manso

We introduce a technique for 3D human keypoint estimation that directly models the notion of spatial uncertainty of a keypoint. Our technique employs a principled approach to modelling spatial uncertainty inspired from techniques in robust…

Computer Vision and Pattern Recognition · Computer Science 2020-12-22 Francis Williams , Or Litany , Avneesh Sud , Kevin Swersky , Andrea Tagliasacchi

The emerging trend in computer vision emphasizes developing universal models capable of simultaneously addressing multiple diverse tasks. Such universality typically requires joint training across multi-domain datasets to ensure effective…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Eunsoo Im , Changhyun Jee , Jung Kwon Lee

Monocular depth estimation is often described as an ill-posed and inherently ambiguous problem. Estimating depth from 2D images is a crucial step in scene reconstruction, 3Dobject recognition, segmentation, and detection. The problem can be…

Computer Vision and Pattern Recognition · Computer Science 2019-01-29 Amlaan Bhoi

Deep learning methods for unsupervised registration often rely on objectives that assume a uniform noise level across the spatial domain (e.g. mean-squared error loss), but noise distributions are often heteroscedastic and input-dependent…

Image and Video Processing · Electrical Eng. & Systems 2024-07-19 Xiaoran Zhang , Daniel H. Pak , Shawn S. Ahn , Xiaoxiao Li , Chenyu You , Lawrence H. Staib , Albert J. Sinusas , Alex Wong , James S. Duncan

Convolutional Neural Network based approaches for monocular 3D human pose estimation usually require a large amount of training images with 3D pose annotations. While it is feasible to provide 2D joint annotations for large corpora of…

Computer Vision and Pattern Recognition · Computer Science 2019-04-09 Ikhsanul Habibie , Weipeng Xu , Dushyant Mehta , Gerard Pons-Moll , Christian Theobalt

Detecting and localizing objects in the real 3D space, which plays a crucial role in scene understanding, is particularly challenging given only a monocular image due to the geometric information loss during imagery projection. We propose…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Zengyi Qin , Jinglu Wang , Yan Lu

Relative localization is an important ability for multiple robots to perform cooperative tasks in GPS-denied environment. This paper presents a novel autonomous positioning framework for monocular relative localization of multiple tiny…

Robotics · Computer Science 2021-09-23 Shushuai Li , Christophe De Wagter , Guido C. H. E. de Croon

Traditional and modern machine learning-based path loss models typically assume a constant prediction variance. We propose a neural network that jointly predicts the mean and link-specific variance by minimizing a Gaussian negative…

Machine Learning · Computer Science 2025-12-01 Jonathan Ethier

Monocular depth estimation, similar to other image-based tasks, is prone to erroneous predictions due to ambiguities in the image, for example, caused by dynamic objects or shadows. For this reason, pixel-wise uncertainty assessment is…

Computer Vision and Pattern Recognition · Computer Science 2025-02-11 Julia Hornauer , Amir El-Ghoussani , Vasileios Belagiannis

This paper studies 3-D distributed network localization using mixed types of local relative measurements. Each node holds a local coordinate frame without a common orientation and can only measure one type of information (relative position,…

Signal Processing · Electrical Eng. & Systems 2023-12-19 Xu Fang , Xiaolei Li , Lihua Xie

Conventional human trajectory prediction models rely on clean curated data, requiring specialized equipment or manual labeling, which is often impractical for robotic applications. The existing predictors tend to overfit to clean…

Computer Vision and Pattern Recognition · Computer Science 2025-03-06 Po-Chien Luan , Yang Gao , Celine Demonsant , Alexandre Alahi

Self-localization on a 3D map by using an inexpensive monocular camera is required to realize autonomous driving. Self-localization based on a camera often uses a convolutional neural network (CNN) that can extract local features that are…

Robotics · Computer Science 2025-12-19 Satoshi Kikuchi , Masaya Kato , Tsuyoshi Tasaki

With the advancements made in deep learning, computer vision problems like object detection and segmentation have seen a great improvement in performance. However, in many real-world applications such as autonomous driving vehicles, the…

Computer Vision and Pattern Recognition · Computer Science 2021-08-10 Kumari Deepshikha , Sai Harsha Yelleni , P. K. Srijith , C Krishna Mohan

Popular research areas like autonomous driving and augmented reality have renewed the interest in image-based camera localization. In this work, we address the task of predicting the 6D camera pose from a single RGB image in a given 3D…

Computer Vision and Pattern Recognition · Computer Science 2018-03-28 Eric Brachmann , Carsten Rother