English
Related papers

Related papers: Size Matters: Reconstructing Real-Scale 3D Models …

200 papers

Malnutrition is a multidomain problem affecting 54% of older adults in long-term care (LTC). Monitoring nutritional intake in LTC is laborious and subjective, limiting clinical inference capabilities. Recent advances in automatic…

Computer Vision and Pattern Recognition · Computer Science 2022-02-02 Kaylen J Pfisterer , Robert Amelard , Audrey G Chung , Braeden Syrnyk , Alexander MacLean , Heather H Keller , Alexander Wong

3D face reconstruction is a fundamental Computer Vision problem of extraordinary difficulty. Current systems often assume the availability of multiple facial images (sometimes from the same subject) as input, and must address a number of…

Computer Vision and Pattern Recognition · Computer Science 2017-09-11 Aaron S. Jackson , Adrian Bulat , Vasileios Argyriou , Georgios Tzimiropoulos

There have been numerous recently proposed methods for monocular depth prediction (MDP) coupled with the equally rapid evolution of benchmarking tools. However, we argue that MDP is currently witnessing benchmark over-fitting and relying on…

Computer Vision and Pattern Recognition · Computer Science 2022-03-16 Evin Pınar Örnek , Shristi Mudgal , Johanna Wald , Yida Wang , Nassir Navab , Federico Tombari

Existing inverse physics methods recover physical parameters from multi-view videos, where geometric constraints across views resolve scale and 3D structure. In monocular settings, however, such constraints are absent, leading to severe…

Computer Vision and Pattern Recognition · Computer Science 2026-05-29 Daniel Rho , Jun Myeong Choi , Matthew Thornton , Biswadip Dey , Roni Sengupta

This research aims to improve dietetic-nutritional treatment using state-of-the-art RGB-D sensors and virtual reality (VR) technology. Recent studies show that adherence to treatment can be improved using multimedia technologies. However,…

Computer Vision and Pattern Recognition · Computer Science 2020-07-03 Andrés Fuster-Guilló , Jorge Azorín-López , Marcelo Saval-Calvo , Juan Miguel Castillo-Zaragoza , Nahuel Garcia-DUrso , Robert B Fisher

Visual SLAM inside the human body will open the way to computer-assisted navigation in endoscopy. However, due to space limitations, medical endoscopes only provide monocular images, leading to systems lacking true scale. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2022-04-21 Victor M. Batlle , J. M. M. Montiel , Juan D. Tardos

Many standard robotic platforms are equipped with at least a fixed 2D laser range finder and a monocular camera. Although those platforms do not have sensors for 3D depth sensing capability, knowledge of depth is an essential part in many…

Computer Vision and Pattern Recognition · Computer Science 2016-11-08 Yiyi Liao , Lichao Huang , Yue Wang , Sarath Kodagoda , Yinan Yu , Yong Liu

Accurate height estimation from monocular aerial imagery presents a significant challenge due to its inherently ill-posed nature. This limitation is rooted in the absence of adequate geometric constraints available to the model when…

Computer Vision and Pattern Recognition · Computer Science 2023-11-07 Xiaomou Hou , Wanshui Gan , Naoto Yokoya

We introduce MapAnything, a unified transformer-based feed-forward model that ingests one or more images along with optional geometric inputs such as camera intrinsics, poses, depth, or partial reconstructions, and then directly regresses…

We propose a method for 3D object reconstruction and 6D-pose estimation from 2D images that uses knowledge about object shape as the primary key. In the proposed pipeline, recognition and labeling of objects in 2D images deliver 2D segment…

Computer Vision and Pattern Recognition · Computer Science 2022-03-03 Marcell Wolnitza , Osman Kaya , Tomas Kulvicius , Florentin Wörgötter , Babette Dellen

3D object reconstruction is a fundamental task of many robotics and AI problems. With the aid of deep convolutional neural networks (CNNs), 3D object reconstruction has witnessed a significant progress in recent years. However, possibly due…

Computer Vision and Pattern Recognition · Computer Science 2018-09-11 Hanqing Wang , Jiaolong Yang , Wei Liang , Xin Tong

This paper presents a new system to obtain dense object reconstructions along with 6-DoF poses from a single image. Geared towards high fidelity reconstruction, several recent approaches leverage implicit surface representations and deep…

Computer Vision and Pattern Recognition · Computer Science 2020-04-28 Aniket Pokale , Aditya Aggarwal , K. Madhava Krishna

Recovering the 3D structure of an object from a single image is a challenging task due to its ill-posed nature. One approach is to utilize the plentiful photos of the same object category to learn a strong 3D shape prior for the object.…

Computer Vision and Pattern Recognition · Computer Science 2021-09-08 Long-Nhat Ho , Anh Tuan Tran , Quynh Phung , Minh Hoai

Monocular cameras are extensively employed in indoor robotics, but their performance is limited in visual odometry, depth estimation, and related applications due to the absence of scale information.Depth estimation refers to the process of…

Robotics · Computer Science 2023-09-15 Yehao Liu , Ruoyan Xia , Xiaosu Xu , Zijian Wang , Yiqing Ya , Mingze Fan

Face reconstruction and tracking is a building block of numerous applications in AR/VR, human-machine interaction, as well as medical applications. Most of these applications rely on a metrically correct prediction of the shape, especially,…

Computer Vision and Pattern Recognition · Computer Science 2022-10-20 Wojciech Zielonka , Timo Bolkart , Justus Thies

A major challenge in monocular 3D object detection is the limited diversity and quantity of objects in real datasets. While augmenting real scenes with virtual objects holds promise to improve both the diversity and quantity of the objects,…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Yunhao Ge , Hong-Xing Yu , Cheng Zhao , Yuliang Guo , Xinyu Huang , Liu Ren , Laurent Itti , Jiajun Wu

3D human shape and pose estimation from monocular images has been an active area of research in computer vision, having a substantial impact on the development of new applications, from activity recognition to creating virtual avatars.…

Computer Vision and Pattern Recognition · Computer Science 2020-08-11 Xiangyu Xu , Hao Chen , Francesc Moreno-Noguer , Laszlo A. Jeni , Fernando De la Torre

Direct computer vision based-nutrient content estimation is a demanding task, due to deformation and occlusions of ingredients, as well as high intra-class and low inter-class variability between meal classes. In order to tackle these…

Information Retrieval · Computer Science 2019-11-06 Matthias Fontanellaz , Stergios Christodoulidis , Stavroula Mougiakakou

Estimating precise metric depth and scene reconstruction from monocular endoscopy is a fundamental task for surgical navigation in robotic surgery. However, traditional stereo matching adopts binocular images to perceive the depth…

Robotics · Computer Science 2022-11-29 Ruofeng Wei , Bin Li , Hangjie Mo , Fangxun Zhong , Yonghao Long , Qi Dou , Yun-Hui Liu , Dong Sun

Three-dimensional shape reconstruction of 2D landmark points on a single image is a hallmark of human vision, but is a task that has been proven difficult for computer vision algorithms. We define a feed-forward deep neural network…

Computer Vision and Pattern Recognition · Computer Science 2016-09-29 Ruiqi Zhao , Yan Wang , Aleix Martinez