English
Related papers

Related papers: Parcel3D: Shape Reconstruction from Single RGB Ima…

200 papers

In the realm of robotic grasping, achieving accurate and reliable interactions with the environment is a pivotal challenge. Traditional methods of grasp planning methods utilizing partial point clouds derived from depth image often suffer…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Lei Zhou , Haozhe Wang , Zhengshen Zhang , Zhiyang Liu , Francis EH Tay , adn Marcelo H. Ang.

We present Spann3R, a novel approach for dense 3D reconstruction from ordered or unordered image collections. Built on the DUSt3R paradigm, Spann3R uses a transformer-based architecture to directly regress pointmaps from images without any…

Computer Vision and Pattern Recognition · Computer Science 2024-08-30 Hengyi Wang , Lourdes Agapito

Instance-level object re-identification is a fundamental computer vision task, with applications from image retrieval to intelligent monitoring and fraud detection. In this work, we propose the novel task of damaged object…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Luca Piano , Filippo Gabriele Pratticò , Alessandro Sebastian Russo , Lorenzo Lanari , Lia Morra , Fabrizio Lamberti

Reconstructing three-dimensional (3D) scenes with semantic understanding is vital in many robotic applications. Robots need to identify which objects, along with their positions and shapes, to manipulate them precisely with given tasks.…

Robotics · Computer Science 2024-12-17 Khang Nguyen , Tuan Dang , Manfred Huber

We study end-to-end learning strategies for 3D shape inference from images, in particular from a single image. Several approaches in this direction have been investigated that explore different shape representations and suitable learning…

Computer Vision and Pattern Recognition · Computer Science 2019-08-21 Roman Klokov , Jakob Verbeek , Edmond Boyer

We propose PSSNet, a network architecture for generating diverse plausible 3D reconstructions from a single 2.5D depth image. Existing methods tend to produce only small variations on a single shape, even when multiple shapes are consistent…

Robotics · Computer Science 2020-11-19 Brad Saund , Dmitry Berenson

RGB images differentiate from depth images as they carry more details about the color and texture information, which can be utilized as a vital complementary to depth for boosting the performance of 3D semantic scene completion (SSC). SSC…

Computer Vision and Pattern Recognition · Computer Science 2019-05-02 Jie Li , Yu Liu , Dong Gong , Qinfeng Shi , Xia Yuan , Chunxia Zhao , Ian Reid

Endovascular surgical tool reconstruction represents an important factor in advancing endovascular tool navigation, which is an important step in endovascular surgery. However, the lack of publicly available datasets significantly restricts…

Image and Video Processing · Electrical Eng. & Systems 2024-10-30 Tudor Jianu , Baoru Huang , Hoan Nguyen , Binod Bhattarai , Tuong Do , Erman Tjiputra , Quang Tran , Pierre Berthet-Rayne , Ngan Le , Sebastiano Fichera , Anh Nguyen

This paper presents an approach to estimating the continuous 6-DoF pose of an object from a single RGB image. The approach combines semantic keypoints predicted by a convolutional network (convnet) with a deformable shape model. Unlike…

Computer Vision and Pattern Recognition · Computer Science 2022-04-13 Karl Schmeckpeper , Philip R. Osteen , Yufu Wang , Georgios Pavlakos , Kenneth Chaney , Wyatt Jordan , Xiaowei Zhou , Konstantinos G. Derpanis , Kostas Daniilidis

Recovering the 3D shape of an object from single or multiple images with deep neural networks has been attracting increasing attention in the past few years. Mainstream works (e.g. 3D-R2N2) use recurrent neural networks (RNNs) to…

Computer Vision and Pattern Recognition · Computer Science 2020-10-06 Haozhe Xie , Hongxun Yao , Shengping Zhang , Shangchen Zhou , Wenxiu Sun

We present UrbanScene3D, a large-scale data platform for research of urban scene perception and reconstruction. UrbanScene3D contains over 128k high-resolution images covering 16 scenes including large-scale real urban regions and synthetic…

Computer Vision and Pattern Recognition · Computer Science 2022-07-20 Liqiang Lin , Yilin Liu , Yue Hu , Xingguang Yan , Ke Xie , Hui Huang

3D fragment reassembly aims to recover the rigid poses of unordered fragment point clouds or meshes in a common object coordinate system to reconstruct the complete shape. The problem becomes particularly challenging as the number of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Hanze Jia , Chunshi Wang , Yuxiao Yang , Zhonghua Jiang , Yawei Luo , Shuainan Ye , Tan Tang

Generating 3D shapes from single RGB images is essential in various applications such as robotics. Current approaches typically target images containing clear and complete visual descriptions of the object, without considering common…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Yiheng Xiong , Angela Dai

Indoor 3D object detection is an essential task in single image scene understanding, impacting spatial cognition fundamentally in visual reasoning. Existing works on 3D object detection from a single image either pursue this goal through…

Computer Vision and Pattern Recognition · Computer Science 2023-11-21 Yanjun Liu , Wenming Yang

We address the important problem of generalizing robotic rearrangement to clutter without any explicit object models. We first generate over 650K cluttered scenes - orders of magnitude more than prior work - in diverse everyday…

Robotics · Computer Science 2023-04-20 Adithyavairavan Murali , Arsalan Mousavian , Clemens Eppner , Adam Fishman , Dieter Fox

Convolutional Neural Networks (CNN) have been regarded as a powerful class of models for image recognition problems. Nevertheless, it is not trivial when utilizing a CNN for learning spatio-temporal video representation. A few studies have…

Computer Vision and Pattern Recognition · Computer Science 2017-11-29 Zhaofan Qiu , Ting Yao , Tao Mei

Various datasets have been proposed for simultaneous localization and mapping (SLAM) and related problems. Existing datasets often include small environments, have incomplete ground truth, or lack important sensor data, such as depth and…

Computer Vision and Pattern Recognition · Computer Science 2023-01-04 Janne Mustaniemi , Juho Kannala , Esa Rahtu , Li Liu , Janne Heikkilä

Detecting 3D objects from a single RGB image is intrinsically ambiguous, thus requiring appropriate prior knowledge and intermediate representations as constraints to reduce the uncertainties and improve the consistencies between the 2D…

Computer Vision and Pattern Recognition · Computer Science 2019-12-18 Siyuan Huang , Yixin Chen , Tao Yuan , Siyuan Qi , Yixin Zhu , Song-Chun Zhu

We introduce POP3D, a novel framework that creates a full $360^\circ$-view 3D model from a single image. POP3D resolves two prominent issues that limit the single-view reconstruction. Firstly, POP3D offers substantial generalizability to…

Computer Vision and Pattern Recognition · Computer Science 2023-09-20 Nuri Ryu , Minsu Gong , Geonung Kim , Joo-Haeng Lee , Sunghyun Cho

Humans perceive the 3D world as a set of distinct objects that are characterized by various low-level (geometry, reflectance) and high-level (connectivity, adjacency, symmetry) properties. Recent methods based on convolutional neural…

Computer Vision and Pattern Recognition · Computer Science 2020-04-03 Despoina Paschalidou , Luc van Gool , Andreas Geiger
‹ Prev 1 4 5 6 7 8 10 Next ›