中文
相关论文

相关论文: Parcel3D: Shape Reconstruction from Single RGB Ima…

200 篇论文

In the realm of robotic grasping, achieving accurate and reliable interactions with the environment is a pivotal challenge. Traditional methods of grasp planning methods utilizing partial point clouds derived from depth image often suffer…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Lei Zhou , Haozhe Wang , Zhengshen Zhang , Zhiyang Liu , Francis EH Tay , adn Marcelo H. Ang.

We present Spann3R, a novel approach for dense 3D reconstruction from ordered or unordered image collections. Built on the DUSt3R paradigm, Spann3R uses a transformer-based architecture to directly regress pointmaps from images without any…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Hengyi Wang , Lourdes Agapito

Instance-level object re-identification is a fundamental computer vision task, with applications from image retrieval to intelligent monitoring and fraud detection. In this work, we propose the novel task of damaged object…

计算机视觉与模式识别 · 计算机科学 2023-04-18 Luca Piano , Filippo Gabriele Pratticò , Alessandro Sebastian Russo , Lorenzo Lanari , Lia Morra , Fabrizio Lamberti

Reconstructing three-dimensional (3D) scenes with semantic understanding is vital in many robotic applications. Robots need to identify which objects, along with their positions and shapes, to manipulate them precisely with given tasks.…

机器人学 · 计算机科学 2024-12-17 Khang Nguyen , Tuan Dang , Manfred Huber

We study end-to-end learning strategies for 3D shape inference from images, in particular from a single image. Several approaches in this direction have been investigated that explore different shape representations and suitable learning…

计算机视觉与模式识别 · 计算机科学 2019-08-21 Roman Klokov , Jakob Verbeek , Edmond Boyer

We propose PSSNet, a network architecture for generating diverse plausible 3D reconstructions from a single 2.5D depth image. Existing methods tend to produce only small variations on a single shape, even when multiple shapes are consistent…

机器人学 · 计算机科学 2020-11-19 Brad Saund , Dmitry Berenson

RGB images differentiate from depth images as they carry more details about the color and texture information, which can be utilized as a vital complementary to depth for boosting the performance of 3D semantic scene completion (SSC). SSC…

计算机视觉与模式识别 · 计算机科学 2019-05-02 Jie Li , Yu Liu , Dong Gong , Qinfeng Shi , Xia Yuan , Chunxia Zhao , Ian Reid

Endovascular surgical tool reconstruction represents an important factor in advancing endovascular tool navigation, which is an important step in endovascular surgery. However, the lack of publicly available datasets significantly restricts…

图像与视频处理 · 电气工程与系统科学 2024-10-30 Tudor Jianu , Baoru Huang , Hoan Nguyen , Binod Bhattarai , Tuong Do , Erman Tjiputra , Quang Tran , Pierre Berthet-Rayne , Ngan Le , Sebastiano Fichera , Anh Nguyen

This paper presents an approach to estimating the continuous 6-DoF pose of an object from a single RGB image. The approach combines semantic keypoints predicted by a convolutional network (convnet) with a deformable shape model. Unlike…

Recovering the 3D shape of an object from single or multiple images with deep neural networks has been attracting increasing attention in the past few years. Mainstream works (e.g. 3D-R2N2) use recurrent neural networks (RNNs) to…

计算机视觉与模式识别 · 计算机科学 2020-10-06 Haozhe Xie , Hongxun Yao , Shengping Zhang , Shangchen Zhou , Wenxiu Sun

We present UrbanScene3D, a large-scale data platform for research of urban scene perception and reconstruction. UrbanScene3D contains over 128k high-resolution images covering 16 scenes including large-scale real urban regions and synthetic…

计算机视觉与模式识别 · 计算机科学 2022-07-20 Liqiang Lin , Yilin Liu , Yue Hu , Xingguang Yan , Ke Xie , Hui Huang

3D fragment reassembly aims to recover the rigid poses of unordered fragment point clouds or meshes in a common object coordinate system to reconstruct the complete shape. The problem becomes particularly challenging as the number of…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Hanze Jia , Chunshi Wang , Yuxiao Yang , Zhonghua Jiang , Yawei Luo , Shuainan Ye , Tan Tang

Generating 3D shapes from single RGB images is essential in various applications such as robotics. Current approaches typically target images containing clear and complete visual descriptions of the object, without considering common…

计算机视觉与模式识别 · 计算机科学 2024-11-05 Yiheng Xiong , Angela Dai

Indoor 3D object detection is an essential task in single image scene understanding, impacting spatial cognition fundamentally in visual reasoning. Existing works on 3D object detection from a single image either pursue this goal through…

计算机视觉与模式识别 · 计算机科学 2023-11-21 Yanjun Liu , Wenming Yang

We address the important problem of generalizing robotic rearrangement to clutter without any explicit object models. We first generate over 650K cluttered scenes - orders of magnitude more than prior work - in diverse everyday…

机器人学 · 计算机科学 2023-04-20 Adithyavairavan Murali , Arsalan Mousavian , Clemens Eppner , Adam Fishman , Dieter Fox

Convolutional Neural Networks (CNN) have been regarded as a powerful class of models for image recognition problems. Nevertheless, it is not trivial when utilizing a CNN for learning spatio-temporal video representation. A few studies have…

计算机视觉与模式识别 · 计算机科学 2017-11-29 Zhaofan Qiu , Ting Yao , Tao Mei

Various datasets have been proposed for simultaneous localization and mapping (SLAM) and related problems. Existing datasets often include small environments, have incomplete ground truth, or lack important sensor data, such as depth and…

计算机视觉与模式识别 · 计算机科学 2023-01-04 Janne Mustaniemi , Juho Kannala , Esa Rahtu , Li Liu , Janne Heikkilä

Detecting 3D objects from a single RGB image is intrinsically ambiguous, thus requiring appropriate prior knowledge and intermediate representations as constraints to reduce the uncertainties and improve the consistencies between the 2D…

计算机视觉与模式识别 · 计算机科学 2019-12-18 Siyuan Huang , Yixin Chen , Tao Yuan , Siyuan Qi , Yixin Zhu , Song-Chun Zhu

We introduce POP3D, a novel framework that creates a full $360^\circ$-view 3D model from a single image. POP3D resolves two prominent issues that limit the single-view reconstruction. Firstly, POP3D offers substantial generalizability to…

计算机视觉与模式识别 · 计算机科学 2023-09-20 Nuri Ryu , Minsu Gong , Geonung Kim , Joo-Haeng Lee , Sunghyun Cho

Humans perceive the 3D world as a set of distinct objects that are characterized by various low-level (geometry, reflectance) and high-level (connectivity, adjacency, symmetry) properties. Recent methods based on convolutional neural…

计算机视觉与模式识别 · 计算机科学 2020-04-03 Despoina Paschalidou , Luc van Gool , Andreas Geiger