中文
相关论文

相关论文: ROCA: Robust CAD Model Retrieval and Alignment fro…

200 篇论文

We present a novel, end-to-end approach to align CAD models to an 3D scan of a scene, enabling transformation of a noisy, incomplete 3D scan to a compact, CAD reconstruction with clean, complete object geometry. Our main contribution lies…

计算机视觉与模式识别 · 计算机科学 2019-06-12 Armen Avetisyan , Angela Dai , Matthias Nießner

Reconstructing 3D shape and pose of static objects from a single image is an essential task for various industries, including robotics, augmented reality, and digital content creation. This can be done by directly predicting 3D shape in…

计算机视觉与模式识别 · 计算机科学 2023-10-18 Florian Langer , Ignas Budvytis , Roberto Cipolla

We present an automated and efficient approach for retrieving high-quality CAD models of objects and their poses in a scene captured by a moving RGB-D camera. We first investigate various objective functions to measure similarity between a…

计算机视觉与模式识别 · 计算机科学 2023-09-13 Stefan Ainetter , Sinisa Stekovic , Friedrich Fraundorfer , Vincent Lepetit

One practical approach to infer 3D scene structure from a single image is to retrieve a closely matching 3D model from a database and align it with the object in the image. Existing methods rely on supervised training with images and pose…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Pattaramanee Arsomngern , Sasikarn Khwanmuang , Matthias Nießner , Supasorn Suwajanakorn

Estimating 3D shapes and poses of static objects from a single image has important applications for robotics, augmented reality and digital content creation. Often this is done through direct mesh predictions which produces unrealistic,…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Florian Langer , Gwangbin Bae , Ignas Budvytis , Roberto Cipolla

Visual retrieval systems face significant challenges when updating models with improved representations due to misalignment between the old and new representations. The costly and resource-intensive backfilling process involves…

计算机视觉与模式识别 · 计算机科学 2024-08-19 Simone Ricci , Niccolò Biondi , Federico Pernici , Alberto Del Bimbo

This paper introduces RaCo, a lightweight neural network designed to learn robust and versatile keypoints suitable for a variety of 3D computer vision tasks. The model integrates three key components: the repeatable keypoint detector, a…

计算机视觉与模式识别 · 计算机科学 2026-02-18 Abhiram Shenoi , Philipp Lindenberger , Paul-Edouard Sarlin , Marc Pollefeys

Without using extra 3-D data like points cloud or depth images for providing 3-D information, we retrieve the 3-D object information from single monocular images. The high-quality predicted depth images are recovered from single monocular…

计算机视觉与模式识别 · 计算机科学 2020-02-14 Zifan Yu , Suya You

Instance shape reconstruction from a 3D scene involves recovering the full geometries of multiple objects at the semantic instance level. Many methods leverage data-driven learning due to the intricacies of scene complexity and significant…

计算机视觉与模式识别 · 计算机科学 2023-12-20 Haolin Liu , Chongjie Ye , Yinyu Nie , Yingfan He , Xiaoguang Han

We propose the Compact Clustering Attention (COCA) layer, an effective building block that introduces a hierarchical strategy for object-centric representation learning, while solving the unsupervised object discovery task on single images.…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Can Küçüksözen , Yücel Yemez

CAD model retrieval to real-world scene observations has shown strong promise as a basis for 3D perception of objects and a clean, lightweight mesh-based scene representation; however, current approaches to retrieve CAD models to a query…

计算机视觉与模式识别 · 计算机科学 2022-03-25 Tim Beyer , Angela Dai

We introduce 3D-COCO, an extension of the original MS-COCO dataset providing 3D models and 2D-3D alignment annotations. 3D-COCO was designed to achieve computer vision tasks such as 3D reconstruction or image detection configurable with…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Maxence Bideaux , Alice Phe , Mohamed Chaouch , Bertrand Luvison , Quoc-Cuong Pham

Collaborative autonomous driving with multiple vehicles usually requires the data fusion from multiple modalities. To ensure effective fusion, the data from each individual modality shall maintain a reasonably high quality. However, in…

人工智能 · 计算机科学 2024-08-02 Zhe Huang , Shuo Wang , Yongcai Wang , Wanting Li , Deying Li , Lei Wang

Digitising the 3D world into a clean, CAD model-based representation has important applications for augmented reality and robotics. Current state-of-the-art methods are computationally intensive as they individually encode each detected…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Florian Langer , Jihong Ju , Georgi Dikov , Gerhard Reitmayr , Mohsen Ghafoorian

We present Scan2CAD, a novel data-driven method that learns to align clean 3D CAD models from a shape database to the noisy and incomplete geometry of a commodity RGB-D scan. For a 3D reconstruction of an indoor scene, our method takes as…

计算机视觉与模式识别 · 计算机科学 2018-11-29 Armen Avetisyan , Manuel Dahnert , Angela Dai , Manolis Savva , Angel X. Chang , Matthias Nießner

We propose a method to detect and reconstruct multiple 3D objects from a single RGB image. The key idea is to optimize for detection, alignment and shape jointly over all objects in the RGB image, while focusing on realistic and physically…

计算机视觉与模式识别 · 计算机科学 2021-06-23 Francis Engelmann , Konstantinos Rematas , Bastian Leibe , Vittorio Ferrari

Flow-matching methods for 3D shape assembly learn point-wise velocity fields that transport parts toward assembled configurations, yet they receive no explicit guidance about which cross-part interactions should drive the motion. We…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Nahyuk Lee , Zhiang Chen , Marc Pollefeys , Sunghwan Hong

The three-dimensional representation of objects or scenes starting from a set of images has been a widely discussed topic for years and has gained additional attention after the diffusion of NeRF-based approaches. However, an underestimated…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Davide Di Nucci , Alessandro Simoni , Matteo Tomei , Luca Ciuffreda , Roberto Vezzani , Rita Cucchiara

Scene rearrangement, like table tidying, is a challenging task in robotic manipulation due to the complexity of predicting diverse object arrangements. Web-scale trained generative models such as Stable Diffusion can aid by generating…

机器人学 · 计算机科学 2024-12-03 Shutong Jin , Ruiyu Wang , Kuangyi Chen , Florian T. Pokorny

We present DenseRaC, a novel end-to-end framework for jointly estimating 3D human pose and body shape from a monocular RGB image. Our two-step framework takes the body pixel-to-surface correspondence map (i.e., IUV map) as proxy…

计算机视觉与模式识别 · 计算机科学 2019-10-10 Yuanlu Xu , Song-Chun Zhu , Tony Tung
‹ 上一页 1 2 3 10 下一页 ›