中文
相关论文

相关论文: SD-6DoF-ICLK: Sparse and Deep Inverse Compositiona…

200 篇论文

Efficient and high-fidelity polarization demosaicking is critical for industrial applications of the division of focal plane (DoFP) polarization imaging systems. However, existing methods have an unsatisfactory balance of speed, accuracy,…

图像与视频处理 · 电气工程与系统科学 2024-09-02 Guangsen Liu , Peng Rao , Xin Chen , Yao Li , Haixin Jiang

Category-level object pose estimation, aiming to predict the 6D pose and 3D size of objects from known categories, typically struggles with large intra-class shape variation. Existing works utilizing mean shapes often fall short of…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Yamei Chen , Yan Di , Guangyao Zhai , Fabian Manhardt , Chenyangguang Zhang , Ruida Zhang , Federico Tombari , Nassir Navab , Benjamin Busam

We present a novel inverse rendering-based framework to estimate the 3D shape (per-pixel surface normals and depth) of objects and scenes from single-view polarization images, the problem popularly known as Shape from Polarization (SfP).…

计算机视觉与模式识别 · 计算机科学 2024-07-15 Ashish Tiwari , Shanmuganathan Raman

This paper proposes PoseLecTr, a graph-based encoder-decoder framework that integrates a novel Legendre convolution with attention mechanisms for six-degree-of-freedom (6-DOF) object pose estimation from monocular RGB images. Conventional…

计算机视觉与模式识别 · 计算机科学 2026-01-08 Alexander Du , Xiujin Liu

In this paper, we propose a dense monocular SLAM system, named DeepRelativeFusion, that is capable to recover a globally consistent 3D structure. To this end, we use a visual SLAM algorithm to reliably recover the camera poses and…

计算机视觉与模式识别 · 计算机科学 2021-07-13 Shing Yan Loo , Syamsiah Mashohor , Sai Hong Tang , Hong Zhang

For robust visual-inertial SLAM in perceptually-challenging indoor environments,recent studies exploit line features to extract descriptive information about scene structure to deal with the degeneracy of point features. But existing…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Wanting Li , Shuo Wang , Yongcai Wang , Yu Shao , Xuewei Bai , Deying Li

LiDAR-based SLAM is a core technology for autonomous vehicles and robots. One key contribution of this work to 3D LiDAR SLAM and localization is a fierce defense of view-based maps (pose graphs with time-stamped sensor readings) as the…

机器人学 · 计算机科学 2025-08-19 José Luis Blanco-Claraco

This paper proposes a generalizable, end-to-end deep learning-based method for relative pose regression between two images. Given two images of the same scene captured from different viewpoints, our method predicts the relative rotation and…

计算机视觉与模式识别 · 计算机科学 2024-04-17 Fadi Khatib , Yuval Margalit , Meirav Galun , Ronen Basri

For high-level geo-spatial applications and intelligent robotics, accurate global pose information is of crucial importance. Map-aided localization is a universal approach to overcome the limitations of global navigation satellite system…

机器人学 · 计算机科学 2025-11-19 Yuxuan Zhou , Xingxing Li , Shengyu Li , Chunxi Xia , Xuanbin Wang , Shaoquan Feng

This article presents HOTFLoc++, an end-to-end hierarchical framework for LiDAR place recognition, re-ranking, and 6-DoF metric localisation in forests. Leveraging an octree-based transformer, our approach extracts features at multiple…

计算机视觉与模式识别 · 计算机科学 2026-04-10 Ethan Griffiths , Maryam Haghighat , Simon Denman , Clinton Fookes , Milad Ramezani

One practical approach to infer 3D scene structure from a single image is to retrieve a closely matching 3D model from a database and align it with the object in the image. Existing methods rely on supervised training with images and pose…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Pattaramanee Arsomngern , Sasikarn Khwanmuang , Matthias Nießner , Supasorn Suwajanakorn

We propose a method for 6DoF pose estimation of rigid objects that uses a state-of-the-art deep learning based instance detector to segment object instances in an RGB image, followed by a point-pair based voting method to recover the…

计算机视觉与模式识别 · 计算机科学 2020-11-12 Rebecca König , Bertram Drost

LiDAR relocalization has attracted increasing attention as it can deliver accurate 6-DoF pose estimation in complex 3D environments. Recent learning-based regression methods offer efficient solutions by directly predicting global poses…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Jianshi Wu , Minghang Zhu , Dunqiang Liu , Wen Li , Sheng Ao , Siqi Shen , Chenglu Wen , Cheng Wang

Accurate spatial understanding is essential for image-guided surgery, augmented reality integration and context awareness. In minimally invasive procedures, where visual input is the sole intraoperative modality, establishing precise…

计算机视觉与模式识别 · 计算机科学 2025-12-12 Alberto Rota , Elena De Momi

We solve the problem of 6-DoF localisation and 3D dense reconstruction in spatial environments as approximate Bayesian inference in a deep state-space model. Our approach leverages both learning and domain knowledge from multiple-view…

机器学习 · 统计学 2021-03-16 Atanas Mirchev , Baris Kayalibay , Patrick van der Smagt , Justin Bayer

Object 6DoF (6D) pose estimation is essential for robotic perception, especially in industrial settings. It enables robots to interact with the environment and manipulate objects. However, existing benchmarks on object 6D pose estimation…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Ruimin Ma , Sebastian Zudaire , Zhen Li , Chi Zhang

We propose an algorithm for rotational sparse coding along with an efficient implementation using steerability. Sparse coding (also called dictionary learning) is an important technique in image processing, useful in inverse problems,…

图像与视频处理 · 电气工程与系统科学 2020-01-31 Michael T. McCann , Vincent Andrearczyk , Michael Unser , Adrien Depeursinge

Our world is full of identical objects (\emphe.g., cans of coke, cars of same model). These duplicates, when seen together, provide additional and strong cues for us to effectively reason about 3D. Inspired by this observation, we introduce…

计算机视觉与模式识别 · 计算机科学 2024-01-11 Tianhang Cheng , Wei-Chiu Ma , Kaiyu Guan , Antonio Torralba , Shenlong Wang

We introduce S2C-3D, a novel sparse-view 3D reconstruction framework for high-fidelity and complete scene reconstruction from as few as six to eight images. Our framework features three components: a specialized diffusion model for…

计算机视觉与模式识别 · 计算机科学 2026-05-08 Yiyang Shen , Yin Yang , Kun Zhou , Tianjia Shao

Detecting objects and their 6D poses from only RGB images is an important task for many robotic applications. While deep learning methods have made significant progress in visual object detection and segmentation, the object pose estimation…

计算机视觉与模式识别 · 计算机科学 2018-03-01 Thanh-Toan Do , Ming Cai , Trung Pham , Ian Reid