English
Related papers

Related papers: Deep Convolutional Neural Network for 6-DOF Image …

200 papers

We present Supermarket-6DoF, a real-world dataset of 1500 grasp attempts across 20 supermarket objects with publicly available 3D models. Unlike most existing grasping datasets that rely on analytical metrics or simulation for grasp…

Robotics · Computer Science 2025-02-25 Jason Toskov , Akansel Cosgun

Recently, deep clustering, which is able to perform feature learning that favors clustering tasks via deep neural networks, has achieved remarkable performance in image clustering applications. However, the existing deep clustering…

Machine Learning · Computer Science 2018-12-12 Yazhou Ren , Ni Wang , Mingxia Li , Zenglin Xu

We address the problem of estimating the relative 6D pose, i.e., position and orientation, of a target spacecraft, from a monocular image, a key capability for future autonomous Rendezvous and Proximity Operations. Due to the difficulty of…

Computer Vision and Pattern Recognition · Computer Science 2025-09-19 Antoine Legrand , Renaud Detry , Christophe De Vleeschouwer

We introduce UprightNet, a learning-based approach for estimating 2DoF camera orientation from a single RGB image of an indoor scene. Unlike recent methods that leverage deep learning to perform black-box regression from image to…

Computer Vision and Pattern Recognition · Computer Science 2019-08-21 Wenqi Xian , Zhengqi Li , Matthew Fisher , Jonathan Eisenmann , Eli Shechtman , Noah Snavely

In this work, we introduce pose interpreter networks for 6-DoF object pose estimation. In contrast to other CNN-based approaches to pose estimation that require expensively annotated object pose data, our pose interpreter network is trained…

Most state-of-the-art localization algorithms rely on robust relative pose estimation and geometry verification to obtain moving object agnostic camera poses in complex indoor environments. However, this approach is prone to mistakes if a…

Computer Vision and Pattern Recognition · Computer Science 2022-09-22 Martina Dubenova , Anna Zderadickova , Ondrej Kafka , Tomas Pajdla , Michal Polic

6D object pose estimation is a prerequisite for many applications. In recent years, monocular pose estimation has attracted much research interest because it does not need depth measurements. In this work, we introduce ConvPoseCNN, a fully…

Computer Vision and Pattern Recognition · Computer Science 2019-12-17 Catherine Capellen , Max Schwarz , Sven Behnke

Public cameras often have limited metadata describing their attributes. A key missing attribute is the precise location of the camera, using which it is possible to precisely pinpoint the location of events seen in the camera. In this…

Computer Vision and Pattern Recognition · Computer Science 2020-03-25 Pradipta Ghosh , Xiaochen Liu , Hang Qiu , Marcos A. M. Vieira , Gaurav S. Sukhatme , Ramesh Govindan

This work proposes a new end-to-end DCNN based approach for motion segmentation, especially for video sequences captured with such non-static cameras, called MOSNET. While other approaches focus on spatial or temporal context only, the…

Computer Vision and Pattern Recognition · Computer Science 2021-02-23 Markus Bosch

We apply convolutional neural networks (CNN) to the problem of image orientation detection in the context of determining the correct orientation (from 0, 90, 180, and 270 degrees) of a consumer photo. The problem is especially important for…

Computer Vision and Pattern Recognition · Computer Science 2023-06-02 Ujash Joshi , Michael Guerzhoy

Visual relocalization has been a widely discussed problem in 3D vision: given a pre-constructed 3D visual map, the 6 DoF (Degrees-of-Freedom) pose of a query image is estimated. Relocalization in large-scale indoor environments enables…

Computer Vision and Pattern Recognition · Computer Science 2022-07-27 Jiahui Zhang , Shitao Tang , Kejie Qiu , Rui Huang , Chuan Fang , Le Cui , Zilong Dong , Siyu Zhu , Ping Tan

Holography encodes the three dimensional (3D) information of a sample in the form of an intensity-only recording. However, to decode the original sample image from its hologram(s), auto-focusing and phase-recovery are needed, which are in…

Computer Vision and Pattern Recognition · Computer Science 2018-12-31 Yichen Wu , Yair Rivenson , Yibo Zhang , Zhensong Wei , Harun Gunaydin , Xing Lin , Aydogan Ozcan

We propose a random convolutional neural network to generate a feature space in which we study image classification and retrieval performance. Put briefly we apply random convolutional blocks followed by global average pooling to generate a…

Computer Vision and Pattern Recognition · Computer Science 2019-03-19 Yunzhe Xue , Usman Roshan

Image classification from independent and identically distributed random variables is considered. Image classifiers are defined which are based on a linear combination of deep convolutional networks with max-pooling layer. Here all the…

Statistics Theory · Mathematics 2025-03-06 Michael Kohler , Adam Krzyzak , Alisha Sänger

Despite the significant progress in 6-DoF visual localization, researchers are mostly driven by ground-level benchmarks. Compared with aerial oblique photography, ground-level map collection lacks scalability and complete coverage. In this…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Shen Yan , Xiaoya Cheng , Yuxiang Liu , Juelin Zhu , Rouwan Wu , Yu Liu , Maojun Zhang

Estimating the location where an image was taken based solely on the contents of the image is a challenging task, even for humans, as properly labeling an image in such a fashion relies heavily on contextual information, and is not as…

Computer Vision and Pattern Recognition · Computer Science 2017-12-29 Jesse M. Johns , Jeremiah Rounds , Michael J. Henry

Many works in collaborative robotics and human-robot interaction focuses on identifying and predicting human behaviour while considering the information about the robot itself as given. This can be the case when sensors and the robot are…

Deep learning has been applied to camera relocalization, in particular, PoseNet and its extended work are the convolutional neural networks which regress the camera pose from a single image. However there are many problems, one of them is…

Robotics · Computer Science 2018-02-27 Qiang Fang , Tianjiang Hu

Novel view synthesis is a challenging problem in computer vision and robotics. Different from the existing works, which need the reference images or 3D models of the scene to generate images under novel views, we propose a novel paradigm to…

Computer Vision and Pattern Recognition · Computer Science 2020-10-23 Xiang Guo , Bo Li , Yuchao Dai , Tongxin Zhang , Hui Deng

Automated classification of human anatomy is an important prerequisite for many computer-aided diagnosis systems. The spatial complexity and variability of anatomy throughout the human body makes classification difficult. "Deep learning"…

Computer Vision and Pattern Recognition · Computer Science 2015-09-17 Holger R. Roth , Christopher T. Lee , Hoo-Chang Shin , Ari Seff , Lauren Kim , Jianhua Yao , Le Lu , Ronald M. Summers
‹ Prev 1 8 9 10 Next ›