English
Related papers

Related papers: GKNet: Graph-based Keypoints Network for Monocular…

200 papers

Prior work on 6-DoF object pose estimation has largely focused on instance-level processing, in which a textured CAD model is available for each object being detected. Category-level 6-DoF pose estimation represents an important step toward…

Computer Vision and Pattern Recognition · Computer Science 2022-05-13 Yunzhi Lin , Jonathan Tremblay , Stephen Tyree , Patricio A. Vela , Stan Birchfield

Global point cloud registration is essential in many robotics tasks like loop closing and relocalization. Unfortunately, the registration often suffers from the low overlap between point clouds, a frequent occurrence in practical…

Robotics · Computer Science 2023-07-25 Zhijian Qiao , Zehuan Yu , Huan Yin , Shaojie Shen

Fine-grained RGBT image semantic segmentation is crucial for all-weather unmanned aerial vehicle (UAV) scene understanding. However, UAV RGBT image semantic segmentation faces two coupled challenges: cross-modal spatial misalignment caused…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Fangqiang Fan , Zhicheng Zhao , Xiaoliang Ma , Chenglong Li , Jin Tang

Self-supervised monocular depth estimation has emerged as a promising method because it does not require groundtruth depth maps during training. As an alternative for the groundtruth depth map, the photometric loss enables to provide…

Computer Vision and Pattern Recognition · Computer Science 2021-01-01 Jaehoon Choi , Dongki Jung , Donghwan Lee , Changick Kim

In recent years, a plethora of diverse methods have been proposed for 3D pose estimation. Among these, self-attention mechanisms and graph convolutions have both been proven to be effective and practical methods. Recognizing the strengths…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Sihan Wen , Xiantan Zhu , Zhiming Tan

Bottom-up approaches for image-based multi-person pose estimation consist of two stages: (1) keypoint detection and (2) grouping of the detected keypoints to form person instances. Current grouping approaches rely on learned embedding from…

Computer Vision and Pattern Recognition · Computer Science 2021-04-07 Jiahao Lin , Gim Hee Lee

We propose a Convolutional Neural Network (CNN)-based model "RotationNet," which takes multi-view images of an object as input and jointly estimates its pose and object category. Unlike previous approaches that use known viewpoint labels…

Computer Vision and Pattern Recognition · Computer Science 2018-03-26 Asako Kanezaki , Yasuyuki Matsushita , Yoshifumi Nishida

Despite substantial progress in 3D human pose estimation from a single-view image, prior works rarely explore global and local correlations, leading to insufficient learning of human skeleton representations. To address this issue, we…

Computer Vision and Pattern Recognition · Computer Science 2023-04-28 Ti Wang , Hong Liu , Runwei Ding , Wenhao Li , Yingxuan You , Xia Li

In this paper, we present a novel end-to-end group collaborative learning network, termed GCoNet+, which can effectively and efficiently (250 fps) identify co-salient objects in natural scenes. The proposed GCoNet+ achieves the new…

Computer Vision and Pattern Recognition · Computer Science 2023-04-11 Peng Zheng , Huazhu Fu , Deng-Ping Fan , Qi Fan , Jie Qin , Yu-Wing Tai , Chi-Keung Tang , Luc Van Gool

Accurately matching local features between a pair of images is a challenging computer vision task. Previous studies typically use attention based graph neural networks (GNNs) with fully-connected graphs over keypoints within/across images…

Computer Vision and Pattern Recognition · Computer Science 2023-07-06 Zizhuo Li , Jiayi Ma

We propose a single-stage, category-level 6-DoF pose estimation algorithm that simultaneously detects and tracks instances of objects within a known category. Our method takes as input the previous and current frame from a monocular RGB…

Computer Vision and Pattern Recognition · Computer Science 2022-05-24 Yunzhi Lin , Jonathan Tremblay , Stephen Tyree , Patricio A. Vela , Stan Birchfield

Occlusion and the scarcity of labeled surgical data are significant challenges in disparity estimation for stereo laparoscopic images. To address these issues, this study proposes a Depth Guided Occlusion-Aware Disparity Refinement Network…

Computer Vision and Pattern Recognition · Computer Science 2025-05-14 Ziteng Liu , Dongdong He , Chenghong Zhang , Wenpeng Gao , Yili Fu

Human motion is a continuous physical process in 3D space, governed by complex dynamic and kinematic constraints. Existing methods typically represent the human pose as an abstract graph structure, neglecting the intrinsic physical…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Shuaijin Wan

The rapid development of urban low-altitude unmanned aerial vehicle (UAV) economy poses new challenges for dynamic site selection of UAV landing points and supply stations. Traditional deep reinforcement learning methods face computational…

Machine Learning · Computer Science 2025-07-16 Jianing Zhi , Xinghua Li , Zidong Chen

General object grasping is an important yet unsolved problem in the field of robotics. Most of the current methods either generate grasp poses with few DoF that fail to cover most of the success grasps, or only take the unstable depth image…

Robotics · Computer Science 2021-03-04 Minghao Gou , Hao-Shu Fang , Zhanda Zhu , Sheng Xu , Chenxi Wang , Cewu Lu

Category-Agnostic Pose Estimation (CAPE) localizes keypoints across diverse object categories with a single model, using one or a few annotated support images. Recent works have shown that using a pose graph (i.e., treating keypoints as…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Or Hirschorn , Shai Avidan

This paper tackles category-level pose estimation of articulated objects in robotic manipulation tasks and introduces a new benchmark dataset. While recent methods estimate part poses and sizes at the category level, they often rely on…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Jingshun Huang , Haitao Lin , Tianyu Wang , Yanwei Fu , Xiangyang Xue , Yi Zhu

In human pose estimation methods based on graph convolutional architectures, the human skeleton is usually modeled as an undirected graph whose nodes are body joints and edges are connections between neighboring joints. However, most of…

Computer Vision and Pattern Recognition · Computer Science 2023-08-09 Tanvir Hassan , A. Ben Hamza

Estimating 3D poses from a monocular video is still a challenging task, despite the significant progress that has been made in recent years. Generally, the performance of existing methods drops when the target person is too small/large, or…

Computer Vision and Pattern Recognition · Computer Science 2020-04-27 Yu Cheng , Bo Yang , Bo Wang , Robby T. Tan

We present GraphDepth, a monocular depth estimation architecture that synergistically integrates Graph Neural Networks (GNNs) within a convolutional encoder-decoder framework. Our approach embeds efficient GraphSAGE layers at multiple…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Ishan Narayan
‹ Prev 1 4 5 6 7 8 10 Next ›