English
Related papers

Related papers: Direction-Aware Hybrid Representation Learning for…

200 papers

Since the introduction of modern deep learning methods for object pose estimation, test accuracy and efficiency has increased significantly. For training, however, large amounts of annotated training data are required for good performance.…

Computer Vision and Pattern Recognition · Computer Science 2021-08-18 Frederik Hagelskjaer , Anders Glent Buch

Deep Implicit Functions (DIFs) have gained popularity in 3D computer vision due to their compactness and continuous representation capabilities. However, addressing dense correspondences and semantic relationships across DIF-encoded shapes…

Computer Vision and Pattern Recognition · Computer Science 2023-07-06 Kun Han , Shanlin Sun , Xiaohui Xie

The recent success of deep networks has significantly advanced 3D human pose estimation from 2D images. The diversity of capturing viewpoints and the flexibility of the human poses, however, remain some significant challenges. In this…

Computer Vision and Pattern Recognition · Computer Science 2019-01-31 Guoqiang Wei , Cuiling Lan , Wenjun Zeng , Zhibo Chen

We propose a new bottom-up method for multi-person 2D human pose estimation that is particularly well suited for urban mobility such as self-driving cars and delivery robots. The new method, PifPaf, uses a Part Intensity Field (PIF) to…

Computer Vision and Pattern Recognition · Computer Science 2019-04-08 Sven Kreiss , Lorenzo Bertoni , Alexandre Alahi

This paper proposes an approach to learn generic multi-modal mesh surface representations using a novel scheme for fusing texture and geometric data. Our approach defines an inverse mapping between different geometric descriptors computed…

Computer Vision and Pattern Recognition · Computer Science 2019-04-10 Bilal Taha , Munawar Hayat , Stefano Berretti , Naoufel Werghi

Camera and LiDAR sensor modalities provide complementary appearance and geometric information useful for detecting 3D objects for autonomous vehicle applications. However, current end-to-end fusion methods are challenging to train and…

Computer Vision and Pattern Recognition · Computer Science 2022-10-28 Anas Mahmoud , Jordan S. K. Hu , Steven L. Waslander

In this paper, we tackle the challenging problem of 3D keypoint estimation of general objects using a novel implicit representation. Previous works have demonstrated promising results for keypoint prediction through direct coordinate…

Computer Vision and Pattern Recognition · Computer Science 2023-06-21 Xiangyu Zhu , Dong Du , Haibin Huang , Chongyang Ma , Xiaoguang Han

Acquiring spatio-temporal states of an action is the most crucial step for action classification. In this paper, we propose a data level fusion strategy, Motion Fused Frames (MFFs), designed to fuse motion information into static images as…

Computer Vision and Pattern Recognition · Computer Science 2018-04-27 Okan Köpüklü , Neslihan Köse , Gerhard Rigoll

We consider the problem of estimating human pose and trajectory by an aerial robot with a monocular camera in near real time. We present a preliminary solution whose distinguishing feature is a dynamic classifier selection architecture. In…

Computer Vision and Pattern Recognition · Computer Science 2018-12-18 Asanka G Perera , Yee Wei Law , Javaan Chahl

Existing monocular 3D pose estimation methods primarily rely on joint positional features, while overlooking intrinsic directional and angular correlations within the skeleton. As a result, they often produce implausible poses under joint…

Computer Vision and Pattern Recognition · Computer Science 2025-06-18 Ming Xu , Xu Zhang

Grasping objects with limited or no prior knowledge about them is a highly relevant skill in assistive robotics. Still, in this general setting, it has remained an open problem, especially when it comes to only partial observability and…

Robotics · Computer Science 2026-01-21 Matthias Humt , Dominik Winkelbauer , Ulrich Hillenbrand , Berthold Bäuml

How to effectively represent camera pose is an essential problem in 3D computer vision, especially in tasks such as camera pose regression and novel view synthesis. Traditionally, 3D position of the camera is represented by Cartesian…

Computer Vision and Pattern Recognition · Computer Science 2021-05-11 Yaxuan Zhu , Ruiqi Gao , Siyuan Huang , Song-Chun Zhu , Ying Nian Wu

We develop a system for modeling hand-object interactions in 3D from RGB images that show a hand which is holding a novel object from a known category. We design a Convolutional Neural Network (CNN) for Hand-held Object Pose and Shape…

Computer Vision and Pattern Recognition · Computer Science 2019-11-12 Mia Kokic , Danica Kragic , Jeannette Bohg

This work proposes an end-to-end approach to estimate full 3D hand pose from stereo cameras. Most existing methods of estimating hand pose from stereo cameras apply stereo matching to obtain depth map and use depth-based solution to…

Computer Vision and Pattern Recognition · Computer Science 2022-06-06 Yuncheng Li , Zehao Xue , Yingying Wang , Liuhao Ge , Zhou Ren , Jonathan Rodriguez

We propose an entirely data-driven approach to estimating the 3D pose of a hand given a depth image. We show that we can correct the mistakes made by a Convolutional Neural Network trained to predict an estimate of the 3D pose by using a…

Computer Vision and Pattern Recognition · Computer Science 2016-10-03 Markus Oberweger , Paul Wohlhart , Vincent Lepetit

State-of-the-art LiDAR-camera 3D object detectors usually focus on feature fusion. However, they neglect the factor of depth while designing the fusion strategy. In this work, we are the first to observe that different modalities play…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Mingqian Ji , Jian Yang , Shanshan Zhang

Many 3D tasks such as pose alignment, animation, motion transfer, and 3D reconstruction rely on establishing correspondences between 3D shapes. This challenge has recently been approached by pairwise matching of semantic features from…

Computer Vision and Pattern Recognition · Computer Science 2025-10-14 Lukas Uzolas , Elmar Eisemann , Petr Kellnhofer

Recent work achieved impressive progress towards joint reconstruction of hands and manipulated objects from monocular color images. Existing methods focus on two alternative representations in terms of either parametric meshes or signed…

Computer Vision and Pattern Recognition · Computer Science 2022-07-27 Zerui Chen , Yana Hasson , Cordelia Schmid , Ivan Laptev

Fusing 3D LiDAR features with 2D camera features is a promising technique for enhancing the accuracy of 3D detection, thanks to their complementary physical properties. While most of the existing methods focus on directly fusing camera…

Computer Vision and Pattern Recognition · Computer Science 2024-01-17 Lemeng Wu , Dilin Wang , Meng Li , Yunyang Xiong , Raghuraman Krishnamoorthi , Qiang Liu , Vikas Chandra

The ability of humans to infer head poses from face shapes, and vice versa, indicates a strong correlation between the two. Accordingly, recent studies on face alignment have employed head pose information to predict facial landmarks in…

Computer Vision and Pattern Recognition · Computer Science 2023-08-28 Jaehyun So , Youngjoon Han