English
Related papers

Related papers: Hand Keypoint Detection in Single Images using Mul…

200 papers

This paper proposes a method to extract the position and pose of vehicles in the 3D world from a single traffic camera. Most previous monocular 3D vehicle detection algorithms focused on cameras on vehicles from the perspective of a driver,…

Computer Vision and Pattern Recognition · Computer Science 2022-01-06 Minghan Zhu , Songan Zhang , Yuanxin Zhong , Pingping Lu , Huei Peng , John Lenneman

Capturing challenging human motions is critical for numerous applications, but it suffers from complex motion patterns and severe self-occlusion under the monocular setting. In this paper, we propose ChallenCap -- a template-based approach…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Yannan He , Anqi Pang , Xin Chen , Han Liang , Minye Wu , Yuexin Ma , Lan Xu

Accurate 7DoF prediction of vehicles at an intersection is an important task for assessing potential conflicts between road users. In principle, this could be achieved by a single camera system that is capable of detecting the pose of each…

Computer Vision and Pattern Recognition · Computer Science 2022-11-30 Matthew Howe , Ian Reid , Jamie Mackenzie

Autonomous driving perception tasks rely heavily on cameras as the primary sensor for Object Detection, Semantic Segmentation, Instance Segmentation, and Object Tracking. However, RGB images captured by cameras lack depth information, which…

Computer Vision and Pattern Recognition · Computer Science 2023-08-02 Marcelo Eduardo Pederiva , José Mario De Martino , Alessandro Zimmer

Understanding the geometry and pose of objects in 2D images is a fundamental necessity for a wide range of real world applications. Driven by deep neural networks, recent methods have brought significant improvements to object pose…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Jogendra Nath Kundu , Rahul M. V. , Aditya Ganeshan , R. Venkatesh Babu

In this work, we propose a novel single-shot and keypoints-based framework for monocular 3D objects detection using only RGB images, called KM3D-Net. We design a fully convolutional model to predict object keypoints, dimension, and…

Computer Vision and Pattern Recognition · Computer Science 2020-09-03 Peixuan Li

This paper presents a novel 3D human pose estimation approach using a single stream of asynchronous events as input. Most of the state-of-the-art approaches solve this task with RGB cameras, however struggling when subjects are moving fast.…

Computer Vision and Pattern Recognition · Computer Science 2021-04-22 Gianluca Scarpellini , Pietro Morerio , Alessio Del Bue

Occlusion is one of the challenging issues when estimating 3D hand pose. This problem becomes more prominent when hand interacts with an object or two hands are involved. In the past works, much attention has not been given to these…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Mallika Garg , Debashis Ghosh , Pyari Mohan Pradhan

Jointly considering multiple camera views (multi-view) is very effective for pedestrian detection under occlusion. For such multi-view systems, it is critical to have well-designed camera configurations, including camera locations,…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Yunzhong Hou , Xingjian Leng , Tom Gedeon , Liang Zheng

Keypoint detection is the foundation of many computer vision tasks, including image registration, structure-from-motion, 3D reconstruction, visual odometry, and SLAM. Traditional detectors (SIFT, ORB, BRISK, FAST, etc.) and learning-based…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Shaharyar Ahmed Khan Tareen , Filza Khan Tareen , Xiaojing Yuan

We present the development and evaluation of a hand tracking algorithm based on single depth images captured from an overhead perspective for use in the COACH prompting system. We train a random decision forest body part classifier using…

Computer Vision and Pattern Recognition · Computer Science 2015-03-10 Stephen Czarnuch , Alex Mihailidis

The typical bottom-up human pose estimation framework includes two stages, keypoint detection and grouping. Most existing works focus on developing grouping algorithms, e.g., associative embedding, and pixel-wise keypoint regression that we…

Computer Vision and Pattern Recognition · Computer Science 2020-06-30 Ke Sun , Zigang Geng , Depu Meng , Bin Xiao , Dong Liu , Zhaoxiang Zhang , Jingdong Wang

Markerless human motion capture (mocap) from multiple RGB cameras is a widely studied problem. Existing methods either need calibrated cameras or calibrate them relative to a static camera, which acts as the reference frame for the mocap…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Nitin Saini , Chun-hao P. Huang , Michael J. Black , Aamir Ahmad

We present a learning framework that learns to recover the 3D shape, pose and texture from a single image, trained on an image collection without any ground truth 3D shape, multi-view, camera viewpoints or keypoint supervision. We approach…

Computer Vision and Pattern Recognition · Computer Science 2020-07-22 Shubham Goel , Angjoo Kanazawa , Jitendra Malik

Multi-sensor frameworks provide opportunities for ensemble learning and sensor fusion to make use of redundancy and supplemental information, helpful in real-world safety applications such as continuous driver state monitoring which…

Machine Learning · Computer Science 2023-10-02 Ross Greer , Mohan Trivedi

To support minimally-invasive intraoperative mitral valve repair, quantitative measurements from the valve can be obtained using an infra-red tracked stylus. It is desirable to view such manually measured points together with the endoscopic…

Computer Vision and Pattern Recognition · Computer Science 2022-01-27 Lukas Burger , Lalith Sharan , Samantha Fischer , Julian Brand , Maximillian Hehl , Gabriele Romano , Matthias Karck , Raffaele De Simone , Ivo Wolf , Sandy Engelhardt

As robotics progresses toward general manipulation, dexterous hands are becoming increasingly critical. However, proprioception in dexterous hands remains a bottleneck due to limitations in volume and generality. In this work, we present…

Robotics · Computer Science 2025-05-14 Junda Huang , Jianshu Zhou , Honghao Guo , Yunhui Liu

This paper addresses the problem of multi-view people occupancy map estimation. Existing solutions for this problem either operate per-view, or rely on a background subtraction pre-processing. Both approaches lessen the detection…

Computer Vision and Pattern Recognition · Computer Science 2017-07-25 Tatjana Chavdarova , François Fleuret

Precise calibration is the basis for the vision-guided robot system to achieve high-precision operations. Systems with multiple eyes (cameras) and multiple hands (robots) are particularly sensitive to calibration errors, such as…

Robotics · Computer Science 2023-05-05 Zishun Zhou , Liping Ma , Xilong Liu , Zhiqiang Cao , Junzhi Yu

Object perception from multi-view cameras is crucial for intelligent systems, particularly in indoor environments, e.g., warehouses, retail stores, and hospitals. Most traditional multi-target multi-camera (MTMC) detection and tracking…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Yizhou Wang , Tim Meinhardt , Orcun Cetintas , Cheng-Yen Yang , Sameer Satish Pusegaonkar , Benjamin Missaoui , Sujit Biswas , Zheng Tang , Laura Leal-Taixé