English
Related papers

Related papers: Benchmarking Monocular 3D Dog Pose Estimation Usin…

200 papers

We explore 3D human pose estimation from a single RGB image. While many approaches try to directly predict 3D pose from image measurements, we explore a simple architecture that reasons through intermediate 2D pose predictions. Our approach…

Computer Vision and Pattern Recognition · Computer Science 2017-04-12 Ching-Hang Chen , Deva Ramanan

We present EgoHumans, a new multi-view multi-human video benchmark to advance the state-of-the-art of egocentric human 3D pose estimation and tracking. Existing egocentric benchmarks either capture single subject or indoor-only scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Rawal Khirodkar , Aayush Bansal , Lingni Ma , Richard Newcombe , Minh Vo , Kris Kitani

We introduce an automatic, end-to-end method for recovering the 3D pose and shape of dogs from monocular internet images. The large variation in shape between dog breeds, significant occlusion and low quality of internet images makes this a…

Computer Vision and Pattern Recognition · Computer Science 2021-02-12 Benjamin Biggs , Oliver Boyne , James Charles , Andrew Fitzgibbon , Roberto Cipolla

Human pose estimation is a critical task in computer vision and sports biomechanics, with applications spanning sports science, rehabilitation, and biomechanical research. While significant progress has been made in monocular 3D pose…

Computer Vision and Pattern Recognition · Computer Science 2025-07-14 Calvin Yeung , Tomohiro Suzuki , Ryota Tanaka , Zhuoer Yin , Keisuke Fujii

One major challenge for 3D pose estimation from a single RGB image is the acquisition of sufficient training data. In particular, collecting large amounts of training data that contain unconstrained images and are annotated with accurate 3D…

Computer Vision and Pattern Recognition · Computer Science 2016-03-29 Hashim Yasin , Umar Iqbal , Björn Krüger , Andreas Weber , Juergen Gall

RGB-based 3D pose estimation methods have been successful with the development of deep learning and the emergence of high-quality 3D pose datasets. However, most existing methods do not operate well for testing images whose distribution is…

Computer Vision and Pattern Recognition · Computer Science 2025-02-26 Hansoo Park , Chanwoo Kim , Jihyeon Kim , Hoseong Cho , Nhat Nguyen Bao Truong , Taehwan Kim , Seungryul Baek

Monocular 3D pose estimators produce camera-centered skeletons, creating view-dependent kinematic signals that complicate comparative analysis in applications such as health and sports science. We present 3DPCNet, a compact,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Tharindu Ekanayake , Constantino Álvarez Casado , Miguel Bordallo López

In nature, the collective behavior of animals, such as flying birds is dominated by the interactions between individuals of the same species. However, the study of such behavior among the bird species is a complex process that humans cannot…

Computer Vision and Pattern Recognition · Computer Science 2022-08-01 Seyed Mojtaba Marvasti-Zadeh , Mohammad N. S. Jahromi , Javad Khaghani , Devin Goodsman , Nilanjan Ray , Nadir Erbilgin

Monocular estimation of 3d human pose has attracted increased attention with the availability of large ground-truth motion capture datasets. However, the diversity of training data available is limited and it is not clear to what extent…

Computer Vision and Pattern Recognition · Computer Science 2020-04-08 Zhe Wang , Daeyun Shin , Charless C. Fowlkes

Robotic manipulation requires accurate perception of the environment, which poses a significant challenge due to its inherent complexity and constantly changing nature. In this context, RGB image and point-cloud observations are two…

Robotics · Computer Science 2024-09-10 Boshi An , Yiran Geng , Kai Chen , Xiaoqi Li , Qi Dou , Hao Dong

Low-cost autonomous agents including autonomous driving vehicles chiefly adopt monocular 3D object detection to perceive surrounding environment. This paper studies 3D intermediate representation methods which generate intermediate 3D…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Qian Ye , Ling Jiang , Wang Zhen , Yuyang Du

Although fiducial markers give an accurate pose estimation in laboratory conditions, where the noisy factors are controlled, using them in field robotic applications remains a challenge. This is constrained to the fiducial maker systems,…

Robotics · Computer Science 2020-01-24 Luis A. Mateos

Understanding camera motion is a fundamental problem in embodied perception and 3D scene understanding. While visual methods have advanced rapidly, they often struggle under visually degraded conditions such as motion blur or occlusions. In…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Daniel Adebi , Sagnik Majumder , Kristen Grauman

Despite the significant improvement in the performance of monocular pose estimation approaches and their ability to generalize to unseen environments, multi-view (MV) approaches are often lagging behind in terms of accuracy and are specific…

Computer Vision and Pattern Recognition · Computer Science 2019-10-09 Abdolrahim Kadkhodamohammadi , Nicolas Padoy

Object pose estimation is a core perception task that enables, for example, object grasping and scene understanding. The widely available, inexpensive and high-resolution RGB sensors and CNNs that allow for fast inference based on this…

6D object pose estimation is one of the fundamental problems in computer vision and robotics research. While a lot of recent efforts have been made on generalizing pose estimation to novel object instances within the same category, namely…

Computer Vision and Pattern Recognition · Computer Science 2022-07-01 Yang Fu , Xiaolong Wang

Until recently Intelligence, Surveillance, and Reconnaissance (ISR) focused on acquiring behavioral information of the targets and their activities. Continuous evolution of intelligence being gathered of the human centric activities has put…

Computer Vision and Pattern Recognition · Computer Science 2014-10-07 Atul Kanaujia

Monocular 3D object detection (M3OD) is intrinsically ill-posed, hence training a high-performance deep learning based M3OD model requires a humongous amount of labeled data with complicated visual variation from diverse scenes, variety of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Zhaonian Kuang , Rui Ding , Meng Yang , Xinhu Zheng , Gang Hua

Monocular RGB cameras mounted on drones are widely used for wildlife monitoring, yet most analytical pipelines remain confined to two-dimensional image space, leaving geometric information in video underexploited. We present WildLIFT, a…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Vandita Shukla , Fabio Remondino , Blair Costelloe , Benjamin Risse

We introduce a model for monocular RGB relative pose estimation of a ground robot that trains from scratch without pose labels nor prior knowledge about the robot's shape or appearance. At training time, we assume: (i) a robot fitted with…

Robotics · Computer Science 2025-09-15 Nicholas Carlotti , Mirko Nava , Alessandro Giusti
‹ Prev 1 3 4 5 6 7 10 Next ›