English
Related papers

Related papers: EgoPoseFormer v2: Accurate Egocentric Human Motion…

200 papers

3D human pose estimation (HPE) in autonomous vehicles (AV) differs from other use cases in many factors, including the 3D resolution and range of data, absence of dense depth maps, failure modes for LiDAR, relative location between the…

Computer Vision and Pattern Recognition · Computer Science 2021-12-23 Jingxiao Zheng , Xinwei Shi , Alexander Gorban , Junhua Mao , Yang Song , Charles R. Qi , Ting Liu , Visesh Chari , Andre Cornman , Yin Zhou , Congcong Li , Dragomir Anguelov

3D hand pose estimation in everyday egocentric images is challenging for several reasons: poor visual signal (occlusion from the object of interaction, low resolution & motion blur), large perspective distortion (hands are close to the…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Aditya Prakash , Ruisen Tu , Matthew Chang , Saurabh Gupta

Egomotion estimation is crucial for applications such as autonomous navigation and robotics, where accurate and real-time motion tracking is required. However, traditional methods relying on inertial sensors are highly sensitive to external…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Hugh Greatorex , Michele Mastella , Madison Cotteret , Ole Richter , Elisabetta Chicca

We propose a stereo vision-based approach for tracking the camera ego-motion and 3D semantic objects in dynamic autonomous driving scenarios. Instead of directly regressing the 3D bounding box using end-to-end approaches, we propose to use…

Computer Vision and Pattern Recognition · Computer Science 2018-11-30 Peiliang Li , Tong Qin , Shaojie Shen

We present EgoTAP, a heatmap-to-3D pose lifting method for highly accurate stereo egocentric 3D pose estimation. Severe self-occlusion and out-of-view limbs in egocentric camera views make accurate pose estimation a challenging problem. To…

Computer Vision and Pattern Recognition · Computer Science 2024-02-29 Taeho Kang , Youngki Lee

Egocentric video gaze estimation requires models to capture individual gaze patterns while adapting to diverse user data. Our approach leverages a transformer-based architecture, integrating it into a PFL framework where only the most…

Computer Vision and Pattern Recognition · Computer Science 2025-02-26 Yuhu Feng , Keisuke Maeda , Takahiro Ogawa , Miki Haseyama

Existing volumetric methods for predicting 3D human pose estimation are accurate, but computationally expensive and optimized for single time-step prediction. We present TEMPO, an efficient multi-view pose estimation model that learns a…

Computer Vision and Pattern Recognition · Computer Science 2023-09-15 Rohan Choudhury , Kris Kitani , Laszlo A. Jeni

Robotic generalization relies on physical intelligence: the ability to reason about state changes, contact-rich interactions, and long-horizon planning under egocentric perception and action. Vision Language Models (VLMs) are essential to…

Egocentric gestures are the most natural form of communication for humans to interact with wearable devices such as VR/AR helmets and glasses. A major issue in such scenarios for real-world applications is that may easily become necessary…

Computer Vision and Pattern Recognition · Computer Science 2020-04-21 Zhengwei Wang , Qi She , Tejo Chalasani , Aljosa Smolic

This paper introduces EgoMAGIC (Medical Assistance, Guidance, Instruction, and Correction), an egocentric medical activity dataset collected as part of DARPA's Perceptually-enabled Task Guidance (PTG) program. This dataset comprises 3,355…

Computer Vision and Pattern Recognition · Computer Science 2026-04-27 Brian VanVoorst , Nicholas Walczak , Christopher Gilleo , Charles Meissner , Fabio Felix , Iran Roman , Bea Steers , Claudio Silva , Yuhan Shen , Zijia Lu , Shih-Po Lee , Ehsan Elhamifar

Egocentric human video data, which captures rich human-environment interactions and can be collected at scale, has become a key driver of embodied intelligence research. However, existing egocentric datasets typically lack tactile sensing,…

AI personal assistants deployed via robots or wearables require embodied understanding to collaborate with humans effectively. However, current Vision-Language Models (VLMs) primarily focus on third-person view videos, neglecting the…

Computer Vision and Pattern Recognition · Computer Science 2024-06-24 Alessandro Suglia , Claudio Greco , Katie Baker , Jose L. Part , Ioannis Papaioannou , Arash Eshghi , Ioannis Konstas , Oliver Lemon

In this report, we present our approach and empirical results of applying masked autoencoders in two egocentric video understanding tasks, namely, Object State Change Classification and PNR Temporal Localization, of Ego4D Challenge 2022. As…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Jiachen Lei , Shuang Ma , Zhongjie Ba , Sai Vemprala , Ashish Kapoor , Kui Ren

Consistent motion estimation is fundamental for all mobile autonomous systems. While this sounds like an easy task, often, it is not the case because of changing environmental conditions affecting odometry obtained from vision, Lidar, or…

Robotics · Computer Science 2022-04-20 Karim Haggag , Sven Lange , Tim Pfeifer , Peter Protzel

To enable a safe and effective human-robot cooperation, it is crucial to develop models for the identification of human activities. Egocentric vision seems to be a viable solution to solve this problem, and therefore many works provide deep…

Computer Vision and Pattern Recognition · Computer Science 2023-03-13 Gabriele Goletto , Mirco Planamente , Barbara Caputo , Giuseppe Averta

The advancement of robot learning is currently hindered by the scarcity of large-scale, high-quality datasets. While established data collection methods such as teleoperation and universal manipulation interfaces dominate current datasets,…

Today's Mixed Reality head-mounted displays track the user's head pose in world space as well as the user's hands for interaction in both Augmented Reality and Virtual Reality scenarios. While this is adequate to support user input, it…

Computer Vision and Pattern Recognition · Computer Science 2022-07-29 Jiaxi Jiang , Paul Streli , Huajian Qiu , Andreas Fender , Larissa Laich , Patrick Snape , Christian Holz

Embodied agents in household environments must plan under partial observation: they need to remember objects, track state changes, and recover when actions fail. Existing benchmarks only partially test this ability. Egocentric video…

Artificial Intelligence · Computer Science 2026-05-14 Qinchuan Cheng , Zhantao Gong , Pengzhan Sun , Angela Yao , Xulei Yang , Shijie Li

We envision a future time when wearable cameras are worn by the masses and recording first-person point-of-view videos of everyday life. While these cameras can enable new assistive technologies and novel research challenges, they also…

Computer Vision and Pattern Recognition · Computer Science 2017-11-30 Ryo Yonetani , Kris M. Kitani , Yoichi Sato

We focus on the task of everyday hand pose estimation from egocentric viewpoints. For this task, we show that depth sensors are particularly informative for extracting near-field interactions of the camera wearer with his/her environment.…

Computer Vision and Pattern Recognition · Computer Science 2014-12-02 Gregory Rogez , James S. Supancic , Maryam Khademi , Jose Maria Martinez Montiel , Deva Ramanan
‹ Prev 1 8 9 10 Next ›