English
Related papers

Related papers: AROS: Affordance Recognition with One-Shot Human S…

200 papers

Pose estimation of the human body and hands is a fundamental problem in computer vision, and learning-based solutions require a large amount of annotated data. In this work, we improve the efficiency of the data annotation process for 3D…

Computer Vision and Pattern Recognition · Computer Science 2023-01-19 Qi Feng , Kun He , He Wen , Cem Keskin , Yuting Ye

Understanding object and its context are very important for robots when dealing with objects for completion of a mission. In this paper, an Affordance-based Ontology (ABO) is proposed for easy robot dealing with substantive and…

Robotics · Computer Science 2012-05-08 Sidiq S. Hidayat , Bong Keung Kim , Kohtaro Ohba

Despite strong results on recognition and segmentation, current 3D visual pre-training methods often underperform on robotic manipulation. We attribute this gap to two factors: the lack of state-action-state dynamics modeling and the…

Robotics · Computer Science 2026-03-11 Qiwei Liang , Boyang Cai , Minghao Lai , Sitong Zhuang , Tao Lin , Yan Qin , Yixuan Ye , Jiaming Liang , Renjing Xu

The accuracy of 3D Human Pose and Shape reconstruction (HPS) from an image is progressively improving. Yet, no known method is robust across all image distortion. To address issues due to variations of camera poses, we introduce SHARE, a…

Computer Vision and Pattern Recognition · Computer Science 2024-01-02 Shreelekha Revankar , Shijia Liao , Yu Shen , Junbang Liang , Huaishu Peng , Ming Lin

We introduce Audio-Visual Affordance Grounding (AV-AG), a new task that segments object interaction regions from action sounds. Unlike existing approaches that rely on textual instructions or demonstration videos, which often limited by…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Lidong Lu , Guo Chen , Zhu Wei , Yicheng Liu , Tong Lu

The ability to successfully grasp objects is crucial in robotics, as it enables several interactive downstream applications. To this end, most approaches either compute the full 6D pose for the object of interest or learn to predict a set…

This paper presents a new system to obtain dense object reconstructions along with 6-DoF poses from a single image. Geared towards high fidelity reconstruction, several recent approaches leverage implicit surface representations and deep…

Computer Vision and Pattern Recognition · Computer Science 2020-04-28 Aniket Pokale , Aditya Aggarwal , K. Madhava Krishna

Recent methods for 6D pose estimation of objects assume either textured 3D models or real images that cover the entire range of target poses. However, it is difficult to obtain textured 3D models and annotate the poses of objects in real…

Computer Vision and Pattern Recognition · Computer Science 2020-11-04 Kiru Park , Timothy Patten , Markus Vincze

Estimating 3D human motion from an egocentric video sequence plays a critical role in human behavior understanding and has various applications in VR/AR. However, naively learning a mapping between egocentric videos and human motions is…

Computer Vision and Pattern Recognition · Computer Science 2023-08-29 Jiaman Li , C. Karen Liu , Jiajun Wu

Generating realistic full-body motion interacting with objects is critical for applications in robotics, virtual reality, and human-computer interaction. While existing methods can generate full-body motion within 3D scenes, they often lack…

Computer Vision and Pattern Recognition · Computer Science 2025-10-28 Kunal Bhosikar , Siddharth Katageri , Vivek Madhavaram , Kai Han , Charu Sharma

We consider the problem of estimating human pose and trajectory by an aerial robot with a monocular camera in near real time. We present a preliminary solution whose distinguishing feature is a dynamic classifier selection architecture. In…

Computer Vision and Pattern Recognition · Computer Science 2018-12-18 Asanka G Perera , Yee Wei Law , Javaan Chahl

In this work we study the use of 3D hand poses to recognize first-person dynamic hand actions interacting with 3D objects. Towards this goal, we collected RGB-D video sequences comprised of more than 100K frames of 45 daily hand action…

Computer Vision and Pattern Recognition · Computer Science 2018-04-11 Guillermo Garcia-Hernando , Shanxin Yuan , Seungryul Baek , Tae-Kyun Kim

Deducing a 3D human pose from a single 2D image is inherently challenging because multiple 3D poses can correspond to the same 2D representation. 3D data can resolve this pose ambiguity, but it is expensive to record and requires an…

Computer Vision and Pattern Recognition · Computer Science 2025-05-09 Christian Keilstrup Ingwersen , Rasmus Tirsgaard , Rasmus Nylander , Janus Nørtoft Jensen , Anders Bjorholm Dahl , Morten Rieger Hannemose

In this work we propose to utilize information about human actions to improve pose estimation in monocular videos. To this end, we present a pictorial structure model that exploits high-level information about activities to incorporate…

Computer Vision and Pattern Recognition · Computer Science 2017-02-13 Umar Iqbal , Martin Garbade , Juergen Gall

The wide application of pre-trained models is driving the trend of once-for-all training in one-shot neural architecture search (NAS). However, training within a huge sample space damages the performance of individual subnets and requires…

Networking and Internet Architecture · Computer Science 2023-06-19 Haibin Wang , Ce Ge , Hesen Chen , Xiuyu Sun

Detecting beneficial feature interactions is essential in recommender systems, and existing approaches achieve this by examining all the possible feature interactions. However, the cost of examining all the possible higher-order feature…

Information Retrieval · Computer Science 2022-06-29 Yixin Su , Yunxiang Zhao , Sarah Erfani , Junhao Gan , Rui Zhang

This paper tackles the task of semi-supervised video object segmentation, i.e., the separation of an object from the background in a video, given the mask of the first frame. We present One-Shot Video Object Segmentation (OSVOS), based on a…

Computer Vision and Pattern Recognition · Computer Science 2017-04-14 Sergi Caelles , Kevis-Kokitsi Maninis , Jordi Pont-Tuset , Laura Leal-Taixé , Daniel Cremers , Luc Van Gool

The 3D world limits the human body pose and the human body pose conveys information about the surrounding objects. Indeed, from a single image of a person placed in an indoor scene, we as humans are adept at resolving ambiguities of the…

Computer Vision and Pattern Recognition · Computer Science 2021-04-19 Zhenzhen Weng , Serena Yeung

We propose a training-free method, Articulate3D, to pose a 3D asset through language control. Despite advances in vision and language models, this task remains surprisingly challenging. To achieve this goal, we decompose the problem into…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 Oishi Deb , Anjun Hu , Ashkan Khakzar , Philip Torr , Christian Rupprecht

The dominant paradigm in 3D human pose estimation that lifts a 2D pose sequence to 3D heavily relies on long-term temporal clues (i.e., using a daunting number of video frames) for improved accuracy, which incurs performance saturation,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-10 Qitao Zhao , Ce Zheng , Mengyuan Liu , Chen Chen
‹ Prev 1 8 9 10 Next ›