中文
相关论文

相关论文: kPAM: KeyPoint Affordances for Category-Level Robo…

200 篇论文

Object pose estimation is an important component of most vision pipelines for embodied agents, as well as in 3D vision more generally. In this paper we tackle the problem of estimating the pose of novel object categories in a zero-shot…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Walter Goodwin , Sagar Vaze , Ioannis Havoutis , Ingmar Posner

Clothes manipulation, such as folding or hanging, is a critical capability for home service robots. Despite recent advances, most existing methods remain limited to specific clothes types and tasks, due to the complex, high-dimensional…

机器人学 · 计算机科学 2025-10-20 Yuhong Deng , Chao Tang , Cunjun Yu , Linfeng Li , David Hsu

Category-level 6D object pose estimation aims to estimate the rotation, translation and size of unseen instances within specific categories. In this area, dense correspondence-based methods have achieved leading performance. However, they…

计算机视觉与模式识别 · 计算机科学 2024-03-29 Xiao Lin , Wenfei Yang , Yuan Gao , Tianzhu Zhang

We consider a category-level perception problem, where one is given 2D or 3D sensor data picturing an object of a given category (e.g., a car), and has to reconstruct the 3D pose and shape of the object despite intra-class variability…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Jingnan Shi , Heng Yang , Luca Carlone

Humans perceive and interact with the world with the awareness of equivariance, facilitating us in manipulating different objects in diverse poses. For robotic manipulation, such equivariance also exists in many scenarios. For example, no…

机器人学 · 计算机科学 2024-08-08 Yue Chen , Chenrui Tie , Ruihai Wu , Hao Dong

In order to meaningfully interact with the world, robot manipulators must be able to interpret objects they encounter. A critical aspect of this interpretation is pose estimation: inferring quantities that describe the position and…

机器人学 · 计算机科学 2023-05-23 Walter Goodwin , Ioannis Havoutis , Ingmar Posner

End-to-end robot manipulation policies offer significant potential for enabling embodied agents to understand and interact with the world. Unlike traditional modular pipelines, end-to-end learning mitigates key limitations such as…

机器人学 · 计算机科学 2025-09-26 Dekun Lu , Wei Gao , Kui Jia

Category-level object pose and shape estimation from a single depth image has recently drawn research attention due to its potential utility for tasks such as robotics manipulation. The task is particularly challenging because the three…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Yihao Zhang , Harpreet S. Sawhney , John J. Leonard

Category-level articulated object pose estimation aims to estimate a hierarchy of articulation-aware object poses of an unseen articulated object from a known category. To reduce the heavy annotations needed for supervised learning methods,…

计算机视觉与模式识别 · 计算机科学 2023-03-01 Xueyi Liu , Ji Zhang , Ruizhen Hu , Haibin Huang , He Wang , Li Yi

Empowering autonomous agents with 3D understanding for daily objects is a grand challenge in robotics applications. When exploring in an unknown environment, existing methods for object pose estimation are still not satisfactory due to the…

计算机视觉与模式识别 · 计算机科学 2023-02-02 Guanglin Li , Yifeng Li , Zhichao Ye , Qihang Zhang , Tao Kong , Zhaopeng Cui , Guofeng Zhang

Category-level object pose estimation aims to find 6D object poses of previously unseen object instances from known categories without access to object CAD models. To reduce the huge amount of pose annotations needed for category-level…

计算机视觉与模式识别 · 计算机科学 2021-11-02 Xiaolong Li , Yijia Weng , Li Yi , Leonidas Guibas , A. Lynn Abbott , Shuran Song , He Wang

Keypoint detection is an essential building block for many robotic applications like motion capture and pose estimation. Historically, keypoints are detected using uniquely engineered markers such as checkerboards or fiducials. More…

机器人学 · 计算机科学 2023-02-28 Jingpei Lu , Florian Richter , Michael Yip

Category-agnostic pose estimation (CAPE) aims to predict keypoints for arbitrary classes given a few support images annotated with keypoints. Existing methods only rely on the features extracted at support keypoints to predict or refine the…

计算机视觉与模式识别 · 计算机科学 2024-03-21 Junjie Chen , Jiebin Yan , Yuming Fang , Li Niu

Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background features. Existing keypoint-based approaches can focus on manipulation-centric features and be…

机器人学 · 计算机科学 2026-04-17 Anukriti Singh , Kasra Torshizi , Khuzema Habib , Kelin Yu , Ruohan Gao , Pratap Tokekar

For years, researchers have been devoted to generalizable object perception and manipulation, where cross-category generalizability is highly desired yet underexplored. In this work, we propose to learn such cross-category skills via…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Haoran Geng , Helin Xu , Chengyang Zhao , Chao Xu , Li Yi , Siyuan Huang , He Wang

Promising results have been achieved recently in category-level manipulation that generalizes across object instances. Nevertheless, it often requires expensive real-world data collection and manual specification of semantic keypoints for…

机器人学 · 计算机科学 2022-05-09 Bowen Wen , Wenzhao Lian , Kostas Bekris , Stefan Schaal

Reorienting objects by using supports is a practical yet challenging manipulation task. Owing to the intricate geometry of objects and the constrained feasible motions of the robot, multiple manipulation steps are required for object…

机器人学 · 计算机科学 2023-08-30 Peng Xu , Hu Cheng , Jiankun Wang , Max Q. -H. Meng

This paper studies the task of any objects grasping from the known categories by free-form language instructions. This task demands the technique in computer vision, natural language processing, and robotics. We bring these disciplines…

机器人学 · 计算机科学 2022-05-10 Chilam Cheang , Haitao Lin , Yanwei Fu , Xiangyang Xue

Bimanual manipulation is a challenging yet crucial robotic capability, demanding precise spatial localization and versatile motion trajectories, which pose significant challenges to existing approaches. Existing approaches fall into two…

机器人学 · 计算机科学 2025-04-25 Yuyin Yang , Zetao Cai , Yang Tian , Jia Zeng , Jiangmiao Pang

Robotic manipulation in everyday scenarios, especially in unstructured environments, requires skills in pose-aware object manipulation (POM), which adapts robots' grasping and handling according to an object's 6D pose. Recognizing an…

机器人学 · 计算机科学 2024-03-21 Qiaojun Yu , Ce Hao , Junbo Wang , Wenhai Liu , Liu Liu , Yao Mu , Yang You , Hengxu Yan , Cewu Lu