English
Related papers

Related papers: AKB-48: A Real-World Articulated Object Knowledge …

200 papers

A deep understanding of kinematic structures and movable components is essential for enabling robots to manipulate objects and model their own articulated forms. Such understanding is captured through articulated objects, which are…

Robotics · Computer Science 2026-03-04 Jiawei Wang , Dingyou Wang , Jiaming Hu , Qixuan Zhang , Jingyi Yu , Lan Xu

We propose Neural 3D Articulation Prior (NAP), the first 3D deep generative model to synthesize 3D articulated object models. Despite the extensive research on generating 3D objects, compositions, or scenes, there remains a lack of focus on…

Computer Vision and Pattern Recognition · Computer Science 2023-05-26 Jiahui Lei , Congyue Deng , Bokui Shen , Leonidas Guibas , Kostas Daniilidis

Robots in human environments will need to interact with a wide variety of articulated objects such as cabinets, drawers, and dishwashers while assisting humans in performing day-to-day tasks. Existing methods either require objects to be…

Robotics · Computer Science 2021-07-21 Ajinkya Jain , Rudolf Lioutikov , Caleb Chuck , Scott Niekum

Advancing robotic manipulation of deformable objects can enable automation of repetitive tasks across multiple industries, from food processing to textiles and healthcare. Yet robots struggle with the high dimensionality of deformable…

Robotics · Computer Science 2024-09-26 Jan Obrist , Miguel Zamora , Hehui Zheng , Juan Zarate , Robert K. Katzschmann , Stelian Coros

Given a single image of a general object such as a chair, could we also restore its articulated 3D shape similar to human modeling, so as to animate its plausible articulations and diverse motions? This is an interesting new question that…

Computer Vision and Pattern Recognition · Computer Science 2022-07-07 Ji Yang , Xinxin Zuo , Sen Wang , Zhenbo Yu , Xingyu Li , Bingbing Ni , Minglun Gong , Li Cheng

A vision model with general-purpose object-level 3D understanding should be capable of inferring both 2D (e.g., class name and bounding box) and 3D information (e.g., 3D location and 3D viewpoint) for arbitrary rigid objects in natural…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Wufei Ma , Guanning Zeng , Guofeng Zhang , Qihao Liu , Letian Zhang , Adam Kortylewski , Yaoyao Liu , Alan Yuille

The ability to understand and generate similes is an imperative step to realize human-level AI. However, there is still a considerable gap between machine intelligence and human cognition in similes, since deep models based on statistical…

Computation and Language · Computer Science 2022-12-13 Qianyu He , Xintao Wang , Jiaqing Liang , Yanghua Xiao

We propose an unsupervised vision-based system to estimate the joint configurations of the robot arm from a sequence of RGB or RGB-D images without knowing the model a priori, and then adapt it to the task of category-independent…

Computer Vision and Pattern Recognition · Computer Science 2020-12-02 Qihao Liu , Weichao Qiu , Weiyao Wang , Gregory D. Hager , Alan L. Yuille

Rigging and skinning are essential steps to create realistic 3D animations, often requiring significant expertise and manual effort. Traditional attempts at automating these processes rely heavily on geometric heuristics and often struggle…

Graphics · Computer Science 2025-07-08 Yufan Deng , Yuhao Zhang , Chen Geng , Shangzhe Wu , Jiajun Wu

We present ConceptFactory, a novel scope to facilitate more efficient annotation of 3D object knowledge by recognizing 3D objects through generalized concepts (i.e. object conceptualization), aiming at promoting machine intelligence to…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Jianhua Sun , Yuxuan Li , Longfei Xu , Nange Wang , Jiude Wei , Yining Zhang , Cewu Lu

To interact with daily-life articulated objects of diverse structures and functionalities, understanding the object parts plays a central role in both user instruction comprehension and task execution. However, the possible discordance…

Robotics · Computer Science 2024-04-02 Haoran Geng , Songlin Wei , Congyue Deng , Bokui Shen , He Wang , Leonidas Guibas

In order for robots to operate effectively in homes and workplaces, they must be able to manipulate the articulated objects common within environments built for and by humans. Previous work learns kinematic models that prescribe this…

Robotics · Computer Science 2016-07-04 Zhengyang Wu , Mohit Bansal , Matthew R. Walter

Visual object counting is a fundamental computer vision task underpinning numerous real-world applications, from cell counting in biomedicine to traffic and wildlife monitoring. However, existing methods struggle to handle the challenge of…

Computer Vision and Pattern Recognition · Computer Science 2025-07-31 Corentin Dumery , Noa Etté , Aoxiang Fan , Ren Li , Jingyi Xu , Hieu Le , Pascal Fua

Architecture embodies aesthetic, cultural, and historical values, standing as a tangible testament to human civilization. Researchers have long leveraged virtual reality (VR), mixed reality (MR), and augmented reality (AR) to enable…

Graphics · Computer Science 2025-09-26 Yuze Wang , Luo Yang , Junyi Wang , Yue Qi

Surgery monitoring in Mixed Reality (MR) environments has recently received substantial focus due to its importance in image-based decisions, skill assessment, and robot-assisted surgery. Tracking hands and articulated surgical instruments…

Computer Vision and Pattern Recognition · Computer Science 2024-10-03 Ahmed Tawfik Aboukhadra , Nadia Robertini , Jameel Malik , Ahmed Elhayek , Gerd Reis , Didier Stricker

Humans commonly work with multiple objects in daily life and can intuitively transfer manipulation skills to novel objects by understanding object functional regularities. However, existing technical approaches for analyzing and…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Yun Liu , Haolin Yang , Xu Si , Ling Liu , Zipeng Li , Yuxiang Zhang , Yebin Liu , Li Yi

Recent advances in Vision-Language-Action (VLA) and world-model methods have improved generalization in tasks such as robotic manipulation and object interaction. However, Successful execution of such tasks depends on large, costly…

Robotics · Computer Science 2026-03-16 Yulu Wu , Jiujun Cheng , Haowen Wang , Dengyang Suo , Pei Ren , Qichao Mao , Shangce Gao , Yakun Huang

Rigged objects are commonly used in artist pipelines, as they can flexibly adapt to different scenes and postures. However, articulating the rigs into realistic affordance-aware postures (e.g., following the context, respecting the physics…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Yu-Chu Yu , Chieh Hubert Lin , Hsin-Ying Lee , Chaoyang Wang , Yu-Chiang Frank Wang , Ming-Hsuan Yang

The complexity of the visual world creates significant challenges for comprehensive visual understanding. In spite of recent successes in visual recognition, today's vision systems would still struggle to deal with visual queries that…

Computer Vision and Pattern Recognition · Computer Science 2015-11-11 Yuke Zhu , Ce Zhang , Christopher Ré , Li Fei-Fei

This paper studies the problem of fixing malfunctional 3D objects. While previous works focus on building passive perception models to learn the functionality from static 3D objects, we argue that functionality is reckoned with respect to…

Computer Vision and Pattern Recognition · Computer Science 2022-05-06 Yining Hong , Kaichun Mo , Li Yi , Leonidas J. Guibas , Antonio Torralba , Joshua B. Tenenbaum , Chuang Gan
‹ Prev 1 4 5 6 7 8 10 Next ›