English
Related papers

Related papers: CAGE: Controllable Articulation GEneration

200 papers

Capturing and re-animating the 3D structure of articulated objects present significant barriers. On one hand, methods requiring extensively calibrated multi-view setups are prohibitively complex and resource-intensive, limiting their…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Heng Yu , Joel Julin , Zoltán Á. Milacski , Koichiro Niinuma , László A. Jeni

Interactive 3D simulated objects are crucial in AR/VR, animations, and robotics, driving immersive experiences and advanced automation. However, creating these articulated objects requires extensive human effort and expertise, limiting…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Long Le , Jason Xie , William Liang , Hung-Ju Wang , Yue Yang , Yecheng Jason Ma , Kyle Vedder , Arjun Krishna , Dinesh Jayaraman , Eric Eaton

Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods. In this work, we introduce Part$^{2}$GS, a novel framework for modeling articulated digital…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Tianjiao Yu , Vedant Shah , Muntasir Wahed , Ying Shen , Kiet A. Nguyen , Ismini Lourentzou

We tackle the challenge of concurrent reconstruction at the part level with the RGB appearance and estimation of motion parameters for building digital twins of articulated objects using the 3D Gaussian Splatting (3D-GS) method. With two…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Junfu Guo , Yu Xin , Gaoyi Liu , Kai Xu , Ligang Liu , Ruizhen Hu

Animating an object in 3D often requires an articulated structure, e.g. a kinematic chain or skeleton of the manipulated object with proper skinning weights, to obtain smooth movements and surface deformations. However, existing models that…

Computer Vision and Pattern Recognition · Computer Science 2023-04-17 Tianshu Kuai , Akash Karthikeyan , Yash Kant , Ashkan Mirzaei , Igor Gilitschenski

We propose a novel task of text-controlled human object interaction generation in 3D scenes with movable objects. Existing human-scene interaction datasets suffer from insufficient interaction categories and typically only consider…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Xinhao Cai , Minghang Zheng , Xin Jin , Yang Liu

Object geometry is key information for robot manipulation. Yet, object reconstruction is a challenging task because cameras only capture partial observations of objects, especially when occlusion occurs. In this paper, we leverage two extra…

Computer Vision and Pattern Recognition · Computer Science 2025-12-05 Minghan Zhu , Zhiyi Wang , Qihang Sun , Maani Ghaffari , Michael Posa

Object functionality is often expressed through part articulation -- as when the two rigid parts of a scissor pivot against each other to perform the cutting function. Such articulations are often similar across objects within the same…

Computer Vision and Pattern Recognition · Computer Science 2018-09-21 Li Yi , Haibin Huang , Difan Liu , Evangelos Kalogerakis , Hao Su , Leonidas Guibas

Reconstructing articulated objects is essential for building digital twins of interactive environments. However, prior methods typically decouple geometry and motion by first reconstructing object shape in distinct states and then…

Computer Vision and Pattern Recognition · Computer Science 2025-11-13 Licheng Shen , Saining Zhang , Honghan Li , Peilin Yang , Zihao Huang , Zongzheng Zhang , Hao Zhao

This paper introduces the first text-guided work for generating the sequence of hand-object interaction in 3D. The main challenge arises from the lack of labeled data where existing ground-truth datasets are nowhere near generalizable in…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Junuk Cha , Jihyeon Kim , Jae Shin Yoon , Seungryul Baek

As humans, we aspire to create media content that is both freely willed and readily controlled. Thanks to the prominent development of generative techniques, we now can easily utilize 2D diffusion methods to synthesize images controlled by…

Graphics · Computer Science 2024-05-15 Wenqi Dong , Bangbang Yang , Lin Ma , Xiao Liu , Liyuan Cui , Hujun Bao , Yuewen Ma , Zhaopeng Cui

Human-object interactions with articulated objects are common in everyday life. Despite much progress in single-view 3D reconstruction, it is still challenging to infer an articulated 3D object model from an RGB video showing a person…

Computer Vision and Pattern Recognition · Computer Science 2022-09-14 Sanjay Haresh , Xiaohao Sun , Hanxiao Jiang , Angel X. Chang , Manolis Savva

With the onset of diffusion-based generative models and their ability to generate text-conditioned images, content generation has received a massive invigoration. Recently, these models have been shown to provide useful guidance for the…

Computer Vision and Pattern Recognition · Computer Science 2023-11-30 Alexander Vilesov , Pradyumna Chari , Achuta Kadambi

Our work aims to reconstruct hand-held objects given a single RGB image. In contrast to prior works that typically assume known 3D templates and reduce the problem to 3D pose estimation, our work reconstructs generic hand-held object…

Computer Vision and Pattern Recognition · Computer Science 2022-04-15 Yufei Ye , Abhinav Gupta , Shubham Tulsiani

Understanding and manipulating articulated objects, such as doors and drawers, is crucial for robots operating in human environments. We wish to develop a system that can learn to articulate novel objects with no prior interaction, after…

Robotics · Computer Science 2024-05-03 Harry Zhang , Ben Eisner , David Held

In the realm of digital creativity, our potential to craft intricate 3D worlds from imagination is often hampered by the limitations of existing digital tools, which demand extensive expertise and efforts. To narrow this disparity, we…

Computer Vision and Pattern Recognition · Computer Science 2024-06-21 Longwen Zhang , Ziyu Wang , Qixuan Zhang , Qiwei Qiu , Anqi Pang , Haoran Jiang , Wei Yang , Lan Xu , Jingyi Yu

We tackle the problem of text-driven 3D generation from a geometry alignment perspective. Given a set of text prompts, we aim to generate a collection of objects with semantically corresponding parts aligned across them. Recent methods…

3D medical image generation is essential for data augmentation and patient privacy, calling for reliable and efficient models suited for clinical practice. However, current methods suffer from limited anatomical fidelity, restricted axial…

Computer Vision and Pattern Recognition · Computer Science 2025-12-01 Minye Shao , Xingyu Miao , Haoran Duan , Zeyu Wang , Jingkun Chen , Yawen Huang , Xian Wu , Jingjing Deng , Yang Long , Yefeng Zheng

We present DIPO, a novel framework for the controllable generation of articulated 3D objects from a pair of images: one depicting the object in a resting state and the other in an articulated state. Compared to the single-image approach,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-29 Ruiqi Wu , Xinjie Wang , Liu Liu , Chunle Guo , Jiaxiong Qiu , Chongyi Li , Lichao Huang , Zhizhong Su , Ming-Ming Cheng

Humans intuitively perceive object shape and orientation from a single image, guided by strong priors about canonical poses. However, existing 3D generative models often produce misaligned results due to inconsistent training data, limiting…

Computer Vision and Pattern Recognition · Computer Science 2025-11-26 Yichong Lu , Yuzhuo Tian , Zijin Jiang , Yikun Zhao , Yuanbo Yang , Hao Ouyang , Haoji Hu , Huimin Yu , Yujun Shen , Yiyi Liao