English
Related papers

Related papers: JODA: Composable Joint Dynamics for Articulated Ob…

200 papers

3D models of manufactured objects are important for populating virtual worlds and for synthetic data generation for vision and robotics. To be most useful, such objects should be articulated: their parts should move when interacted with.…

Graphics · Computer Science 2022-06-20 Xianghao Xu , Yifan Ruan , Srinath Sridhar , Daniel Ritchie

Flow-matching methods for 3D shape assembly learn point-wise velocity fields that transport parts toward assembled configurations, yet they receive no explicit guidance about which cross-part interactions should drive the motion. We…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Nahyuk Lee , Zhiang Chen , Marc Pollefeys , Sunghwan Hong

Recent advancements in customized video generation have led to significant improvements in the simultaneous adaptation of appearance and motion. Typically, decoupling the appearance and motion training, prior methods often introduce concept…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Fangda Chen , Shanshan Zhao , Chuanfu Xu , Long Lan

Independently trained vision and language models inhabit disjoint representational spaces, shaped by their respective modalities, objectives, and architectures. The Platonic Representation Hypothesis (PRH) suggests these models may…

Machine Learning · Computer Science 2026-05-18 Lauren Hyoseo Yoon , Yisong Yue , Been Kim

Humans learn abstract concepts through multisensory synergy, and once formed, such representations can often be recalled from a single modality. Inspired by this principle, we introduce Concerto, a minimalist simulation of human concept…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Yujia Zhang , Xiaoyang Wu , Yixing Lao , Chengyao Wang , Zhuotao Tian , Naiyan Wang , Hengshuang Zhao

Dynamics simulation with frictional contacts is important for a wide range of applications, from cloth simulation to object manipulation. Recent methods using smoothed lagged friction forces have enabled robust and differentiable simulation…

Graphics · Computer Science 2024-08-01 Egor Larionov , Andreas Longva , Uri M. Ascher , Jan Bender , Dinesh K. Pai

This work introduces JEMA (Joint Embedding with Multimodal Alignment), a novel co-learning framework tailored for laser metal deposition (LMD), a pivotal process in metal additive manufacturing. As Industry 5.0 gains traction in industrial…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Joao Sousa , Roya Darabi , Armando Sousa , Frank Brueckner , Luís Paulo Reis , Ana Reis

Autonomous robots operating in real-world environments encounter a variety of objects that can be both rigid and articulated in nature. Having knowledge of these specific object properties not only helps in designing appropriate…

Computer Vision and Pattern Recognition · Computer Science 2023-05-12 Ayush Aggarwal , Rustam Stolkin , Naresh Marturi

Despite progress, Vision-Language-Action models (VLAs) are limited by a scarcity of large-scale, diverse robot data. While human manipulation videos offer a rich alternative, existing methods are forced to choose between small,…

Robotics · Computer Science 2026-02-26 Hao Luo , Ye Wang , Wanpeng Zhang , Haoqi Yuan , Yicheng Feng , Haiweng Xu , Sipeng Zheng , Zongqing Lu

The availability of low-cost range sensors and the development of relatively robust algorithms for the extraction of skeleton joint locations have inspired many researchers to develop human activity recognition methods using the 3-D data.…

Computer Vision and Pattern Recognition · Computer Science 2018-08-14 Saeed Ghodsi , Hoda Mohammadzade , Erfan Korki

We present Dojo, a differentiable physics engine for robotics that prioritizes stable simulation, accurate contact physics, and differentiability with respect to states, actions, and system parameters. Dojo models hard contact and friction…

We propose ChainHOI, a novel approach for text-driven human-object interaction (HOI) generation that explicitly models interactions at both the joint and kinetic chain levels. Unlike existing methods that implicitly model interactions using…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Ling-An Zeng , Guohong Huang , Yi-Lin Wei , Shengbo Gu , Yu-Ming Tang , Jingke Meng , Wei-Shi Zheng

Manipulating dynamic objects remains an open challenge for Vision-Language-Action (VLA) models, which, despite strong generalization in static manipulation, struggle in dynamic scenarios requiring rapid perception, temporal anticipation,…

Robotics · Computer Science 2026-01-30 Haozhe Xie , Beichen Wen , Jiarui Zheng , Zhaoxi Chen , Fangzhou Hong , Haiwen Diao , Ziwei Liu

Estimating 3D articulated shapes like animal bodies from monocular images is inherently challenging due to the ambiguities of camera viewpoint, pose, texture, lighting, etc. We propose ARTIC3D, a self-supervised framework to reconstruct…

Computer Vision and Pattern Recognition · Computer Science 2023-06-08 Chun-Han Yao , Amit Raj , Wei-Chih Hung , Yuanzhen Li , Michael Rubinstein , Ming-Hsuan Yang , Varun Jampani

Modern interactive applications increasingly demand dynamic 3D content, yet the transformation of static 3D models into animated assets constitutes a significant bottleneck in content creation pipelines. While recent advances in generative…

Computer Vision and Pattern Recognition · Computer Science 2025-08-15 Chaoyue Song , Xiu Li , Fan Yang , Zhongcong Xu , Jiacheng Wei , Fayao Liu , Jiashi Feng , Guosheng Lin , Jianfeng Zhang

Compositional Customized Image Generation aims to customize multiple target concepts within generation content, which has gained attention for its wild application. Existing approaches mainly concentrate on the target entity's appearance…

Computer Vision and Pattern Recognition · Computer Science 2025-08-29 Zhu Xu , Zhaowen Wang , Yuxin Peng , Yang Liu

Robotic manipulation can greatly benefit from the data efficiency, robustness, and predictability of model-based methods if robots can quickly generate models of novel objects they encounter. This is especially difficult when effects like…

Robotics · Computer Science 2023-10-19 Bibit Bianchini , Mathew Halm , Michael Posa

Semantics has enabled 3D scene understanding and affordance-driven object interaction. However, robots operating in real-world environments face a critical limitation: they cannot anticipate how objects move. Long-horizon mobile…

Controllable 3D human avatars have found widespread applications in 3D games, the metaverse, and AR/VR scenarios. The conventional approach to creating such a 3D avatar requires a lengthy, intricate pipeline encompassing appearance…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Jiahe Zhu , Xinyao Wang , Yiyu Zhuang , Yanwen Wang , Jing Tian , Yao Yao , Hao Zhu

We present DIMO, a generative approach capable of generating diverse 3D motions for arbitrary objects from a single image. The core idea of our work is to leverage the rich priors in well-trained video models to extract the common motion…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Linzhan Mou , Jiahui Lei , Chen Wang , Lingjie Liu , Kostas Daniilidis