中文
相关论文

相关论文: DeformMaster: An Interactive Physics-Neural World …

200 篇论文

Action-conditioned video models offer a promising path to building general-purpose robot simulators that can improve directly from data. Yet, despite training on large-scale robot datasets, current state-of-the-art video models still…

Controlling the shape of deformable linear objects using robots and constraints provided by environmental fixtures has diverse industrial applications. In order to establish robust contacts with these fixtures, accurate estimation of the…

机器人学 · 计算机科学 2024-01-31 Kejia Chen , Zhenshan Bing , Yansong Wu , Fan Wu , Liding Zhang , Sami Haddadin , Alois Knoll

We study the problem of imitating object interactions from Internet videos. This requires understanding the hand-object interactions in 4D, spatially in 3D and over time, which is challenging due to mutual hand-object occlusions. In this…

计算机视觉与模式识别 · 计算机科学 2022-11-24 Austin Patel , Andrew Wang , Ilija Radosavovic , Jitendra Malik

Recent video diffusion models have achieved impressive capabilities as large-scale generative world models. However, these models often struggle with fine-grained physical consistency, exhibiting physically implausible dynamics over time.…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Haoran Lu , Shang Wu , Jianshu Zhang , Maojiang Su , Guo Ye , Chenwei Xu , Lie Lu , Pranav Maneriker , Fan Du , Manling Li , Zhaoran Wang , Han Liu

World models, which predict future transitions from past observation and action sequences, have shown great promise for improving data efficiency in sequential decision-making. However, existing world models often require extensive…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Siqiao Huang , Jialong Wu , Qixing Zhou , Shangchen Miao , Mingsheng Long

Deformable object manipulation is a classical and challenging research area in robotics. Compared with rigid object manipulation, this problem is more complex due to the deformation properties including elastic, plastic, and elastoplastic…

机器人学 · 计算机科学 2024-05-14 Jianhua Shan , Yuhao Sun , Shixin Zhang , Fuchun Sun , Zixi Chen , Zirong Shen , Cesare Stefanini , Yiyong Yang , Shan Luo , Bin Fang

We introduce a novel, data-driven approach for reconstructing temporally coherent 3D motion from unstructured and potentially partial observations of non-rigidly deforming shapes. Our goal is to achieve high-fidelity motion reconstructions…

计算机视觉与模式识别 · 计算机科学 2025-09-09 Aymen Merrouche , Stefanie Wuhrer , Edmond Boyer

We address the challenge of learning to manipulate deformable objects with unknown dynamics. In non-rigid objects, the dynamics parameters define how they react to interactions -- how they stretch, bend, compress, and move -- and they are…

机器人学 · 计算机科学 2026-03-20 Bohan Wu , Roberto Martín-Martín , Li Fei-Fei

Accurate 3D shape abstraction from a single 2D image is a long-standing problem in computer vision and graphics. By leveraging a set of primitives to represent the target shape, recent methods have achieved promising results. However, these…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Di Liu , Xiang Yu , Meng Ye , Qilong Zhangli , Zhuowei Li , Zhixing Zhang , Dimitris N. Metaxas

Manipulation of deformable objects is a challenging task for a robot. It will be problematic to use a single sensory input to track the behaviour of such objects: vision can be subjected to occlusions, whereas tactile inputs cannot capture…

机器人学 · 计算机科学 2023-05-01 Leszek Pecyna , Siyuan Dong , Shan Luo

We introduce PhysWorld, a framework that enables robot learning from video generation through physical world modeling. Recent video generation models can synthesize photorealistic visual demonstrations from language commands and images,…

We present an efficient spacetime optimization method to automatically generate animations for a general volumetric, elastically deformable body. Our approach can model the interactions between the body and the environment and automatically…

图形学 · 计算机科学 2017-09-11 Zherong Pan , Dinesh Manocha

Manipulating volumetric deformable objects in the real world, like plush toys and pizza dough, bring substantial challenges due to infinite shape variations, non-rigid motions, and partial observability. We introduce ACID, an…

计算机视觉与模式识别 · 计算机科学 2022-08-09 Bokui Shen , Zhenyu Jiang , Christopher Choy , Leonidas J. Guibas , Silvio Savarese , Anima Anandkumar , Yuke Zhu

Neural rendering has demonstrated remarkable success in dynamic scene reconstruction. Thanks to the expressiveness of neural representations, prior works can accurately capture the motion and achieve high-fidelity reconstruction of the…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Hengyi Wang , Jingwen Wang , Lourdes Agapito

In this paper, we aim to create physical digital twins of deformable objects under interaction. Existing methods focus more on the physical learning of current state modeling, but generalize worse to future prediction. This is because…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Qingshan Xu , Jiao Liu , Shangshu Yu , Yuxuan Wang , Yuan Zhou , Junbao Zhou , Jiequan Cui , Yew-Soon Ong , Hanwang Zhang

Collocated tactile sensing is a fundamental enabling technology for dexterous manipulation. However, deformable sensors introduce complex dynamics between the robot, grasped object, and environment that must be considered for fine…

机器人学 · 计算机科学 2022-09-28 Miquel Oller , Mireia Planas , Dmitry Berenson , Nima Fazeli

Video transformers have recently emerged as an effective alternative to convolutional networks for action classification. However, most prior video transformers adopt either global space-time attention or hand-defined strategies to compare…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Jue Wang , Lorenzo Torresani

Demystifying complex human-ground interactions is essential for accurate and realistic 3D human motion reconstruction from RGB videos, as it ensures consistency between the humans and the ground plane. Prior methods have modeled…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Sihan Ma , Qiong Cao , Hongwei Yi , Jing Zhang , Dacheng Tao

Generative models trained on internet data have revolutionized how text, image, and video content can be created. Perhaps the next milestone for generative models is to simulate realistic experience in response to actions taken by humans,…

Teaching robots to fold, drape, or reposition deformable objects such as cloth will unlock a variety of automation applications. While remarkable progress has been made for rigid object manipulation, manipulating deformable objects poses…