English
Related papers

Related papers: DeformNet: Latent Space Modeling and Dynamics Pred…

200 papers

In the field of robotic manipulation, the proficiency of deformable object manipulation lags behind human capabilities due to the inherent characteristics of deformable objects. These objects have infinite degrees of freedom, resulting in…

Robotics · Computer Science 2023-11-17 Peng Zhou

We present a new model DrNET that learns disentangled image representations from video. Our approach leverages the temporal coherence of video and a novel adversarial loss to learn a representation that factorizes each frame into a…

Machine Learning · Computer Science 2024-03-15 Remi Denton , Vighnesh Birodkar

Understanding the dynamics of generic 3D scenes is fundamentally challenging in computer vision, essential in enhancing applications related to scene reconstruction, motion tracking, and avatar creation. In this work, we address the task as…

Computer Vision and Pattern Recognition · Computer Science 2024-06-07 Yan Zhang , Sergey Prokudin , Marko Mihajlovic , Qianli Ma , Siyu Tang

This work proposes DOFS, a pilot dataset of 3D deformable objects (DOs) (e.g., elasto-plastic objects) with full spatial information (i.e., top, side, and bottom information) using a novel and low-cost data collection platform with a…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Zhen Zhang , Xiangyu Chu , Yunxi Tang , K. W. Samuel Au

In this work, we bridge the gap between recent pose estimation and tracking work to develop a powerful method for robots to track objects in their surroundings. Motion-Nets use a segmentation model to segment the scene, and separate…

Robotics · Computer Science 2019-10-31 Felix Leeb , Arunkumar Byravan , Dieter Fox

Convolution neural networks (CNNs) and Transformers have their own advantages and both have been widely used for dense prediction in multi-task learning (MTL). Most of the current studies on MTL solely rely on CNN or Transformer. In this…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Yangyang Xu , Yibo Yang , Lefei Zhang

Recent years have witnessed significant progress in the field of neural surface reconstruction. While the extensive focus was put on volumetric and implicit approaches, a number of works have shown that explicit graphics primitives such as…

Computer Vision and Pattern Recognition · Computer Science 2023-04-07 Sergey Prokudin , Qianli Ma , Maxime Raafat , Julien Valentin , Siyu Tang

Neural rendering techniques combining machine learning with geometric reasoning have arisen as one of the most promising approaches for synthesizing novel views of a scene from a sparse set of images. Among these, stands out the Neural…

Computer Vision and Pattern Recognition · Computer Science 2020-12-01 Albert Pumarola , Enric Corona , Gerard Pons-Moll , Francesc Moreno-Noguer

In this paper, we present an attention-guided deformable convolutional network for hand-held multi-frame high dynamic range (HDR) imaging, namely ADNet. This problem comprises two intractable challenges of how to handle saturation and noise…

Computer Vision and Pattern Recognition · Computer Science 2021-05-25 Zhen Liu , Wenjie Lin , Xinpeng Li , Qing Rao , Ting Jiang , Mingyan Han , Haoqiang Fan , Jian Sun , Shuaicheng Liu

We introduce PhysXNet, a learning-based approach to predict the dynamics of deformable clothes given 3D skeleton motion sequences of humans wearing these clothes. The proposed model is adaptable to a large variety of garments and changing…

Computer Vision and Pattern Recognition · Computer Science 2021-11-16 Jordi Sanchez-Riera , Albert Pumarola , Francesc Moreno-Noguer

This paper proposes an anchor-based deformation model, namely AnchorDEF, to predict 3D garment animation from a body motion sequence. It deforms a garment mesh template by a mixture of rigid transformations with extra nonlinear…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Fang Zhao , Zekun Li , Shaoli Huang , Junwu Weng , Tianfei Zhou , Guo-Sen Xie , Jue Wang , Ying Shan

Representation learning becomes especially important for complex systems with multimodal data sources such as cameras or sensors. Recent advances in reinforcement learning and optimal control make it possible to design control algorithms on…

3D world models (i.e., learning-based 3D dynamics models) offer a promising approach to generalizable robotic manipulation by capturing the underlying physics of environment evolution conditioned on robot actions. However, existing 3D world…

Robotics · Computer Science 2025-08-27 Suning Huang , Qianzhong Chen , Xiaohan Zhang , Jiankai Sun , Mac Schwager

Robotic manipulation of deformable linear objects (DLOs) is an active area of research, though emerging applications, like automotive wire harness installation, introduce constraints that have not been considered in prior work. Confined…

Robotics · Computer Science 2024-02-19 Tyler Toner , Vahidreza Molazadeh , Miguel Saez , Dawn M. Tilbury , Kira Barton

Neuroimaging data, particularly from techniques like MRI or PET, offer rich but complex information about brain structure and activity. To manage this complexity, latent representation models - such as Autoencoders, Generative Adversarial…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 C. Vázquez-García , F. J. Martínez-Murcia , F. Segovia Román , Juan M. Górriz

In this paper, we propose deformable deep convolutional neural networks for generic object detection. This new deep learning object detection framework has innovations in multiple aspects. In the proposed new deep architecture, a new…

Computer Vision and Pattern Recognition · Computer Science 2015-06-03 Wanli Ouyang , Xiaogang Wang , Xingyu Zeng , Shi Qiu , Ping Luo , Yonglong Tian , Hongsheng Li , Shuo Yang , Zhe Wang , Chen-Change Loy , Xiaoou Tang

In many real-world settings, image observations of freely rotating 3D rigid bodies may be available when low-dimensional measurements are not. However, the high-dimensionality of image data precludes the use of classical estimation…

Computer Vision and Pattern Recognition · Computer Science 2024-04-12 Justice Mason , Christine Allen-Blanchette , Nicholas Zolman , Elizabeth Davison , Naomi Ehrich Leonard

Data-driven deep learning approaches to image registration can be less accurate than conventional iterative approaches, especially when training data is limited. To address this whilst retaining the fast inference speed of deep learning, we…

Computer Vision and Pattern Recognition · Computer Science 2024-10-28 Xi Jia , Alexander Thorley , Wei Chen , Huaqi Qiu , Linlin Shen , Iain B Styles , Hyung Jin Chang , Ales Leonardis , Antonio de Marvao , Declan P. O'Regan , Daniel Rueckert , Jinming Duan

The last several years have seen significant progress in using depth cameras for tracking articulated objects such as human bodies, hands, and robotic manipulators. Most approaches focus on tracking skeletal parameters of a fixed shape…

Computer Vision and Pattern Recognition · Computer Science 2017-11-23 Aaron Walsman , Weilin Wan , Tanner Schmidt , Dieter Fox

3D shape representations that accommodate learning-based 3D reconstruction are an open problem in machine learning and computer graphics. Previous work on neural 3D reconstruction demonstrated benefits, but also limitations, of point cloud,…

Computer Vision and Pattern Recognition · Computer Science 2020-11-25 Jun Gao , Wenzheng Chen , Tommy Xiang , Clement Fuji Tsang , Alec Jacobson , Morgan McGuire , Sanja Fidler