中文
相关论文

相关论文: CloDS: Visual-Only Unsupervised Cloth Dynamics Lea…

200 篇论文

A key challenge for an agent learning to interact with the world is to reason about physical properties of objects and to foresee their dynamics under the effect of applied forces. In order to scale learning through interaction to many…

机器人学 · 计算机科学 2020-08-04 Iman Nematollahi , Oier Mees , Lukas Hermann , Wolfram Burgard

Robotic manipulation of cloth is a highly complex task because of its infinite-dimensional shape-state space that makes cloth state estimation very difficult. In this paper we introduce the dGLI Cloth Coordinates, a low-dimensional…

We propose a deep videorealistic 3D human character model displaying highly realistic shape, motion, and dynamic appearance learned in a new weakly supervised way from multi-view imagery. In contrast to previous work, our controllable 3D…

计算机视觉与模式识别 · 计算机科学 2021-09-01 Marc Habermann , Lingjie Liu , Weipeng Xu , Michael Zollhoefer , Gerard Pons-Moll , Christian Theobalt

Recent advances in 3D Gaussian Splatting (3DGS) enable real-time, high-fidelity novel view synthesis (NVS) with explicit 3D representations. However, performance degradation and instability remain significant under sparse-view conditions.…

计算机视觉与模式识别 · 计算机科学 2025-10-10 Meixi Song , Xin Lin , Dizhe Zhang , Haodong Li , Xiangtai Li , Bo Du , Lu Qi

Dynamic scene reconstruction is a long-term challenge in the field of 3D vision. Recently, the emergence of 3D Gaussian Splatting has provided new insights into this problem. Although subsequent efforts rapidly extend static 3D Gaussian to…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Ruijie Zhu , Yanzhe Liang , Hanzhi Chang , Jiacheng Deng , Jiahao Lu , Wenfei Yang , Tianzhu Zhang , Yongdong Zhang

Out-of-distribution (OOD) 3D relighting requires novel view synthesis under unseen lighting conditions that differ significantly from the observed images. Existing relighting methods, which assume consistent light source distributions…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Yumeng He , Yunbo Wang , Xiaokang Yang

Learning without supervision how to predict 3D scene flows from point clouds is essential to many perception systems. We propose a novel learning framework for this task which improves the necessary regularization. Relying on the assumption…

计算机视觉与模式识别 · 计算机科学 2024-08-14 Patrik Vacek , David Hurych , Karel Zimmermann , Patrick Perez , Tomas Svoboda

Recent neural, physics-based modeling of garment deformations allows faster and visually aesthetic results as opposed to the existing methods. Material-specific parameters are used by the formulation to control the garment inextensibility.…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Ruochen Chen , Liming Chen , Shaifali Parashar

For visual estimation of optical flow, a crucial function for many vision tasks, unsupervised learning, using the supervision of view synthesis has emerged as a promising alternative to supervised methods, since ground-truth flow is not…

计算机视觉与模式识别 · 计算机科学 2023-04-17 Zitang Sun , Shin'ya Nishida , Zhengbo Luo

3D Gaussian splatting has demonstrated impressive performance in real-time novel view synthesis. However, achieving successful reconstruction from RGB images generally requires multiple input views captured under static conditions. To…

计算机视觉与模式识别 · 计算机科学 2024-05-31 Wei Sun , Qi Zhang , Yanzhao Zhou , Qixiang Ye , Jianbin Jiao , Yuan Li

Multi-modal scene reconstruction integrating RGB and thermal infrared data is essential for robust environmental perception across diverse lighting and weather conditions. However, extending 3D Gaussian Splatting (3DGS) to multi-spectral…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Zhaoqi Su , Shihai Chen , Xinyan Lin , Liqin Huang , Zhipeng Su , Xiaoqiang Lu

We propose UnCLe, the first standardized benchmark for Unsupervised Continual Learning of a multimodal 3D reconstruction task: Depth completion aims to infer a dense depth map from a pair of synchronized RGB image and sparse depth map. We…

计算机视觉与模式识别 · 计算机科学 2025-06-11 Xien Chen , Rit Gangopadhyay , Michael Chu , Patrick Rim , Hyoungseob Park , Alex Wong

Recent works in 3D multimodal learning have made remarkable progress. However, typically 3D multimodal models are only capable of handling point clouds. Compared to the emerging 3D representation technique, 3D Gaussian Splatting (3DGS), the…

计算机视觉与模式识别 · 计算机科学 2026-01-13 Siyu Jiao , Haoye Dong , Yuyang Yin , Zequn Jie , Yinlong Qian , Yao Zhao , Humphrey Shi , Yunchao Wei

Gait recognition is instrumental in crime prevention and social security, for it can be conducted at a long distance to figure out the identity of persons. However, existing datasets and methods cannot satisfactorily deal with the most…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Xuqian Ren , Saihui Hou , Chunshui Cao , Xu Liu , Yongzhen Huang

Learning to predict scene depth from RGB inputs is a challenging task both for indoor and outdoor robot navigation. In this work we address unsupervised learning of scene depth and robot ego-motion where supervision is provided by monocular…

计算机视觉与模式识别 · 计算机科学 2018-11-16 Vincent Casser , Soeren Pirk , Reza Mahjourian , Anelia Angelova

In a constant evolving world, change detection is of prime importance to keep updated maps. To better sense areas with complex geometry (urban areas in particular), considering 3D data appears to be an interesting alternative to classical…

计算机视觉与模式识别 · 计算机科学 2024-10-28 Iris de Gélis , Sébastien Lefèvre , Thomas Corpetti

We present Contextualized Local Visual Embeddings (CLoVE), a self-supervised convolutional-based method that learns representations suited for dense prediction tasks. CLoVE deviates from current methods and optimizes a single loss function…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Thalles Santos Silva , Helio Pedrini , Adín Ramírez Rivera

Recent advances in 3D Gaussian Splatting have shown promising results. Existing methods typically assume static scenes and/or multiple images with prior poses. Dynamics, sparse views, and unknown poses significantly increase the problem…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Weihang Li , Weirong Chen , Shenhan Qian , Jiajie Chen , Daniel Cremers , Haoang Li

Perception of the environment is a critical component for enabling autonomous driving. It provides the vehicle with the ability to comprehend its surroundings and make informed decisions. Depth prediction plays a pivotal role in this…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Houssem Boulahbal

We address the unsupervised learning of several interconnected problems in low-level vision: single view depth prediction, camera motion estimation, optical flow, and segmentation of a video into the static scene and moving regions. Our key…

计算机视觉与模式识别 · 计算机科学 2019-03-13 Anurag Ranjan , Varun Jampani , Lukas Balles , Kihwan Kim , Deqing Sun , Jonas Wulff , Michael J. Black