中文
相关论文

相关论文: C3DPO: Canonical 3D Pose Networks for Non-Rigid St…

200 篇论文

High-fidelity human 3D models can now be learned directly from videos, typically by combining a template-based surface model with neural representations. However, obtaining a template surface requires expensive multi-view capture systems,…

计算机视觉与模式识别 · 计算机科学 2023-09-04 Shih-Yang Su , Timur Bagautdinov , Helge Rhodin

We introduce PlatonicGAN to discover the 3D structure of an object class from an unstructured collection of 2D images, i.e., where no relation between photos is known, except that they are showing instances of the same category. The key…

计算机视觉与模式识别 · 计算机科学 2021-06-11 Philipp Henzler , Niloy Mitra , Tobias Ritschel

Video generative models have recently achieved notable advancements in synthesis quality. However, generating complex motions remains a critical challenge, as existing models often struggle to produce natural, smooth, and contextually…

计算机视觉与模式识别 · 计算机科学 2025-11-07 Guo Cheng , Danni Yang , Ziqi Huang , Jianlou Si , Chenyang Si , Ziwei Liu

To improve the generalization of 3D human pose estimators, many existing deep learning based models focus on adding different augmentations to training poses. However, data augmentation techniques are limited to the "seen" pose combinations…

计算机视觉与模式识别 · 计算机科学 2023-01-10 Cheng-Yen Yang , Jiajia Luo , Lu Xia , Yuyin Sun , Nan Qiao , Ke Zhang , Zhongyu Jiang , Jenq-Neng Hwang

3D vision foundation models have shown strong generalization in reconstructing key 3D attributes from uncalibrated images through a single feed-forward pass. However, when deployed in online settings such as driving scenarios, predictions…

计算机视觉与模式识别 · 计算机科学 2026-03-19 Fengyi Zhang , Tianjun Zhang , Kasra Khosoussi , Zheng Zhang , Zi Huang , Yadan Luo

We consider the problem of planning views for a robot to acquire images of an object for visual inspection and reconstruction. In contrast to offline methods which require a 3D model of the object as input or online methods which rely on…

机器人学 · 计算机科学 2019-10-07 Selim Engin , Eric Mitchell , Daewon Lee , Volkan Isler , Daniel D. Lee

As the development of deep neural networks, 3D object recognition is becoming increasingly popular in computer vision community. Many multi-view based methods are proposed to improve the category recognition accuracy. These approaches…

计算机视觉与模式识别 · 计算机科学 2019-06-18 Qi Xuan , Fuxian Li , Yi Liu , Yun Xiang

Category-level 3D pose estimation is a fundamentally important problem in computer vision and robotics, e.g. for embodied agents or to train 3D generative models. However, so far methods that estimate the category-level object pose require…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Leonhard Sommer , Artur Jesslen , Eddy Ilg , Adam Kortylewski

We investigate the problem of estimating the 3D shape of an object, given a set of 2D landmarks in a single image. To alleviate the reconstruction ambiguity, a widely-used approach is to confine the unknown 3D shape within a shape space…

计算机视觉与模式识别 · 计算机科学 2015-06-03 Xiaowei Zhou , Spyridon Leonardos , Xiaoyan Hu , Kostas Daniilidis

We present a deep neural network for removing undesirable shading features from an unconstrained portrait image, recovering the underlying texture. Our training scheme incorporates three regularization strategies: masked loss, to emphasize…

计算机视觉与模式识别 · 计算机科学 2022-07-25 Joshua Weir , Junhong Zhao , Andrew Chalmers , Taehyun Rhee

We propose a technique for learning single-view 3D object pose estimation models by utilizing a new source of data -- in-the-wild videos where objects turn. Such videos are prevalent in practice (e.g., cars in roundabouts, airplanes near…

计算机视觉与模式识别 · 计算机科学 2022-12-14 Zezhou Cheng , Matheus Gadelha , Subhransu Maji

In this paper, we propose a two-stage depth ranking based method (DRPose3D) to tackle the problem of 3D human pose estimation. Instead of accurate 3D positions, the depth ranking can be identified by human intuitively and learned using the…

计算机视觉与模式识别 · 计算机科学 2018-05-25 Min Wang , Xipeng Chen , Wentao Liu , Chen Qian , Liang Lin , Lizhuang Ma

We propose a framework that can deform an object in a 2D image as it exists in 3D space. Most existing methods for 3D-aware image manipulation are limited to (1) only changing the global scene information or depth, or (2) manipulating an…

计算机视觉与模式识别 · 计算机科学 2022-03-30 Jihyun Lee , Minhyuk Sung , Hyunjin Kim , Tae-Kyun Kim

3D reconstruction from a single image is a key problem in multiple applications ranging from robotic manipulation to augmented reality. Prior methods have tackled this problem through generative models which predict 3D reconstructions as…

计算机视觉与模式识别 · 计算机科学 2017-08-17 Andrey Kurenkov , Jingwei Ji , Animesh Garg , Viraj Mehta , JunYoung Gwak , Christopher Choy , Silvio Savarese

Recovering the 3D structure of the scene from images yields useful information for tasks such as shape and scene recognition, object detection, or motion planning and object grasping in robotics. In this thesis, we introduce a general…

计算机视觉与模式识别 · 计算机科学 2010-07-20 Hoang Trinh

Near-range portrait photographs often contain perspective distortion artifacts that bias human perception and challenge both facial recognition and reconstruction techniques. We present the first deep learning based approach to remove such…

计算机视觉与模式识别 · 计算机科学 2019-05-21 Yajie Zhao , Zeng Huang , Tianye Li , Weikai Chen , Chloe LeGendre , Xinglei Ren , Jun Xing , Ari Shapiro , Hao Li

Neural 3D reconstruction from multi-view images has recently attracted increasing attention from the community. Existing methods normally learn a neural field for the whole scene, while it is still under-explored how to reconstruct a target…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Xiaobao Wei , Renrui Zhang , Jiarui Wu , Jiaming Liu , Ming Lu , Yandong Guo , Shanghang Zhang

Image segmentation and 3D pose estimation are two key cogs in any algorithm for scene understanding. However, state-of-the-art CRF-based models for image segmentation rely mostly on 2D object models to construct top-down high-order…

计算机视觉与模式识别 · 计算机科学 2016-06-20 Siddharth Mahendran , René Vidal

Most model-free visual object tracking methods formulate the tracking task as object location estimation given by a 2D segmentation or a bounding box in each video frame. We argue that this representation is limited and instead propose to…

计算机视觉与模式识别 · 计算机科学 2023-04-14 Denys Rozumnyi , Jiri Matas , Marc Pollefeys , Vittorio Ferrari , Martin R. Oswald

Detecting anomalies in images has become a well-explored problem in both academia and industry. State-of-the-art algorithms are able to detect defects in increasingly difficult settings and data modalities. However, most current methods are…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Mathis Kruse , Marco Rudolph , Dominik Woiwode , Bodo Rosenhahn