English
Related papers

Related papers: Learning Articulated Shape with Keypoint Pseudo-la…

200 papers

Animating a newly designed character using motion capture (mocap) data is a long standing problem in computer animation. A key consideration is the skeletal structure that should correspond to the available mocap data, and the shape…

Graphics · Computer Science 2021-05-07 Peizhuo Li , Kfir Aberman , Rana Hanocka , Libin Liu , Olga Sorkine-Hornung , Baoquan Chen

For non-rigid objects, predicting the 3D shape from 2D keypoint observations is ill-posed due to occlusions, and the need to disentangle changes in viewpoint and changes in shape. This challenge has often been addressed by embedding…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Shalini Maiti , Lourdes Agapito , Benjamin Graham

In the monocular setting, predicting 3D pose and shape of animals typically relies solely on visual information, which is highly under-constrained. In this work, we explore using audio to enhance 3D shape and motion recovery of horses from…

Computer Vision and Pattern Recognition · Computer Science 2024-07-02 Ci Li , Elin Hernlund , Hedvig Kjellström , Silvia Zuffi

This paper presents a semi-supervised learning framework to train a keypoint detector using multiview image streams given the limited labeled data (typically $<$4\%). We leverage the complementary relationship between multiview geometry and…

Computer Vision and Pattern Recognition · Computer Science 2019-03-26 Yilun Zhang , Hyun Soo Park

Effectively utilizing the vast amounts of ego-centric navigation data that is freely available on the internet can advance generalized intelligent systems, i.e., to robustly scale across perspectives, platforms, environmental conditions,…

Computer Vision and Pattern Recognition · Computer Science 2022-04-22 Jimuyang Zhang , Ruizhao Zhu , Eshed Ohn-Bar

Understanding the geometry and pose of objects in 2D images is a fundamental necessity for a wide range of real world applications. Driven by deep neural networks, recent methods have brought significant improvements to object pose…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Jogendra Nath Kundu , Rahul M. V. , Aditya Ganeshan , R. Venkatesh Babu

Recovery of articulated 3D structure from 2D observations is a challenging computer vision problem with many applications. Current learning-based approaches achieve state-of-the-art accuracy on public benchmarks but are restricted to…

Computer Vision and Pattern Recognition · Computer Science 2019-11-13 Onorina Kovalenko , Vladislav Golyanik , Jameel Malik , Ahmed Elhayek , Didier Stricker

With the recent growth of urban mapping and autonomous driving efforts, there has been an explosion of raw 3D data collected from terrestrial platforms with lidar scanners and color cameras. However, due to high labeling costs, ground-truth…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Kyle Genova , Xiaoqi Yin , Abhijit Kundu , Caroline Pantofaru , Forrester Cole , Avneesh Sud , Brian Brewington , Brian Shucker , Thomas Funkhouser

Much progress has been made in reconstructing garments from an image or a video. However, none of existing works meet the expectations of digitizing high-quality animatable dynamic garments that can be adjusted to various unseen poses. In…

Computer Vision and Pattern Recognition · Computer Science 2023-11-03 Xiongzheng Li , Jinsong Zhang , Yu-Kun Lai , Jingyu Yang , Kun Li

This paper investigates the research task of reconstructing the 3D clothed human body from a monocular image. Due to the inherent ambiguity of single-view input, existing approaches leverage pre-trained SMPL(-X) estimation models or…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 Gangjian Zhang , Nanjie Yao , Shunsi Zhang , Hanfeng Zhao , Guoliang Pang , Jian Shu , Hao Wang

Existing methods for reconstructing animatable 3D animals from videos typically rely on sparse semantic keypoints to fit parametric models. However, obtaining such keypoints is labor-intensive, and keypoint detectors trained on limited…

Computer Vision and Pattern Recognition · Computer Science 2025-07-15 Shanshan Zhong , Jiawei Peng , Zehan Zheng , Zhongzhan Huang , Wufei Ma , Guofeng Zhang , Qihao Liu , Alan Yuille , Jieneng Chen

We present a convolutional neural network for joint 3D shape prediction and viewpoint estimation from a single input image. During training, our network gets the learning signal from a silhouette of an object in the input image - a form of…

Robotics · Computer Science 2019-10-18 Oier Mees , Maxim Tatarchenko , Thomas Brox , Wolfram Burgard

Global visual localization estimates the absolute pose of a camera using a single image, in a previously mapped area. Obtaining the pose from a single image enables many robotics and augmented/virtual reality applications. Inspired by…

Computer Vision and Pattern Recognition · Computer Science 2023-12-19 Mohammad Altillawi , Shile Li , Sai Manoj Prakhya , Ziyuan Liu , Joan Serrat

Understanding the 3D motion of articulated objects is essential in robotic scene understanding, mobile manipulation, and motion planning. Prior methods for articulation estimation have primarily focused on controlled settings, assuming…

Animating an object in 3D often requires an articulated structure, e.g. a kinematic chain or skeleton of the manipulated object with proper skinning weights, to obtain smooth movements and surface deformations. However, existing models that…

Computer Vision and Pattern Recognition · Computer Science 2023-04-17 Tianshu Kuai , Akash Karthikeyan , Yash Kant , Ashkan Mirzaei , Igor Gilitschenski

Online mapping models show remarkable results in predicting vectorized maps from multi-view camera images only. However, all existing approaches still rely on ground-truth high-definition maps during training, which are expensive to obtain…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 Christian Löwens , Thorben Funke , Jingchao Xie , Alexandru Paul Condurache

We present a novel method for precise 3D object localization in single images from a single calibrated camera using only 2D labels. No expensive 3D labels are needed. Thus, instead of using 3D labels, our model is trained with…

Computer Vision and Pattern Recognition · Computer Science 2023-11-30 Daniel Kienzle , Julian Lorenz , Katja Ludwig , Rainer Lienhart

Detecting and matching robust viewpoint-invariant keypoints is critical for visual SLAM and Structure-from-Motion. State-of-the-art learning-based methods generate training samples via homography adaptation to create 2D synthetic views with…

Computer Vision and Pattern Recognition · Computer Science 2020-11-19 Jiexiong Tang , Rares Ambrus , Vitor Guizilini , Sudeep Pillai , Hanme Kim , Patric Jensfelt , Adrien Gaidon

We present a method to learn single-view reconstruction of the 3D shape, pose, and texture of objects from categorized natural images in a self-supervised manner. Since this is a severely ill-posed problem, carefully designing a training…

Computer Vision and Pattern Recognition · Computer Science 2019-11-21 Hiroharu Kato , Tatsuya Harada

A robot's ability to act is fundamentally constrained by what it can perceive. Many existing approaches to visual representation learning utilize general-purpose training criteria, e.g. image reconstruction, smoothness in latent space, or…