English
Related papers

Related papers: SHeaP: Self-Supervised Head Geometry Predictor Lea…

200 papers

3D Gaussian Splatting (3DGS) provides an efficient method for high-quality scene reconstruction using anisotropic Gaussians. Recently, 3DGS-based methods have significantly improved the rendering quality of human avatars while enabling…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Hongzhe Liao , Chuhua Xian , Hongmin Cai , Haiyang Liu , Fa-Ting Hong

Unsupervised object discovery and localization aims to detect or segment objects in an image without any supervision. Recent efforts have demonstrated a notable potential to identify salient foreground objects by utilizing self-supervised…

Computer Vision and Pattern Recognition · Computer Science 2024-01-05 Xin Zhang , Jinheng Xie , Yuan Yuan , Michael Bi Mi , Robby T. Tan

We present a new shear calibration method based on machine learning. The method estimates the individual shear responses of the objects from the combination of several measured properties on the images using supervised learning. The…

Cosmology and Nongalactic Astrophysics · Physics 2020-11-25 Arnau Pujol , Jerome Bobin , Florent Sureau , Axel Guinot , Martin Kilbinger

Photorealistic 3D reconstruction of street scenes is a critical technique for developing real-world simulators for autonomous driving. Despite the efficacy of Neural Radiance Fields (NeRF) for driving scenes, 3D Gaussian Splatting (3DGS)…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Nan Huang , Xiaobao Wei , Wenzhao Zheng , Pengju An , Ming Lu , Wei Zhan , Masayoshi Tomizuka , Kurt Keutzer , Shanghang Zhang

Our method studies the complex task of object-centric 3D understanding from a single RGB-D observation. As it is an ill-posed problem, existing methods suffer from low performance for both 3D shape and 6D pose and size estimation in complex…

Computer Vision and Pattern Recognition · Computer Science 2022-07-28 Muhammad Zubair Irshad , Sergey Zakharov , Rares Ambrus , Thomas Kollar , Zsolt Kira , Adrien Gaidon

3D reconstruction and simulation, although interrelated, have distinct objectives: reconstruction requires a flexible 3D representation that can adapt to diverse scenes, while simulation needs a structured representation to model motion…

Computer Vision and Pattern Recognition · Computer Science 2024-11-25 Shaojie Ma , Yawei Luo , Wei Yang , Yi Yang

High-fidelity reconstruction of head avatars from monocular videos is highly desirable for virtual human applications, but it remains a challenge in the fields of computer graphics and computer vision. In this paper, we propose a two-phase…

Graphics · Computer Science 2025-03-31 Pilseo Park , Ze Zhang , Michel Sarkis , Ning Bi , Xiaoming Liu , Yiying Tong

Creating high-fidelity and editable head avatars is a pivotal challenge in computer vision and graphics, boosting many AR/VR applications. While recent advancements have achieved photorealistic renderings and plausible animation, head…

Computer Vision and Pattern Recognition · Computer Science 2025-08-18 Heyi Sun , Cong Wang , Tian-Xing Xu , Jingwei Huang , Di Kang , Chunchao Guo , Song-Hai Zhang

Deep neural networks trained as image denoisers are widely used as priors for solving imaging inverse problems. While Gaussian denoising is thought sufficient for learning image priors, we show that priors from deep models pre-trained as…

Image and Video Processing · Electrical Eng. & Systems 2024-10-04 Yuyang Hu , Albert Peng , Weijie Gan , Peyman Milanfar , Mauricio Delbracio , Ulugbek S. Kamilov

This paper introduces VisionPAD, a novel self-supervised pre-training paradigm designed for vision-centric algorithms in autonomous driving. In contrast to previous approaches that employ neural rendering with explicit depth supervision,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-23 Haiming Zhang , Wending Zhou , Yiyao Zhu , Xu Yan , Jiantao Gao , Dongfeng Bai , Yingjie Cai , Bingbing Liu , Shuguang Cui , Zhen Li

This paper explores Masked Autoencoders (MAE) with Gaussian Splatting. While reconstructive self-supervised learning frameworks such as MAE learns good semantic abstractions, it is not trained for explicit spatial awareness. Our approach,…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Jathushan Rajasegaran , Xinlei Chen , Rulilong Li , Christoph Feichtenhofer , Jitendra Malik , Shiry Ginosar

This paper presents RoGSplat, a novel approach for synthesizing high-fidelity novel views of unseen human from sparse multi-view images, while requiring no cumbersome per-subject optimization. Unlike previous methods that typically struggle…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Junjin Xiao , Qing Zhang , Yonewei Nie , Lei Zhu , Wei-Shi Zheng

Text-guided 3D human generation has advanced with the development of efficient 3D representations and 2D-lifting methods like Score Distillation Sampling (SDS). However, current methods suffer from prolonged training times and often produce…

Computer Vision and Pattern Recognition · Computer Science 2025-03-17 Zichen Tang , Yuan Yao , Miaomiao Cui , Liefeng Bo , Hongyu Yang

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruction. This raises…

Computer Vision and Pattern Recognition · Computer Science 2025-02-14 Jing Wen , Alexander G. Schwing , Shenlong Wang

Reconstructing 3D human bodies from sparse views has been an appealing topic, which is crucial to broader the related applications. In this paper, we propose a quite challenging but valuable task to reconstruct the human body from only two…

Graphics · Computer Science 2025-08-21 Jia Lu , Taoran Yi , Jiemin Fang , Chen Yang , Chuiyun Wu , Wei Shen , Wenyu Liu , Qi Tian , Xinggang Wang

Surface reconstruction is fundamental to computer vision and graphics, enabling applications in 3D modeling, mixed reality, robotics, and more. Existing approaches based on volumetric rendering obtain promising results, but optimize on a…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Yueh-Cheng Liu , Lukas Höllein , Matthias Nießner , Angela Dai

3D Gaussian Splatting (3D-GS) enables real-time 3D scene reconstruction but lacks robust segmentation for editing tasks such as object removal, extraction, and recoloring. Existing approaches that lift 2D segmentations to the 3D domain…

Computer Vision and Pattern Recognition · Computer Science 2026-05-18 Raushan Joshi , Jean-Yves Guillemaut

While methods that regress 3D human meshes from images have progressed rapidly, the estimated body shapes often do not capture the true human shape. This is problematic since, for many applications, accurate body shape is as important as…

Computer Vision and Pattern Recognition · Computer Science 2022-06-15 Vasileios Choutas , Lea Muller , Chun-Hao P. Huang , Siyu Tang , Dimitrios Tzionas , Michael J. Black

The emergence of neural rendering has significantly advanced the rendering quality of 3D human avatars, with the recently popular 3DGS technique enabling real-time performance. However, SMPL-driven 3DGS human avatars still struggle to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Wangze Xu , Yifan Zhan , Zhihang Zhong , Xiao Sun

In this paper we propose a highly scalable convolutional neural network, end-to-end trainable, for real-time 3D human pose regression from still RGB images. We call this approach the Scalable Sequential Pyramid Networks (SSP-Net) as it is…

Computer Vision and Pattern Recognition · Computer Science 2020-09-07 Diogo Luvizon , Hedi Tabia , David Picard