中文
相关论文

相关论文: Pix2NPHM: Learning to Regress NPHM Reconstructions…

200 篇论文

We describe Human Mesh Recovery (HMR), an end-to-end framework for reconstructing a full 3D mesh of a human body from a single RGB image. In contrast to most current methods that compute 2D or 3D joint locations, we produce a richer and…

计算机视觉与模式识别 · 计算机科学 2018-06-26 Angjoo Kanazawa , Michael J. Black , David W. Jacobs , Jitendra Malik

We present an approach to generate high fidelity 3D face avatar with a high-resolution UV texture map from a single image. To estimate the face geometry, we use a deep neural network to directly predict vertex coordinates of the 3D face…

计算机视觉与模式识别 · 计算机科学 2019-12-10 Ruizhe Wang , Chih-Fan Chen , Hao Peng , Xudong Liu , Oliver Liu , Xin Li

High-quality reconstruction of controllable 3D head avatars from 2D videos is highly desirable for virtual human applications in movies, games, and telepresence. Neural implicit fields provide a powerful representation to model 3D head…

计算机视觉与模式识别 · 计算机科学 2023-04-24 Chuhan Chen , Matthew O'Toole , Gaurav Bharaj , Pablo Garrido

Vision Transformers (ViTs) are built by stacking independently parameterized blocks, but it remains unclear how much of this depth requires layer specific transformations and how much can be realized through recurrent computation. We study…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Michal Byra , Pawel Olszowiec , Grzegorz Stefanski , Grzegorz Gruszczynski , Alberto Presta

Recovering 3D face models from 2D in-the-wild images has gained considerable attention in the computer vision community due to its wide range of potential applications. However, the lack of ground-truth labeled datasets and the complexity…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Danling Cao

Novel View Synthesis (NVS) has traditionally relied on models with explicit 3D inductive biases combined with known camera parameters from Structure-from-Motion (SfM) beforehand. Recent vision foundation models like VGGT take an orthogonal…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Youming Deng , Songyou Peng , Junyi Zhang , Kathryn Heal , Tiancheng Sun , John Flynn , Steve Marschner , Lucy Chai

Recovering 3D human body shape and pose from 2D images is a challenging task due to high complexity and flexibility of human body, and relatively less 3D labeled data. Previous methods addressing these issues typically rely on predicting…

计算机视觉与模式识别 · 计算机科学 2019-03-29 Pengfei Yao , Zheng Fang , Fan Wu , Yao Feng , Jiwei Li

Faithfully reconstructing textured shapes and physical properties from videos presents an intriguing yet challenging problem. Significant efforts have been dedicated to advancing such a system identification problem in this area. Previous…

图形学 · 计算机科学 2025-06-10 Chuhao Chen , Zhiyang Dou , Chen Wang , Yiming Huang , Anjun Chen , Qiao Feng , Jiatao Gu , Lingjie Liu

The reconstruction of cortical surfaces from brain magnetic resonance imaging (MRI) scans is essential for quantitative analyses of cortical thickness and sulcal morphology. Although traditional and deep learning-based algorithmic pipelines…

计算机视觉与模式识别 · 计算机科学 2022-03-21 Fabian Bongratz , Anne-Marie Rickmann , Sebastian Pölsterl , Christian Wachinger

Transformers, which are popular for language modeling, have been explored for solving vision tasks recently, e.g., the Vision Transformer (ViT) for image classification. The ViT model splits each image into a sequence of tokens with fixed…

计算机视觉与模式识别 · 计算机科学 2021-12-01 Li Yuan , Yunpeng Chen , Tao Wang , Weihao Yu , Yujun Shi , Zihang Jiang , Francis EH Tay , Jiashi Feng , Shuicheng Yan

We propose a principled convolutional neural pyramid (CNP) framework for general low-level vision and image processing tasks. It is based on the essential finding that many applications require large receptive fields for structure…

计算机视觉与模式识别 · 计算机科学 2017-04-10 Xiaoyong Shen , Ying-Cong Chen , Xin Tao , Jiaya Jia

Vision Transformers (ViTs) have attracted a lot of popularity in recent years, due to their exceptional capabilities in modeling long-range spatial dependencies and scalability for large scale training. Although the training parallelism of…

计算机视觉与模式识别 · 计算机科学 2024-01-29 Ali Hatamizadeh , Michael Ranzinger , Shiyi Lan , Jose M. Alvarez , Sanja Fidler , Jan Kautz

Terahertz (THz) imaging is one of the hotspots in the field of optics, where the depth information retrieval is a key factor to restore the three-dimensional appearance of objects. Impressive results for depth extraction in visible and…

光学 · 物理学 2024-07-09 Mingjun Xiang , Hui Yuan , Kai Zhou , Hartmut G. Roskos

Three-dimensional shape reconstruction of 2D landmark points on a single image is a hallmark of human vision, but is a task that has been proven difficult for computer vision algorithms. We define a feed-forward deep neural network…

计算机视觉与模式识别 · 计算机科学 2016-09-29 Ruiqi Zhao , Yan Wang , Aleix Martinez

We introduce InverseFaceNet, a deep convolutional inverse rendering framework for faces that jointly estimates facial pose, shape, expression, reflectance and illumination from a single input image. By estimating all parameters from just a…

计算机视觉与模式识别 · 计算机科学 2018-05-17 Hyeongwoo Kim , Michael Zollhöfer , Ayush Tewari , Justus Thies , Christian Richardt , Christian Theobalt

There are two de facto standard architectures in recent computer vision: Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs). Strong inductive biases of convolutions help the model learn sample effectively, but such strong…

计算机视觉与模式识别 · 计算机科学 2022-10-05 Yunsung Lee , Gyuseong Lee , Kwangrok Ryoo , Hyojun Go , Jihye Park , Seungryong Kim

3D face reconstruction plays a very important role in many real-world multimedia applications, including digital entertainment, social media, affection analysis, and person identification. The de-facto pipeline for estimating the parametric…

计算机视觉与模式识别 · 计算机科学 2021-01-07 Jialiang Zhang , Lixiang Lin , Jianke Zhu , Steven C. H. Hoi

We present FaceLift, a novel feed-forward approach for generalizable high-quality 360-degree 3D head reconstruction from a single image. Our pipeline first employs a multi-view latent diffusion model to generate consistent side and back…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Weijie Lyu , Yi Zhou , Ming-Hsuan Yang , Zhixin Shu

Achieving photorealistic 3D view synthesis and relighting of human portraits is pivotal for advancing AR/VR applications. Existing methodologies in portrait relighting demonstrate substantial limitations in terms of generalization and 3D…

Due to their highly structured characteristics, faces are easier to recover than natural scenes for blind image super-resolution. Therefore, we can extract the degradation representation of an image from the low-quality and recovered face…

计算机视觉与模式识别 · 计算机科学 2023-09-18 Zhicun Yin , Ming Liu , Xiaoming Li , Hui Yang , Longan Xiao , Wangmeng Zuo