中文
相关论文

相关论文: HART: Human Aligned Reconstruction Transformer

200 篇论文

We present UniSH, a unified, feed-forward framework for joint metric-scale 3D scene and human reconstruction. A key challenge in this domain is the scarcity of large-scale, annotated real-world data, forcing a reliance on synthetic…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Mengfei Li , Peng Li , Zheng Zhang , Jiahao Lu , Chengfeng Zhao , Wei Xue , Qifeng Liu , Sida Peng , Wenxiao Zhang , Wenhan Luo , Yuan Liu , Yike Guo

Reconstructing an interactive human avatar and the background from a monocular video of a dynamic human scene is highly challenging. In this work we adopt a strategy of point cloud decoupling and joint optimization to achieve the decoupled…

图形学 · 计算机科学 2025-06-30 Da Li , Donggang Jia , Markus Hadwiger , Ivan Viola

Constrained by the low-rank bottleneck inherent in attention mechanisms, current stereo matching transformers suffer from limited nonlinear expressivity, which renders their feature representations sensitive to challenging conditions such…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Ziyang Chen , Wenting Li , Yongjun Zhang , Yabo Wu , Bingshu Wang , Yong Zhao , C. L. Philip Chen

3D Gaussian Splatting (3DGS) has been recognized as a pioneering technique in scene reconstruction and novel view synthesis. Recent work on reconstructing the 3D human body using 3DGS attempts to leverage prior information on human pose to…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Hao Tian , Rui Liu , Wen Shen , Yilong Hu , Zhihao Zheng , Xiaolin Qin

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

计算机视觉与模式识别 · 计算机科学 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

Human Mesh Recovery (HMR) is fundamentally ambiguous: under occlusion or weak depth cues, multiple 3D bodies can explain the same image evidence. This ambiguity is not uniform across the body, as torso pose and root structure are often…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Patrick Kwon , Chen Chen

Accurate, real-time 3D reconstruction of human heads from monocular images and videos underlies numerous visual applications. As 3D ground truth data is hard to come by at scale, previous methods have sought to learn from abundant 2D videos…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Liam Schoneveld , Zhe Chen , Davide Davoli , Jiapeng Tang , Saimon Terazawa , Ko Nishino , Matthias Nießner

Reconstruction of the shape and motion of humans from RGB-D is a challenging problem, receiving much attention in recent years. Recent approaches for full-body reconstruction use a statistic shape model, which is built upon accurate…

计算机视觉与模式识别 · 计算机科学 2018-07-10 Ryosuke Kimura , Akihiko Sayo , Fabian Lorenzo Dayrit , Yuta Nakashima , Hiroshi Kawasaki , Ambrosio Blanco , Katsushi Ikeuchi

Real-time rendering of photorealistic and controllable human avatars stands as a cornerstone in Computer Vision and Graphics. While recent advances in neural implicit rendering have unlocked unprecedented photorealism for digital avatars,…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Haokai Pang , Heming Zhu , Adam Kortylewski , Christian Theobalt , Marc Habermann

We propose a novel Auto-Regressive (AR) image generation approach that models images as hierarchical compositions of interpretable visual layers. While AR models have achieved transformative success in language modeling, replicating this…

计算机视觉与模式识别 · 计算机科学 2025-11-13 Siddharth Roheda , Rohit Chowdhury , Aniruddha Bala , Rohan Jaiswal

We address the problem of clothed human reconstruction from a single image or uncalibrated multi-view images. Existing methods struggle with reconstructing detailed geometry of a clothed human and often require a calibrated setting for…

计算机视觉与模式识别 · 计算机科学 2023-04-04 Yukang Cao , Kai Han , Kwan-Yee K. Wong

Parametric body models offer expressive 3D representation of humans across a wide range of poses, shapes, and facial expressions, typically derived by learning a basis over registered 3D meshes. However, existing human mesh modeling…

计算机视觉与模式识别 · 计算机科学 2025-08-22 Jinhyung Park , Javier Romero , Shunsuke Saito , Fabian Prada , Takaaki Shiratori , Yichen Xu , Federica Bogo , Shoou-I Yu , Kris Kitani , Rawal Khirodkar

Recent advances in optimizing Gaussian Splatting for scene geometry have enabled efficient reconstruction of detailed surfaces from images. However, when input views are sparse, such optimization is prone to overfitting, leading to…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Meiying Gu , Jiawei Zhang , Jiahe Li , Xiaohan Yu , Haonan Luo , Jin Zheng , Xiao Bai

Fully supervised human mesh recovery methods are data-hungry and have poor generalizability due to the limited availability and diversity of 3D-annotated benchmark datasets. Recent progress in self-supervised human mesh recovery has been…

计算机视觉与模式识别 · 计算机科学 2022-09-13 Xuan Gong , Meng Zheng , Benjamin Planche , Srikrishna Karanam , Terrence Chen , David Doermann , Ziyan Wu

Despite recent research advancements in reconstructing clothed humans from a single image, accurately restoring the "unseen regions" with high-level details remains an unsolved challenge that lacks attention. Existing methods often generate…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Yangyi Huang , Hongwei Yi , Yuliang Xiu , Tingting Liao , Jiaxiang Tang , Deng Cai , Justus Thies

3D human reconstruction from a single image is a challenging problem and has been exclusively studied in the literature. Recently, some methods have resorted to diffusion models for guidance, optimizing a 3D representation via Score…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Kaiqiang Xiong , Rui Peng , Jiahao Wu , Zhanke Wang , Jie Liang , Xiaoyun Zheng , Feng Gao , Ronggang Wang

We propose a robust and accurate method for reconstructing 3D hand mesh from monocular images. This is a very challenging problem, as hands are often severely occluded by objects. Previous works often have disregarded 2D hand pose…

计算机视觉与模式识别 · 计算机科学 2024-03-14 Shuaibing Wang , Shunli Wang , Dingkang Yang , Mingcheng Li , Ziyun Qian , Liuzhen Su , Lihua Zhang

Appearance-based gaze estimation, aiming to predict accurate 3D gaze direction from a single facial image, has made promising progress in recent years. However, most methods suffer significant performance degradation in cross-domain…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Qida Tan , Hongyu Yang , Wenchao Du

Human hair reconstruction is a challenging problem in computer vision, with growing importance for applications in virtual reality and digital human modeling. Recent advances in 3D Gaussians Splatting (3DGS) provide efficient and explicit…

计算机视觉与模式识别 · 计算机科学 2025-09-10 Yimin Pan , Matthias Nießner , Tobias Kirschstein

Existing 3D human mesh recovery methods often fail to fully exploit the latent information (e.g., human motion, shape alignment), leading to issues with limb misalignment and insufficient local details in the reconstructed human mesh…

计算机视觉与模式识别 · 计算机科学 2025-10-22 Xiang Zhang , Suping Wu , Sheng Yang