English
Related papers

Related papers: HART: Human Aligned Reconstruction Transformer

200 papers

This paper presents a novel framework to recover detailed human body shapes from a single image. It is a challenging task due to factors such as variations in human shapes, body poses, and viewpoints. Prior methods typically attempt to…

Computer Vision and Pattern Recognition · Computer Science 2019-05-14 Hao Zhu , Xinxin Zuo , Sen Wang , Xun Cao , Ruigang Yang

We propose a scalable neural network framework to reconstruct the 3D mesh of a human body from multi-view images, in the subspace of the SMPL model. Use of multi-view images can significantly reduce the projection ambiguity of the problem,…

Computer Vision and Pattern Recognition · Computer Science 2019-08-27 Junbang Liang , Ming C. Lin

Recent advancements in radiance field rendering show promising results in 3D scene representation, where Gaussian splatting-based techniques emerge as state-of-the-art due to their quality and efficiency. Gaussian splatting is widely used…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Arnab Dey , Cheng-You Lu , Andrew I. Comport , Srinath Sridhar , Chin-Teng Lin , Jean Martinet

Achieving dexterous robotic grasping with multi-fingered hands remains a significant challenge. While existing methods rely on complete 3D scans to predict grasp poses, these approaches face limitations due to the difficulty of acquiring…

3D Gaussian Splatting offers expressive scene reconstruction, modeling a broad range of visual, geometric, and semantic information. However, efficient real-time map reconstruction with data streamed from multiple robots and devices remains…

Robotics · Computer Science 2025-06-04 Javier Yu , Timothy Chen , Mac Schwager

Reconstructing 3D humans from images captured at multiple perspectives typically requires pre-calibration, like using checkerboards or MVS algorithms, which limits scalability and applicability in diverse real-world scenarios. In this work,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-16 Xiaozhen Qiao , Wenjia Wang , Zhiyuan Zhao , Jiacheng Sun , Ping Luo , Hongyuan Zhang , Xuelong Li

Large language models (LLMs) have demonstrated remarkable performance in text generation and knowledge-intensive question answering. Nevertheless, they are prone to producing hallucinated content, which severely undermines their reliability…

Computation and Language · Computer Science 2026-03-09 Shize Liang , Hongzhi Wang

Creating a realistic clothed human from a single-view RGB image is crucial for applications like mixed reality and filmmaking. Despite some progress in recent years, mainstream methods often fail to fully utilize side-view information, as…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Dong Liu , Yifan Yang , Zixiong Huang , Yuxin Gao , Mingkui Tan

In recent years, a variety of ML architectures and techniques have seen success in producing skillful medium range weather forecasts. In particular, Vision Transformer (ViT)-based models (e.g. Pangu-Weather, FuXi) have shown strong…

Computer Vision and Pattern Recognition · Computer Science 2024-03-27 Vivek Ramavajjala

A key challenge in the task of human pose and shape estimation is occlusion, including self-occlusions, object-human occlusions, and inter-person occlusions. The lack of diverse and accurate pose and shape training data becomes a major…

Computer Vision and Pattern Recognition · Computer Science 2022-03-02 Kaibing Yang , Renshu Gu , Maoyu Wang , Masahiro Toyoura , Gang Xu

Recovering the 3D geometry of a scene from a sparse set of uncalibrated images is a long-standing problem in computer vision. While recent learning-based approaches such as DUSt3R and MASt3R have demonstrated impressive results by directly…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Sara Rojas , Matthieu Armando , Bernard Ghamen , Philippe Weinzaepfel , Vincent Leroy , Gregory Rogez

Accurate 3D human pose estimation is fundamental for applications such as augmented reality and human-robot interaction. State-of-the-art multi-view methods learn to fuse predictions across views by training on large annotated datasets,…

Computer Vision and Pattern Recognition · Computer Science 2025-12-03 Laura Bragagnolo , Leonardo Barcellona , Stefano Ghidoni

Human Mesh Recovery (HMR) is an important yet challenging problem with applications across various domains including motion capture, augmented reality, and biomechanics. Accurately predicting human pose parameters from a single image…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Jaewoo Heo , George Hu , Zeyu Wang , Serena Yeung-Levy

Recent advances in neural radiance fields enable novel view synthesis of photo-realistic images in dynamic settings, which can be applied to scenarios with human animation. Commonly used implicit backbones to establish accurate models,…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 HyunJun Jung , Nikolas Brasch , Jifei Song , Eduardo Perez-Pellitero , Yiren Zhou , Zhihao Li , Nassir Navab , Benjamin Busam

Recent techniques on implicit geometry representation learning and neural rendering have shown promising results for 3D clothed human reconstruction from sparse video inputs. However, it is still challenging to reconstruct detailed surface…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Hao Wang , Qingshan Xu , Hongyuan Chen , Rui Ma

We describe Human Mesh Recovery (HMR), an end-to-end framework for reconstructing a full 3D mesh of a human body from a single RGB image. In contrast to most current methods that compute 2D or 3D joint locations, we produce a richer and…

Computer Vision and Pattern Recognition · Computer Science 2018-06-26 Angjoo Kanazawa , Michael J. Black , David W. Jacobs , Jitendra Malik

Detailed and photorealistic 3D human modeling is essential for various applications and has seen tremendous progress. However, full-body reconstruction from a monocular RGB image remains challenging due to the ill-posed nature of the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Peng Li , Wangguandong Zheng , Yuan Liu , Tao Yu , Yangguang Li , Xingqun Qi , Xiaowei Chi , Siyu Xia , Yan-Pei Cao , Wei Xue , Wenhan Luo , Yike Guo

Nonparametric approaches have shown promising results on reconstructing 3D human mesh from a single monocular image. Unlike previous approaches that use a parametric human model like skinned multi-person linear model (SMPL), and attempt to…

Computer Vision and Pattern Recognition · Computer Science 2020-03-03 Kevin Lin , Lijuan Wang , Ying Jin , Zicheng Liu , Ming-Ting Sun

Recently, data-driven single-view reconstruction methods have shown great progress in modeling 3D dressed humans. However, such methods suffer heavily from depth ambiguities and occlusions inherent to single view inputs. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2021-12-07 Pierre Zins , Yuanlu Xu , Edmond Boyer , Stefanie Wuhrer , Tony Tung

Egocentric human mesh recovery (HMR) from monocular head-mounted cameras is increasingly important for AR/VR applications, but remains challenging due to the lack of reliable ground-truth (GT) annotations based on parametric human body…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Soyeon Na , Seung Young Noh , Ju Yong Chang