English
Related papers

Related papers: PEAR: Pixel-aligned Expressive humAn mesh Recovery

200 papers

As deep neural networks evolve from convolutional neural networks (ConvNets) to advanced vision transformers (ViTs), there is an increased need to eliminate redundant data for faster processing without compromising accuracy. Previous…

Computer Vision and Pattern Recognition · Computer Science 2024-07-04 Tanvir Mahmud , Burhaneddin Yaman , Chun-Hao Liu , Diana Marculescu

Single image super-resolution (SISR) deals with a fundamental problem of upsampling a low-resolution (LR) image to its high-resolution (HR) version. Last few years have witnessed impressive progress propelled by deep learning methods.…

Computer Vision and Pattern Recognition · Computer Science 2021-05-24 Wenbo Li , Kun Zhou , Lu Qi , Nianjuan Jiang , Jiangbo Lu , Jiaya Jia

Great progress has been made in estimating 3D human pose and shape from images and video by training neural networks to directly regress the parameters of parametric human models like SMPL. However, existing body models have simplified…

Graphics · Computer Science 2025-09-09 Marilyn Keller , Keenon Werling , Soyong Shin , Scott Delp , Sergi Pujades , C. Karen Liu , Michael J. Black

3D Human Body Reconstruction from a monocular image is an important problem in computer vision with applications in virtual and augmented reality platforms, animation industry, en-commerce domain, etc. While several of the existing works…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Abbhinav Venkat , Chaitanya Patel , Yudhik Agrawal , Avinash Sharma

We present HARP (HAnd Reconstruction and Personalization), a personalized hand avatar creation approach that takes a short monocular RGB video of a human hand as input and reconstructs a faithful hand avatar exhibiting a high-fidelity…

Computer Vision and Pattern Recognition · Computer Science 2023-07-06 Korrawe Karunratanakul , Sergey Prokudin , Otmar Hilliges , Siyu Tang

Reconstructing animatable 3D humans from casually captured images of articulated subjects without camera or pose information is highly practical but remains challenging due to view misalignment, occlusions, and the absence of structural…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Lingteng Qiu , Peihao Li , Heyuan Li , Qi Zuo , Xiaodong Gu , Yuan Dong , Weihao Yuan , Rui Peng , Siyu Zhu , Xiaoguang Han , Guanying Chen , Zilong Dong

Nuanced expressiveness, particularly through fine-grained hand and facial expressions, is pivotal for enhancing the realism and vitality of digital human representations. In this work, we focus on investigating the expressiveness of human…

Computer Vision and Pattern Recognition · Computer Science 2024-07-04 Hezhen Hu , Zhiwen Fan , Tianhao Wu , Yihan Xi , Seoyoung Lee , Georgios Pavlakos , Zhangyang Wang

Open-world 3D generation has recently attracted considerable attention. While many single-image-to-3D methods have yielded visually appealing outcomes, they often lack sufficient controllability and tend to produce hallucinated regions that…

Computer Vision and Pattern Recognition · Computer Science 2024-08-20 Chao Xu , Ang Li , Linghao Chen , Yulin Liu , Ruoxi Shi , Hao Su , Minghua Liu

We present a novel method for reconstructing clothed humans from a sparse set of, e.g., 1 to 6 RGB images. Despite impressive results from recent works employing deep implicit representation, we revisit the volumetric approach and…

Computer Vision and Pattern Recognition · Computer Science 2023-07-26 Sicong Tang , Guangyuan Wang , Qing Ran , Lingzhi Li , Li Shen , Ping Tan

Existing Human NeRF methods for reconstructing 3D humans typically rely on multiple 2D images from multi-view cameras or monocular videos captured from fixed camera views. However, in real-world scenarios, human images are often captured…

Computer Vision and Pattern Recognition · Computer Science 2023-08-17 Shoukang Hu , Fangzhou Hong , Liang Pan , Haiyi Mei , Lei Yang , Ziwei Liu

Existing 3D human pose estimation algorithms trained on distortion-free datasets suffer performance drop when applied to new scenarios with a specific camera distortion. In this paper, we propose a simple yet effective model for 3D human…

Computer Vision and Pattern Recognition · Computer Science 2021-12-06 Hanbyel Cho , Yooshin Cho , Jaemyung Yu , Junmo Kim

Learning accurate and parsimonious point cloud representations of scene surfaces from scratch remains a challenge in 3D representation learning. Existing point-based methods often suffer from the vanishing gradient problem or require a…

Computer Vision and Pattern Recognition · Computer Science 2023-12-08 Yanshu Zhang , Shichong Peng , Alireza Moazeni , Ke Li

Recent image-to-3D reconstruction models have greatly advanced geometry generation, but they still struggle to faithfully generate realistic appearance. To address this, we introduce ARM, a novel method that reconstructs high-quality 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Xiang Feng , Chang Yu , Zoubin Bi , Yintong Shang , Feng Gao , Hongzhi Wu , Kun Zhou , Chenfanfu Jiang , Yin Yang

Parameter-efficient finetuning (PEFT) has become ubiquitous to adapt foundation models to downstream task requirements while retaining their generalization ability. However, the amount of additionally introduced parameters and compute for…

Machine Learning · Computer Science 2024-10-14 Massimo Bini , Karsten Roth , Zeynep Akata , Anna Khoreva

In recent years, there has been a growing interest in developing effective alignment pipelines to generate unified representations from different modalities for multi-modal fusion and generation. As an important component of Human-Centric…

Computer Vision and Pattern Recognition · Computer Science 2025-10-23 Zhongyu Jiang , Wenhao Chai , Lei Li , Zhuoran Zhou , Cheng-Yen Yang , Jenq-Neng Hwang

Meaningful facial parts can convey key cues for both facial action unit detection and expression prediction. Textured 3D face scan can provide both detailed 3D geometric shape and 2D texture appearance cues of the face which are beneficial…

Computer Vision and Pattern Recognition · Computer Science 2018-03-16 Asim Jan , Huaxiong Ding , Hongying Meng , Liming Chen , Huibin Li

We present an approach to recover absolute 3D human poses from multi-view images by incorporating multi-view geometric priors in our model. It consists of two separate steps: (1) estimating the 2D poses in multi-view images and (2)…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Haibo Qiu , Chunyu Wang , Jingdong Wang , Naiyan Wang , Wenjun Zeng

Recovering 3D Human-Object Interaction (HOI) from single color images is challenging due to depth ambiguities, occlusions, and the huge variation in object shape and appearance. Thus, past work requires controlled settings such as known…

Computer Vision and Pattern Recognition · Computer Science 2025-04-25 Alpár Cseke , Shashank Tripathi , Sai Kumar Dwivedi , Arjun Lakshmipathy , Agniv Chatterjee , Michael J. Black , Dimitrios Tzionas

Feedforward monocular face capture methods seek to reconstruct posed faces from a single image of a person. Current state of the art approaches have the ability to regress parametric 3D face models in real-time across a wide range of…

Computer Vision and Pattern Recognition · Computer Science 2024-09-13 Kelian Baert , Shrisha Bharadwaj , Fabien Castan , Benoit Maujean , Marc Christie , Victoria Abrevaya , Adnane Boukhayma

Referring Expression Comprehension (REC), which aims to ground a local visual region via natural language, is a task that heavily relies on multimodal alignment. Most existing methods utilize powerful pre-trained models to transfer…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Ting Liu , Zunnan Xu , Yue Hu , Liangtao Shi , Zhiqiang Wang , Quanjun Yin