中文
相关论文

相关论文: MPT: Mesh Pre-Training with Transformers for Human…

200 篇论文

This paper introduces a novel Pre-trained Spatial Temporal Many-to-One (P-STMO) model for 2D-to-3D human pose estimation task. To reduce the difficulty of capturing spatial and temporal information, we divide this task into two stages:…

计算机视觉与模式识别 · 计算机科学 2022-08-01 Wenkang Shan , Zhenhua Liu , Xinfeng Zhang , Shanshe Wang , Siwei Ma , Wen Gao

Reconstructing posed 3D human models from monocular images has important applications in the sports industry, including performance tracking, injury prevention and virtual training. In this work, we combine 3D human pose and shape…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Lorenza Prospero , Abdullah Hamdi , Joao F. Henriques , Christian Rupprecht

Thanks to the rapid development of CNNs and depth sensors, great progress has been made in 3D hand pose estimation. Nevertheless, it is still far from being solved for its cluttered circumstance and severe self-occlusion of hand. In this…

机器人学 · 计算机科学 2019-11-13 Weiguo Zhou , Xin Jiang , Chen Chen , Sijia Mei , Yun-Hui Liu

Estimating 3D human pose from a single image is a challenging task. This work attempts to address the uncertainty of lifting the detected 2D joints to the 3D space by introducing an intermediate state - Part-Centric Heatmap Triplets…

计算机视觉与模式识别 · 计算机科学 2019-10-29 Kun Zhou , Xiaoguang Han , Nianjuan Jiang , Kui Jia , Jiangbo Lu

Multimodal pre-training models, such as LXMERT, have achieved excellent results in downstream tasks. However, current pre-trained models require large amounts of training data and have huge model sizes, which make them difficult to apply in…

计算与语言 · 计算机科学 2021-08-02 Tongtong Liu , Fangxiang Feng , Xiaojie Wang

We present two novel solutions for multi-view 3D human pose estimation based on new learnable triangulation methods that combine 3D information from multiple 2D views. The first (baseline) solution is a basic differentiable algebraic…

计算机视觉与模式识别 · 计算机科学 2019-05-15 Karim Iskakov , Egor Burkov , Victor Lempitsky , Yury Malkov

We present a novel method to improve the accuracy of the 3D reconstruction of clothed human shape from a single image. Recent work has introduced volumetric, implicit and model-based shape learning frameworks for reconstruction of objects…

计算机视觉与模式识别 · 计算机科学 2020-09-30 Akin Caliskan , Armin Mustafa , Evren Imre , Adrian Hilton

Parameter-efficient fine-tuning (PEFT) techniques have emerged to address overfitting and high computational costs associated with fully fine-tuning in self-supervised learning. Mainstream PEFT methods add a few trainable parameters while…

计算机视觉与模式识别 · 计算机科学 2025-06-06 Xingliang Lei , Yiwen Ye , Zhisong Wang , Ziyang Chen , Minglei Shu , Weidong Cai , Yanning Zhang , Yong Xia

Capturing a 3D human body is one of the important tasks in computer vision with a wide range of applications such as virtual reality and sports analysis. However, conventional frame cameras are limited by their temporal resolution and…

计算机视觉与模式识别 · 计算机科学 2024-04-17 Kai Kohyama , Shintaro Shiba , Yoshimitsu Aoki

While heatmap-based human pose estimation methods have shown strong performance, they suffer from three main problems: (P1) "Commonly used Mean Squared Error (MSE)" Loss may not always improve joint localization because it penalizes all…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Muhammed Can Keles , Bedrettin Cetinkaya , Sinan Kalkan , Emre Akbas

We present SCULPT, a novel 3D generative model for clothed and textured 3D meshes of humans. Specifically, we devise a deep neural network that learns to represent the geometry and appearance distribution of clothed human bodies. Training…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Soubhik Sanyal , Partha Ghosh , Jinlong Yang , Michael J. Black , Justus Thies , Timo Bolkart

In 3D human action recognition, limited supervised data makes it challenging to fully tap into the modeling potential of powerful networks such as transformers. As a result, researchers have been actively investigating effective…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Yunyao Mao , Jiajun Deng , Wengang Zhou , Yao Fang , Wanli Ouyang , Houqiang Li

Although recent studies have made remarkable progress in human mesh recovery, they still exhibit limited robustness to occlusions and often produce inaccurate poses and severe motion jitter due to the insufficient spatial features for…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Tao Tang , Hong Liu , Xinshun Wang , Wanruo Zhang

The accuracy and robustness of 3D human pose estimation (HPE) are limited by 2D pose detection errors and 2D to 3D ill-posed challenges, which have drawn great attention to Multi-Hypothesis HPE research. Most existing MH-HPE methods are…

计算机视觉与模式识别 · 计算机科学 2024-05-06 Xianzhou Zeng , Hao Qin , Ming Kong , Luyuan Chen , Qiang Zhu

We propose an end-to-end unified 3D mesh recovery of humans and quadruped animals trained in a weakly-supervised way. Unlike recent work focusing on a single target class only, we aim to recover 3D mesh of broader classes with a single…

计算机视觉与模式识别 · 计算机科学 2021-11-05 Kim Youwang , Kim Ji-Yeon , Kyungdon Joo , Tae-Hyun Oh

Estimating 3d human pose from monocular images is a challenging problem due to the variety and complexity of human poses and the inherent ambiguity in recovering depth from the single view. Recent deep learning based methods show promising…

计算机视觉与模式识别 · 计算机科学 2019-05-06 Sandika Biswas , Sanjana Sinha , Kavya Gupta , Brojeshwar Bhowmick

Human pose forecasting is inherently multimodal since multiple futures exist for an observed pose sequence. However, evaluating multimodality is challenging since the task is ill-posed. Therefore, we first propose an alternative paradigm to…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Reyhaneh Hosseininejad , Megh Shukla , Saeed Saadatnejad , Mathieu Salzmann , Alexandre Alahi

This work aims to discuss the current landscape of kinematic analysis tools, ranging from the state-of-the-art in sports biomechanics such as inertial measurement units (IMUs) and retroreflective marker-based optical motion capture (MoCap)…

计算机视觉与模式识别 · 计算机科学 2025-03-20 Kai Armstrong , Alexander Rodrigues , Alexander P. Willmott , Lei Zhang , Xujiong Ye

Videos from edited media like movies are a useful, yet under-explored source of information. The rich variety of appearance and interactions between humans depicted over a large temporal context in these films could be a valuable source of…

计算机视觉与模式识别 · 计算机科学 2020-12-18 Georgios Pavlakos , Jitendra Malik , Angjoo Kanazawa

We consider the problem of obese human mesh recovery, i.e., fitting a parametric human mesh to images of obese people. Despite obese person mesh fitting being an important problem with numerous applications (e.g., healthcare), much recent…

计算机视觉与模式识别 · 计算机科学 2021-07-14 Ren Li , Meng Zheng , Srikrishna Karanam , Terrence Chen , Ziyan Wu