中文
相关论文

相关论文: STAR: Sparse Trained Articulated Human Body Regres…

200 篇论文

Current state-of-the-art in 3D human pose and shape recovery relies on deep neural networks and statistical morphable body models, such as the Skinned Multi-Person Linear model (SMPL). However, regardless of the advantages of having both…

计算机视觉与模式识别 · 计算机科学 2019-08-09 Meysam Madadi , Hugo Bertiche , Sergio Escalera

Many human pose estimation methods estimate Skinned Multi-Person Linear (SMPL) models and regress the human joints from these SMPL estimates. In this work, we show that the most widely used SMPL-to-joint linear layer (joint regressor) is…

计算机视觉与模式识别 · 计算机科学 2022-05-03 Eric Hedlin , Helge Rhodin , Kwang Moo Yi

The Skinned Multi-Person Linear (SMPL) model can represent a human body by mapping pose and shape parameters to body meshes. This has been shown to facilitate inferring 3D human pose and shape from images via different learning models.…

计算机视觉与模式识别 · 计算机科学 2021-12-09 Andrey Davydov , Anastasia Remizova , Victor Constantin , Sina Honari , Mathieu Salzmann , Pascal Fua

Existing Transformers for monocular 3D human shape and pose estimation typically have a quadratic computation and memory complexity with respect to the feature length, which hinders the exploitation of fine-grained information in…

计算机视觉与模式识别 · 计算机科学 2024-04-24 Xiangyu Xu , Lijuan Liu , Shuicheng Yan

This paper addresses the problem of monocular 3D human shape and pose estimation from an RGB image. Despite great progress in this field in terms of pose prediction accuracy, state-of-the-art methods often predict inaccurate body shapes. We…

计算机视觉与模式识别 · 计算机科学 2020-09-23 Akash Sengupta , Ignas Budvytis , Roberto Cipolla

Great progress has been made in estimating 3D human pose and shape from images and video by training neural networks to directly regress the parameters of parametric human models like SMPL. However, existing body models have simplified…

图形学 · 计算机科学 2025-09-09 Marilyn Keller , Keenon Werling , Soyong Shin , Scott Delp , Sergi Pujades , C. Karen Liu , Michael J. Black

Model merging is an efficient way of obtaining a multi-task model from several pretrained models without further fine-tuning, and it has gained attention in various domains, including natural language processing (NLP). Despite the…

计算与语言 · 计算机科学 2025-02-17 Yu-Ang Lee , Ching-Yun Ko , Tejaswini Pedapati , I-Hsin Chung , Mi-Yen Yeh , Pin-Yu Chen

Human pose estimation in low-resolution videos presents a fundamental challenge in computer vision. Conventional methods either assume high-quality inputs or employ computationally expensive cascaded processing, which limits their…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Yucheng Jin , Jinyan Chen , Ziyue He , Baojun Han , Furan An

We describe the first method to automatically estimate the 3D pose of the human body as well as its 3D shape from a single unconstrained image. We estimate a full 3D mesh and show that 2D joints alone carry a surprising amount of…

计算机视觉与模式识别 · 计算机科学 2016-07-28 Federica Bogo , Angjoo Kanazawa , Christoph Lassner , Peter Gehler , Javier Romero , Michael J. Black

Statistical 3D shape models of the head, hands, and fullbody are widely used in computer vision and graphics. Despite their wide use, we show that existing models of the head and hands fail to capture the full range of motion for these…

计算机视觉与模式识别 · 计算机科学 2022-10-26 Ahmed A. A. Osman , Timo Bolkart , Dimitrios Tzionas , Michael J. Black

Tensors are becoming prevalent in modern applications such as medical imaging and digital marketing. In this paper, we propose a sparse tensor additive regression (STAR) that models a scalar response as a flexible nonparametric function of…

机器学习 · 统计学 2021-03-08 Botao Hao , Boxiang Wang , Pengyuan Wang , Jingfei Zhang , Jian Yang , Will Wei Sun

Motion retargeting seeks to faithfully replicate the spatio-temporal motion characteristics of a source character onto a target character with a different body shape. Apart from motion semantics preservation, ensuring geometric plausibility…

计算机视觉与模式识别 · 计算机科学 2025-07-31 Xiaohang Yang , Qing Wang , Jiahao Yang , Gregory Slabaugh , Shanxin Yuan

In multi-view 3D human pose estimation, models typically rely on images captured simultaneously from different camera views to predict a pose at a specific moment. While providing accurate spatial information, this traditional approach…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Ling Li , Changjie Chen , Yuyan Wang , Jiaqing Lyu , Kenglun Chang , Yiyun Chen , Zhidong Deng

Reconstructing posed 3D human models from monocular images has important applications in the sports industry, including performance tracking, injury prevention and virtual training. In this work, we combine 3D human pose and shape…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Lorenza Prospero , Abdullah Hamdi , Joao F. Henriques , Christian Rupprecht

Compared to joint position, the accuracy of joint rotation and shape estimation has received relatively little attention in the skinned multi-person linear model (SMPL)-based human mesh reconstruction from multi-view images. The work in…

计算机视觉与模式识别 · 计算机科学 2022-08-25 Sungho Chun , Sungbum Park , Ju Yong Chang

The remarkable success of Large Language Models (LLMs) relies heavily on their substantial scale, which poses significant challenges during model deployment in terms of latency and memory consumption. Recently, numerous studies have…

计算与语言 · 计算机科学 2024-12-19 Weiyu Huang , Yuezhou Hu , Guohao Jian , Jun Zhu , Jianfei Chen

Large language models (LLMs) rely on self-attention for contextual understanding, demanding high-throughput inference and large-scale token parallelism (LTPP). Existing dynamic sparsity accelerators falter under LTPP scenarios due to…

硬件体系结构 · 计算机科学 2025-12-25 Huizheng Wang , Taiquan Wei , Hongbin Wang , Zichuan Wang , Xinru Tang , Zhiheng Yue , Shaojun Wei , Yang Hu , Shouyi Yin

We introduce STAR, a text-to-image model that employs a scale-wise auto-regressive paradigm. Unlike VAR, which is constrained to class-conditioned synthesis for images up to 256$\times$256, STAR enables text-driven image generation up to…

计算机视觉与模式识别 · 计算机科学 2025-02-20 Xiaoxiao Ma , Mohan Zhou , Tao Liang , Yalong Bai , Tiejun Zhao , Biye Li , Huaian Chen , Yi Jin

Learning to regress 3D human body shape and pose (e.g.~SMPL parameters) from monocular images typically exploits losses on 2D keypoints, silhouettes, and/or part-segmentation when 3D training data is not available. Such losses, however, are…

计算机视觉与模式识别 · 计算机科学 2022-02-24 Sai Kumar Dwivedi , Nikos Athanasiou , Muhammed Kocabas , Michael J. Black

We propose a scalable neural network framework to reconstruct the 3D mesh of a human body from multi-view images, in the subspace of the SMPL model. Use of multi-view images can significantly reduce the projection ambiguity of the problem,…

计算机视觉与模式识别 · 计算机科学 2019-08-27 Junbang Liang , Ming C. Lin
‹ 上一页 1 2 3 10 下一页 ›