中文
相关论文

相关论文: Hyperbolic Space Learning Method Leveraging Tempor…

200 篇论文

Learning the representation of data with hierarchical structures in the hyperbolic space attracts increasing attention in recent years. Due to the constant negative curvature, the hyperbolic space resembles tree metrics and captures the…

机器学习 · 计算机科学 2022-02-21 Huiru Xiao , Caigao Jiang , Yangqiu Song , James Zhang , Junwu Xiong

Scene graph representations enable structured visual understanding by modeling objects and their relationships, and have been widely used for multiview and 3D scene reasoning. Existing methods such as MSG learn scene graph embeddings in…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Liyang Wang , Zeyu Zhang , Hao Tang

We present a unified perspective on tackling various human-centric video tasks by learning human motion representations from large-scale and heterogeneous data resources. Specifically, we propose a pretraining stage in which a motion…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Wentao Zhu , Xiaoxuan Ma , Zhaoyang Liu , Libin Liu , Wayne Wu , Yizhou Wang

Monocular video human mesh recovery is essential for digital humans, avatar animation, and embodied simulation, where both temporal stability and expressive whole-body motion are required. Existing video HMR methods produce coherent body…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Wenhao Shen , Ming Zhou , Hengyuan Zhang , Siyuan Bian , Youjiang Xu , Xi Lin

We present a novel method for populating 3D indoor scenes with virtual humans that can navigate in the environment and interact with objects in a realistic manner. Existing approaches rely on training sequences that contain captured human…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Kaifeng Zhao , Yan Zhang , Shaofei Wang , Thabo Beeler , Siyu Tang

Recent years have witnessed significant progress in 3D hand mesh recovery. Nevertheless, because of the intrinsic 2D-to-3D ambiguity, recovering camera-space 3D information from a single RGB image remains challenging. To tackle this…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Xingyu Chen , Yufeng Liu , Chongyang Ma , Jianlong Chang , Huayan Wang , Tian Chen , Xiaoyan Guo , Pengfei Wan , Wen Zheng

Although significant progress has been achieved on monocular maker-less human motion capture in recent years, it is still hard for state-of-the-art methods to obtain satisfactory results in occlusion scenarios. There are two main reasons:…

计算机视觉与模式识别 · 计算机科学 2022-07-13 Buzhen Huang , Yuan Shu , Jingyi Ju , Yangang Wang

We have recently seen tremendous progress in the neural advances for photo-real human modeling and rendering. However, it's still challenging to integrate them into an existing mesh-based pipeline for downstream applications. In this paper,…

We introduce MetricHMSR, a novel framework for recovering metric human meshes and 3D scenes from a single monocular image. Existing methods struggle to recover metric scale due to monocular scale ambiguity and weak-perspective camera…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Chentao Song , He Zhang , Haolei Yuan , Haozhe Lin , Jianhua Tao , Hongwen Zhang , Tao Yu

With 3D data rapidly emerging as an important form of multimedia information, 3D human mesh recovery technology has also advanced accordingly. However, current methods mainly focus on handling humans wearing tight clothing and perform…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Yunqi Gao , Leyuan Liu , Yuhan Li , Changxin Gao , Yuanyuan Liu , Jingying Chen

Finding meaningful representations and distances of hierarchical data is important in many fields. This paper presents a new method for hierarchical data embedding and distance. Our method relies on combining diffusion geometry, a central…

机器学习 · 计算机科学 2023-05-31 Ya-Wei Eileen Lin , Ronald R. Coifman , Gal Mishne , Ronen Talmon

Representation learning has become an invaluable approach for learning from symbolic data such as text and graphs. However, while complex symbolic datasets often exhibit a latent hierarchical structure, state-of-the-art methods typically…

人工智能 · 计算机科学 2017-05-29 Maximilian Nickel , Douwe Kiela

Current approaches in 3D human pose estimation primarily focus on regressing 3D joint locations, often neglecting critical physical constraints such as bone length consistency and body symmetry. This work introduces a recurrent neural…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Chih-Hsiang Hsu , Jyh-Shing Roger Jang

Human motion taxonomies serve as high-level hierarchical abstractions that classify how humans move and interact with their environment. They have proven useful to analyse grasps, manipulation skills, and whole-body support poses. Despite…

机器人学 · 计算机科学 2024-09-17 Noémie Jaquier , Leonel Rozo , Miguel González-Duque , Viacheslav Borovitskiy , Tamim Asfour

We introduce SAM 3D Body (3DB), a promptable model for single-image full-body 3D human mesh recovery (HMR) that demonstrates state-of-the-art performance, with strong generalization and consistent accuracy in diverse in-the-wild conditions.…

Human Mesh Recovery (HMR) is fundamentally ambiguous: under occlusion or weak depth cues, multiple 3D bodies can explain the same image evidence. This ambiguity is not uniform across the body, as torso pose and root structure are often…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Patrick Kwon , Chen Chen

Human motion recovered from monocular videos often appears overly smooth or dynamically inconsistent, even when joint positions are numerically accurate. We observe that this limitation stems from the absence of reliable high-order temporal…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Dingkun Wei , Zehong Shen , Yan Xia , Georgios Pavlakos , Yujun Shen , Xiaowei Zhou

Despite the impressive results achieved by deep learning based 3D reconstruction, the techniques of directly learning to model 4D human captures with detailed geometry have been less studied. This work presents a novel framework that can…

计算机视觉与模式识别 · 计算机科学 2022-04-20 Boyan Jiang , Yinda Zhang , Xingkui Wei , Xiangyang Xue , Yanwei Fu

Although existing video-based 3D human mesh recovery methods have made significant progress, simultaneously estimating human pose and shape from low-resolution image features limits their performance. These image features lack sufficient…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Tao Tang , Hong Liu , Yingxuan You , Ti Wang , Wenhao Li

Monocular 3D human pose and shape estimation is an inherently ill-posed problem due to depth ambiguities, occlusions, and truncations. Recent probabilistic approaches learn a distribution over plausible 3D human meshes by maximizing the…

计算机视觉与模式识别 · 计算机科学 2024-11-26 Tom Wehrbein , Marco Rudolph , Bodo Rosenhahn , Bastian Wandt