English

Generating Attribute-Aware Human Motions from Textual Prompt

Computer Vision and Pattern Recognition 2025-11-14 v2 Multimedia

Abstract

Text-driven human motion generation has recently attracted considerable attention, allowing models to generate human motions based on textual descriptions. However, current methods neglect the influence of human attributes-such as age, gender, weight, and height-which are key factors shaping human motion patterns. This work represents a pilot exploration for bridging this gap. We conceptualize each motion as comprising both attribute information and action semantics, where textual descriptions align exclusively with action semantics. To achieve this, a new framework inspired by Structural Causal Models is proposed to decouple action semantics from human attributes, enabling text-to-semantics prediction and attribute-controlled generation. The resulting model is capable of generating attribute-aware motion aligned with the user's text and attribute inputs. For evaluation, we introduce a comprehensive dataset containing attribute annotations for text-motion pairs, setting the first benchmark for attribute-aware motion generation. Extensive experiments validate our model's effectiveness.

Keywords

Cite

@article{arxiv.2506.21912,
  title  = {Generating Attribute-Aware Human Motions from Textual Prompt},
  author = {Xinghan Wang and Kun Xu and Fei Li and Cao Sheng and Jiazhong Yu and Yadong Mu},
  journal= {arXiv preprint arXiv:2506.21912},
  year   = {2025}
}

Comments

Accepted by AAAI 2026

R2 v1 2026-07-01T03:35:47.263Z