English
Related papers

Related papers: Motion-X: A Large-scale 3D Expressive Whole-body H…

200 papers

In this paper, we introduce Motion-X++, a large-scale multimodal 3D expressive whole-body human motion dataset. Existing motion datasets predominantly capture body-only poses, lacking facial expressions, hand gestures, and fine-grained pose…

Computer Vision and Pattern Recognition · Computer Science 2025-01-10 Yuhong Zhang , Jing Lin , Ailing Zeng , Guanlin Wu , Shunlin Lu , Yurong Fu , Yuanhao Cai , Ruimao Zhang , Haoqian Wang , Lei Zhang

This paper introduces OmniMotion-X, a versatile multimodal framework for whole-body human motion generation, leveraging an autoregressive diffusion transformer in a unified sequence-to-sequence manner. OmniMotion-X efficiently supports…

Computer Vision and Pattern Recognition · Computer Science 2025-10-23 Guowei Xu , Yuxuan Bian , Ailing Zeng , Mingyi Shi , Shaoli Huang , Wen Li , Lixin Duan , Qiang Xu

Generating realistic human motions from textual descriptions has undergone significant advancements. However, existing methods often overlook specific body part movements and their timing. In this paper, we address this issue by enriching…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Bizhu Wu , Jinheng Xie , Meidan Ding , Zhe Kong , Jianfeng Ren , Ruibin Bai , Rong Qu , Linlin Shen

Bimanual human activities inherently involve coordinated movements of both hands and body. However, the impact of this coordination in activity understanding has not been systematically evaluated due to the lack of suitable datasets. Such…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Tatsuro Banno , Takehiko Ohkawa , Ruicong Liu , Ryosuke Furuta , Yoichi Sato

The analysis of the ubiquitous human-human interactions is pivotal for understanding humans as social beings. Existing human-human interaction datasets typically suffer from inaccurate body motions, lack of hand gestures and fine-grained…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Liang Xu , Xintao Lv , Yichao Yan , Xin Jin , Shuwen Wu , Congsheng Xu , Yifan Liu , Yizhou Zhou , Fengyun Rao , Xingdong Sheng , Yunhui Liu , Wenjun Zeng , Xiaokang Yang

We introduce MotionScript, a novel framework for generating highly detailed, natural language descriptions of 3D human motions. Unlike existing motion datasets that rely on broad action labels or generic captions, MotionScript provides…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Payam Jome Yazdian , Rachel Lagasse , Hamid Mohammadi , Eric Liu , Li Cheng , Angelica Lim

Recent advances in 3D human motion and language integration have primarily focused on text-to-motion generation, leaving the task of motion understanding relatively unexplored. We introduce Dense Motion Captioning, a novel task that aims to…

Computer Vision and Pattern Recognition · Computer Science 2025-11-10 Shiyao Xu , Benedetta Liberatori , Gül Varol , Paolo Rota

In this paper, we introduce RoleMotion, a large-scale human motion dataset that encompasses a wealth of role-playing and functional motion data tailored to fit various specific scenes. Existing text datasets are mainly constructed…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Junran Peng , Yiheng Huang , Silei Shen , Zeji Wei , Jingwei Yang , Baojie Wang , Yonghao He , Chuanchen Luo , Man Zhang , Xucheng Yin , Wei Sui

Character image animation, which generates high-quality videos from a reference image and target pose sequence, has seen significant progress in recent years. However, most existing methods only apply to human figures, which usually do not…

Computer Vision and Pattern Recognition · Computer Science 2024-12-12 Shuai Tan , Biao Gong , Xiang Wang , Shiwei Zhang , Dandan Zheng , Ruobing Zheng , Kecheng Zheng , Jingdong Chen , Ming Yang

Synthesizing human motion has advanced rapidly, yet realistic hand motion and bimanual interaction remain underexplored. Whole-body models often miss the fine-grained cues that drive dexterous behavior, finger articulation, contact timing,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Zimu Zhang , Yucheng Zhang , Xiyan Xu , Ziyin Wang , Sirui Xu , Kai Zhou , Bing Zhou , Chuan Guo , Jian Wang , Yu-Xiong Wang , Liang-Yan Gui

3D multi-person motion prediction is a challenging task that involves modeling individual behaviors and interactions between people. Despite the emergence of approaches for this task, comparing them is difficult due to the lack of…

Computer Vision and Pattern Recognition · Computer Science 2023-06-27 Xiaogang Peng , Xiao Zhou , Yikai Luo , Hao Wen , Yu Ding , Zizhao Wu

We introduce X-Dyna, a novel zero-shot, diffusion-based pipeline for animating a single human image using facial expressions and body movements derived from a driving video, that generates realistic, context-aware dynamics for both the…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Di Chang , Hongyi Xu , You Xie , Yipeng Gao , Zhengfei Kuang , Shengqu Cai , Chenxu Zhang , Guoxian Song , Chao Wang , Yichun Shi , Zeyuan Chen , Shijie Zhou , Linjie Luo , Gordon Wetzstein , Mohammad Soleymani

We present X-UniMotion, a unified and expressive implicit latent representation for whole-body human motion, encompassing facial expressions, body poses, and hand gestures. Unlike prior motion transfer methods that rely on explicit skeletal…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Guoxian Song , Hongyi Xu , Xiaochen Zhao , You Xie , Tianpei Gu , Zenan Li , Chenxu Zhang , Linjie Luo

Scalable learning of humanoid robots is crucial for their deployment in real-world applications. While traditional approaches primarily rely on reinforcement learning or teleoperation to achieve whole-body control, they are often limited by…

The Codec Avatars Lab at Meta introduces Embody 3D, a multimodal dataset of 500 individual hours of 3D motion data from 439 participants collected in a multi-camera collection stage, amounting to over 54 million frames of tracked 3D motion.…

In this paper, we introduce a novel path to $\textit{general}$ human motion generation by focusing on 2D space. Traditional methods have primarily generated human motions in 3D, which, while detailed and realistic, are often limited by the…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Yuan Wang , Zhao Wang , Junhao Gong , Di Huang , Tong He , Wanli Ouyang , Jile Jiao , Xuetao Feng , Qi Dou , Shixiang Tang , Dan Xu

The generation of humanoid animation from text prompts can profoundly impact animation production and AR/VR experiences. However, existing methods only generate body motion data, excluding facial expressions and hand movements. This…

Computer Vision and Pattern Recognition · Computer Science 2024-09-23 Mingdian Liu , Yilin Liu , Gurunandan Krishnan , Karl S Bayer , Bing Zhou

While reconstructing human poses in 3D from inexpensive sensors has advanced significantly in recent years, quantifying the dynamics of human motion, including the muscle-generated joint torques and external forces, remains a challenge.…

To understand and analyze human behavior, we need to capture humans moving in, and interacting with, the world. Most existing methods perform 3D human pose estimation without explicitly considering the scene. We observe however that the…

Computer Vision and Pattern Recognition · Computer Science 2019-08-21 Mohamed Hassan , Vasileios Choutas , Dimitrios Tzionas , Michael J. Black

We present AnimaX, a feed-forward 3D animation framework that bridges the motion priors of video diffusion models with the controllable structure of skeleton-based animation. Traditional motion synthesis methods are either restricted to…

Computer Vision and Pattern Recognition · Computer Science 2025-06-25 Zehuan Huang , Haoran Feng , Yangtian Sun , Yuanchen Guo , Yanpei Cao , Lu Sheng
‹ Prev 1 2 3 10 Next ›