中文
相关论文

相关论文: PHUMA: Physically-Grounded Humanoid Locomotion Dat…

200 篇论文

One of the key arguments for building robots that have similar form factors to human beings is that we can leverage the massive human data for training. Yet, doing so has remained challenging in practice due to the complexities in humanoid…

机器人学 · 计算机科学 2024-06-18 Zipeng Fu , Qingqing Zhao , Qi Wu , Gordon Wetzstein , Chelsea Finn

High-quality data collection is a fundamental cornerstone for training humanoid whole-body visuomotor policies. Current data acquisition paradigms predominantly rely on robot teleoperation, which is often hindered by limited hardware…

机器人学 · 计算机科学 2026-05-06 Chenhao Yu , Hongwu Wang , Youhao Hu , Jiachen Zhang , Yuanyuan Li , Shaqi Luo

Current humanoid motion tracking systems can execute routine and moderately dynamic behaviors, yet significant gaps remain near hardware performance limits and algorithmic robustness boundaries. Martial arts represent an extreme case of…

机器人学 · 计算机科学 2026-02-17 Zhongxiang Lei , Lulu Cao , Xuyang Wang , Tianyi Qian , Jinyan Liu , Xuesong Li

Reconstructing physically plausible human motion from monocular videos remains a challenging problem in computer vision and graphics. Existing methods primarily focus on kinematics-based pose estimation, often leading to unrealistic results…

计算机视觉与模式识别 · 计算机科学 2025-10-06 Qiao Feng , Yiming Huang , Yufu Wang , Jiatao Gu , Lingjie Liu

Imitation learning is a promising approach for training humanoid robots to both walk and manipulate, but it requires a large number of demonstrations, which are time-intensive and difficult to collect via teleoperation. Existing…

Visual imitation learning provides a framework for learning complex manipulation behaviors by leveraging human demonstrations. However, current interfaces for imitation such as kinesthetic teaching or teleoperation prohibitively restrict…

机器人学 · 计算机科学 2020-08-12 Sarah Young , Dhiraj Gandhi , Shubham Tulsiani , Abhinav Gupta , Pieter Abbeel , Lerrel Pinto

Humanoids have the potential to be the ideal embodiment in environments designed for humans. Thanks to the structural similarity to the human body, they benefit from rich sources of demonstration data, e.g., collected via teleoperation,…

机器人学 · 计算机科学 2024-11-05 Oleg Kaidanov , Firas Al-Hafez , Yusuf Suvari , Boris Belousov , Jan Peters

Understanding human motion from video is essential for a range of applications, including pose estimation, mesh recovery and action recognition. While state-of-the-art methods predominantly rely on transformer-based architectures, these…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Arnab Kumar Mondal , Stefano Alletto , Denis Tome

Pushing is a motion primitive useful to handle objects that are too large, too heavy, or too cluttered to be grasped. It is at the core of much of robotic manipulation, in particular when physical interaction is involved. It seems…

机器人学 · 计算机科学 2016-08-05 Kuan-Ting Yu , Maria Bauza , Nima Fazeli , Alberto Rodriguez

Humanoid robots are promising to acquire various skills by imitating human behaviors. However, existing algorithms are only capable of tracking smooth, low-speed human motions, even with delicate reward and curriculum design. This paper…

机器人学 · 计算机科学 2025-10-28 Weiji Xie , Jinrui Han , Jiakun Zheng , Huanyu Li , Xinzhe Liu , Jiyuan Shi , Weinan Zhang , Chenjia Bai , Xuelong Li

A long-standing goal in computer vision is to capture, model, and realistically synthesize human behavior. Specifically, by learning from data, our goal is to enable virtual humans to navigate within cluttered indoor scenes and naturally…

计算机视觉与模式识别 · 计算机科学 2021-08-19 Mohamed Hassan , Duygu Ceylan , Ruben Villegas , Jun Saito , Jimei Yang , Yi Zhou , Michael Black

Recently, humanoid robots have made significant advances in their ability to perform challenging tasks due to the deployment of Reinforcement Learning (RL), however, the inherent complexity of humanoid robots, including the difficulty of…

机器人学 · 计算机科学 2024-08-27 Qiang Zhang , Peter Cui , David Yan , Jingkai Sun , Yiqun Duan , Gang Han , Wen Zhao , Weining Zhang , Yijie Guo , Arthur Zhang , Renjing Xu

Multimodal Deep Learning enhances decision-making by integrating diverse information sources, such as texts, images, audio, and videos. To develop trustworthy multimodal approaches, it is essential to understand how uncertainty impacts…

机器学习 · 计算机科学 2025-08-14 Grigor Bezirganyan , Sana Sellami , Laure Berti-Équille , Sébastien Fournier

We present a new method for generating controllable, dynamically responsive, and photorealistic human animations. Given an image of a person, our system allows the user to generate Physically plausible Upper Body Animation (PUBA) using…

计算机视觉与模式识别 · 计算机科学 2022-12-12 Ziyuan Huang , Zhengping Zhou , Yung-Yu Chuang , Jiajun Wu , C. Karen Liu

There are several challenges in developing a model for multi-tasking humanoid control. Reinforcement learning and imitation learning approaches are quite popular in this domain. However, there is a trade-off between the two. Reinforcement…

机器人学 · 计算机科学 2024-06-18 Siddharth Padmanabhan , Kazuki Miyazawa , Takato Horii , Takayuki Nagai

Recent advances in generative video modeling, driven by large-scale datasets and powerful architectures, have yielded remarkable visual realism. However, emerging evidence suggests that simply scaling data and model size does not endow…

计算机视觉与模式识别 · 计算机科学 2026-05-20 Ying Shen , Jerry Xiong , Tianjiao Yu , Ismini Lourentzou

Falls represent a significant cause of injury among the elderly population. Extensive research has been devoted to the utilization of wearable IMU sensors in conjunction with machine learning techniques for fall detection. To address the…

定量方法 · 定量生物学 2023-10-18 Jie Tang , Bin He , Junkai Xu , Tian Tan , Zhipeng Wang , Yanmin Zhou , Shuo Jiang

Accurate and physically feasible human motion prediction is crucial for safe and seamless human-robot collaboration. While recent advancements in human motion capture enable real-time pose estimation, the practical value of many existing…

We present MAMMA, a markerless motion-capture pipeline that accurately recovers SMPL-X parameters from multi-view video of two-person interaction sequences. Traditional motion-capture systems rely on physical markers. Although they offer…

We present ActiveUMI, a framework for a data collection system that transfers in-the-wild human demonstrations to robots capable of complex bimanual manipulation. ActiveUMI couples a portable VR teleoperation kit with sensorized controllers…

机器人学 · 计算机科学 2025-10-03 Qiyuan Zeng , Chengmeng Li , Jude St. John , Zhongyi Zhou , Junjie Wen , Guorui Feng , Yichen Zhu , Yi Xu