中文
相关论文

相关论文: WHAM: Reconstructing World-grounded Humans with Ac…

200 篇论文

We propose TRAM, a two-stage method to reconstruct a human's global trajectory and motion from in-the-wild videos. TRAM robustifies SLAM to recover the camera motion in the presence of dynamic humans and uses the scene background to derive…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Yufu Wang , Ziyun Wang , Lingjie Liu , Kostas Daniilidis

We propose a method to reconstruct global human trajectories from videos in the wild. Our optimization method decouples the camera and human motion, which allows us to place people in the same world coordinate frame. Most existing methods…

计算机视觉与模式识别 · 计算机科学 2023-03-22 Vickie Ye , Georgios Pavlakos , Jitendra Malik , Angjoo Kanazawa

In this paper, we present a novel framework designed to reconstruct long-sequence 3D human motion in the world coordinates from in-the-wild videos with multiple shot transitions. Such long-sequence in-the-wild motions are highly valuable to…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Yuhong Zhang , Guanlin Wu , Ling-Hao Chen , Zhuokai Zhao , Jing Lin , Xiaoke Jiang , Jiamin Wu , Zhuoheng Li , Hao Frank Yang , Haoqian Wang , Lei Zhang

We present a method to estimate human motion in a global scene from moving cameras. This is a highly challenging task due to the coupling of human and camera motions in the video. To address this problem, we propose a joint optimization…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Muhammed Kocabas , Ye Yuan , Pavlo Molchanov , Yunrong Guo , Michael J. Black , Otmar Hilliges , Jan Kautz , Umar Iqbal

Inferring 3D human motion from video remains a challenging problem with many applications. While traditional methods estimate the human in image coordinates, many applications require human motion to be estimated in world coordinates. This…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Joachim Tesch , Giorgio Becherini , Prerana Achar , Anastasios Yiannakidis , Muhammed Kocabas , Priyanka Patel , Michael J. Black

Estimating human and camera trajectories with accurate scale in the world coordinate system from a monocular video is a highly desirable yet challenging and ill-posed problem. In this study, we aim to recover expressive parametric human…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Wanqi Yin , Zhongang Cai , Ruisi Wang , Fanzhou Wang , Chen Wei , Haiyi Mei , Weiye Xiao , Zhitao Yang , Qingping Sun , Atsushi Yamashita , Ziwei Liu , Lei Yang

We present a novel method for recovering world-grounded human motion from monocular video. The main challenge lies in the ambiguity of defining the world coordinate system, which varies between sequences. Previous approaches attempt to…

计算机视觉与模式识别 · 计算机科学 2024-09-11 Zehong Shen , Huaijin Pi , Yan Xia , Zhi Cen , Sida Peng , Zechen Hu , Hujun Bao , Ruizhen Hu , Xiaowei Zhou

Despite the advent in 3D hand pose estimation, current methods predominantly focus on single-image 3D hand reconstruction in the camera frame, overlooking the world-space motion of the hands. Such limitation prohibits their direct use in…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Jinglei Zhang , Jiankang Deng , Chao Ma , Rolandos Alexandros Potamias

Demystifying complex human-ground interactions is essential for accurate and realistic 3D human motion reconstruction from RGB videos, as it ensures consistency between the humans and the ground plane. Prior methods have modeled…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Sihan Ma , Qiong Cao , Hongwei Yi , Jing Zhang , Dacheng Tao

Estimating human motion from video is an active research area due to its many potential applications. Most state-of-the-art methods predict human shape and posture estimates for individual images and do not leverage the temporal information…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Dorian F. Henning , Tristan Laidlow , Stefan Leutenegger

Recovering world-coordinate human motion from monocular videos with humanoid robot retargeting is significant for embodied intelligence and robotics. To avoid complex SLAM pipelines or heavy temporal models, we propose a lightweight,…

机器人学 · 计算机科学 2025-12-29 Zhangzheng Tu , Kailun Su , Shaolong Zhu , Yukun Zheng

Accurate camera motion estimation is essential for recovering global human motion in world coordinates from RGB video inputs. SLAM is widely used for estimating camera trajectory and point cloud, but monocular SLAM does so only up to an…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Fengyuan Yang , Kerui Gu , Ha Linh Nguyen , Tze Ho Elden Tse , Angela Yao

Global human motion reconstruction from in-the-wild monocular videos is increasingly demanded across VR, graphics, and robotics applications, yet requires accurate mapping of human poses from camera to world coordinates-a task challenged by…

计算机视觉与模式识别 · 计算机科学 2025-09-08 Qijun Ying , Zhongyuan Hu , Rui Zhang , Ronghui Li , Yu Lu , Zijiao Zeng

Accurate and robust 3D scene reconstruction from casual, in-the-wild videos can significantly simplify robot deployment to new environments. However, reliable camera pose estimation and scene reconstruction from such unconstrained videos…

计算机视觉与模式识别 · 计算机科学 2025-04-30 Shuo Sun , Torsten Sattler , Malcolm Mielle , Achim J. Lilienthal , Martin Magnusson

In multi-view human body capture systems, the recovered 3D geometry or even the acquired imagery data can be heavily corrupted due to occlusions, noise, limited field of- view, etc. Direct estimation of 3D pose, body shape or motion on…

计算机视觉与模式识别 · 计算机科学 2018-02-02 Zhong Li , Yu Ji , Wei Yang , Jinwei Ye , Jingyi Yu

We present an approach for 3D global human mesh recovery from monocular videos recorded with dynamic cameras. Our approach is robust to severe and long-term occlusions and tracks human bodies even when they go outside the camera's field of…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Ye Yuan , Umar Iqbal , Pavlo Molchanov , Kris Kitani , Jan Kautz

Existing deep models predict 2D and 3D kinematic poses from video that are approximately accurate, but contain visible errors that violate physical constraints, such as feet penetrating the ground and bodies leaning at extreme angles. In…

计算机视觉与模式识别 · 计算机科学 2020-07-27 Davis Rempe , Leonidas J. Guibas , Aaron Hertzmann , Bryan Russell , Ruben Villegas , Jimei Yang

Recent advances in 3D foundation models have led to growing interest in reconstructing humans and their surrounding environments. However, most existing approaches focus on monocular inputs, and extending them to multi-view settings…

计算机视觉与模式识别 · 计算机科学 2026-03-19 Sangmin Kim , Minhyuk Hwang , Geonho Cha , Dongyoon Wee , Jaesik Park

Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motion prediction, virtual reality, and digital twinning. Despite significant interest in…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Yeheng Zong , Pou-Chun Kung , Yike Pan , Seth Isaacson , Yizhou Chen , Ram Vasudevan , Katherine A. Skinner

Analyzing human motion is a challenging task with a wide variety of applications in computer vision and in graphics. One such application, of particular importance in computer animation, is the retargeting of motion from one performer to…

计算机视觉与模式识别 · 计算机科学 2019-05-13 Kfir Aberman , Rundi Wu , Dani Lischinski , Baoquan Chen , Daniel Cohen-Or
‹ 上一页 1 2 3 10 下一页 ›