中文
相关论文

相关论文: Masked Modeling for Human Motion Recovery Under Oc…

200 篇论文

Recurrent neural networks are powerful tools for handling incomplete data problems in computer vision, thanks to their significant generative capabilities. However, the computational demand for these algorithms is too high to work in real…

计算机视觉与模式识别 · 计算机科学 2015-05-07 Ozgur Yilmaz

Video Frame Interpolation (VFI) aims to synthesize intermediate frames between existing frames to enhance visual smoothness and quality. Beyond the conventional methods based on the reconstruction loss, recent works have employed generative…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Jaihyun Lew , Jooyoung Choi , Chaehun Shin , Dahuin Jung , Sungroh Yoon

We introduce an approach for detecting and tracking detailed 3D poses of multiple people from a single monocular camera stream. Our system maintains temporally coherent predictions in crowded scenes filled with difficult poses and…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Alejandro Newell , Peiyun Hu , Lahav Lipson , Stephan R. Richter , Vladlen Koltun

Multimodal representation learning has shown promising improvements on various vision-language tasks. Most existing methods excel at building global-level alignment between vision and language while lacking effective fine-grained image-text…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Zijia Zhao , Longteng Guo , Xingjian He , Shuai Shao , Zehuan Yuan , Jing Liu

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

计算机视觉与模式识别 · 计算机科学 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

Recent advances in image-based human pose estimation make it possible to capture 3D human motion from a single RGB video. However, the inherent depth ambiguity and self-occlusion in a single view prohibit the recovery of as high-quality…

计算机视觉与模式识别 · 计算机科学 2020-08-20 Junting Dong , Qing Shuai , Yuanqing Zhang , Xian Liu , Xiaowei Zhou , Hujun Bao

3D Multi-object tracking (MOT) ensures consistency during continuous dynamic detection, conducive to subsequent motion planning and navigation tasks in autonomous driving. However, camera-based methods suffer in the case of occlusions and…

计算机视觉与模式识别 · 计算机科学 2022-09-13 Li Wang , Xinyu Zhang , Wenyuan Qin , Xiaoyu Li , Lei Yang , Zhiwei Li , Lei Zhu , Hong Wang , Jun Li , Huaping Liu

Recently, occluded person re-identification(Re-ID) remains a challenging task that people are frequently obscured by other people or obstacles, especially in a crowd massing situation. In this paper, we propose a self-supervised deep…

计算机视觉与模式识别 · 计算机科学 2022-02-11 Mi Zhou , Hongye Liu , Zhekun Lv , Wei Hong , Xiai Chen

Rendering the visual appearance of moving humans from occluded monocular videos is a challenging task. Most existing research renders 3D humans under ideal conditions, requiring a clear and unobstructed scene. Those methods cannot be used…

计算机视觉与模式识别 · 计算机科学 2025-08-18 Tiange Xiang , Adam Sun , Scott Delp , Kazuki Kozuka , Li Fei-Fei , Ehsan Adeli

Recently, implicit neural representation has been widely used to generate animatable human avatars. However, the materials and geometry of those representations are coupled in the neural network and hard to edit, which hinders their…

计算机视觉与模式识别 · 计算机科学 2024-05-21 Qifeng Chen , Rengan Xie , Kai Huang , Qi Wang , Wenting Zheng , Rong Li , Yuchi Huo

We present a deblurring method for scenes with occluding objects using a carefully designed layered blur model. Layered blur model is frequently used in the motion deblurring problem to handle locally varying blurs, which is caused by…

计算机视觉与模式识别 · 计算机科学 2016-11-30 Byeongjoo Ahn , Tae Hyun Kim , Wonsik Kim , Kyoung Mu Lee

We present a novel method to learn temporally consistent 3D reconstruction of clothed people from a monocular video. Recent methods for 3D human reconstruction from monocular video using volumetric, implicit or parametric human shape…

计算机视觉与模式识别 · 计算机科学 2021-04-20 Akin Caliskan , Armin Mustafa , Adrian Hilton

High-fidelity 3D scene reconstruction from monocular videos continues to be challenging, especially for complete and fine-grained geometry reconstruction. The previous 3D reconstruction approaches with neural implicit representations have…

计算机视觉与模式识别 · 计算机科学 2022-10-03 Zi-Xin Zou , Shi-Sheng Huang , Yan-Pei Cao , Tai-Jiang Mu , Ying Shan , Hongbo Fu

We introduce a new method that generates photo-realistic humans under novel views and poses given a monocular video as input. Despite the significant progress recently on this topic, with several methods exploring shared canonical neural…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Tiantian Wang , Nikolaos Sarafianos , Ming-Hsuan Yang , Tony Tung

Occlusion remains one of the major challenges in person reidentification (ReID) as a result of the diversity of poses and the variation of appearances. Developing novel architectures to improve the robustness of occlusion-aware person Re-ID…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Syeda Nyma Ferdous , Xin Li

Recent years have witnessed tremendous progress in the 3D reconstruction of dynamic humans from a monocular video with the advent of neural rendering techniques. This task has a wide range of applications, including the creation of virtual…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Kanghao Chen , Zeyu Wang , Lin Wang

Recent advances in transformer-based text-to-motion generation have led to impressive progress in synthesizing high-quality human motion. Nevertheless, jointly achieving high fidelity, streaming capability, real-time responsiveness, and…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Dongjie Fu , Tengjiao Sun , Pengcheng Fang , Xiaohao Cai , Hansung Kim

Human mesh recovery (HMR) models 3D human body from monocular videos, with recent works extending it to world-coordinate human trajectory and motion reconstruction. However, most existing methods remain offline, relying on future frames or…

计算机视觉与模式识别 · 计算机科学 2026-03-19 Yiwen Zhao , Ce Zheng , Yufu Wang , Hsueh-Han Daniel Yang , Liting Wen , Laszlo A. Jeni

This report reviews recent advancements in human motion prediction, reconstruction, and generation. Human motion prediction focuses on forecasting future poses and movements from historical data, addressing challenges like nonlinear…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Canxuan Gang , Yiran Wang

3D human pose estimation (HPE) is crucial in many fields, such as human behavior analysis, augmented reality/virtual reality (AR/VR) applications, and self-driving industry. Videos that contain multiple potentially occluded people captured…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Renshu Gu , Gaoang Wang , Jenq-Neng Hwang