English
Related papers

Related papers: PhySIC: Physically Plausible 3D Human-Scene Intera…

200 papers

We present a fully automatic system that takes a 3D scene and generates plausible 3D human bodies that are posed naturally in that 3D scene. Given a 3D scene without people, humans can easily imagine how people could interact with the scene…

Computer Vision and Pattern Recognition · Computer Science 2020-04-21 Yan Zhang , Mohamed Hassan , Heiko Neumann , Michael J. Black , Siyu Tang

We present a new end-to-end learning framework to obtain detailed and spatially coherent reconstructions of multiple people from a single image. Existing multi-person methods suffer from two main drawbacks: they are often model-based and…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Armin Mustafa , Akin Caliskan , Lourdes Agapito , Adrian Hilton

Reconstructing 3D Human-Object Interaction from an RGB image is essential for perceptive systems. Yet, this remains challenging as it requires capturing the subtle physical coupling between the body and objects. While current methods rely…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Dimitrije Antić , Alvaro Budria , George Paschalidis , Sai Kumar Dwivedi , Dimitrios Tzionas

Humanoid motion control has witnessed significant breakthroughs in recent years, with deep reinforcement learning (RL) emerging as a primary catalyst for achieving complex, human-like behaviors. However, the high dimensionality and…

3D human pose estimation from a monocular video has recently seen significant improvements. However, most state-of-the-art methods are kinematics-based, which are prone to physically implausible motions with pronounced artifacts. Current…

Computer Vision and Pattern Recognition · Computer Science 2022-09-20 Jiefeng Li , Siyuan Bian , Chao Xu , Gang Liu , Gang Yu , Cewu Lu

Recovering high-quality 3D scenes from a single RGB image is a challenging task in computer graphics. Current methods often struggle with domain-specific limitations or low-quality object generation. To address these, we propose CAST…

Computer Vision and Pattern Recognition · Computer Science 2025-05-14 Kaixin Yao , Longwen Zhang , Xinhao Yan , Yan Zeng , Qixuan Zhang , Wei Yang , Lan Xu , Jiayuan Gu , Jingyi Yu

A long-standing goal of 3D human reconstruction is to create lifelike and fully detailed 3D humans from single-view images. The main challenge lies in inferring unknown body shapes, appearances, and clothing details in areas not visible in…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Hsuan-I Ho , Jie Song , Otmar Hilliges

Synthesizing 3D human avatars interacting realistically with a scene is an important problem with applications in AR/VR, video games and robotics. Towards this goal, we address the task of generating a virtual human -- hands and full body…

Robotics · Computer Science 2023-03-30 Purva Tendulkar , Dídac Surís , Carl Vondrick

We present Human3R, a unified, feed-forward framework for online 4D human-scene reconstruction, in the world frame, from casually captured monocular videos. Unlike previous approaches that rely on multi-stage pipelines, iterative…

Computer Vision and Pattern Recognition · Computer Science 2026-03-04 Yue Chen , Xingyu Chen , Yuxuan Xue , Anpei Chen , Yuliang Xiu , Gerard Pons-Moll

This paper presents a novel generative approach that outputs 3D indoor environments solely from a textual description of the scene. Current methods often treat scene synthesis as a mere layout prediction task, leading to rooms with…

Machine Learning · Computer Science 2025-02-12 Yao Wei , Matteo Toso , Pietro Morerio , Michael Ying Yang , Alessio Del Bue

Human video synthesis aims to create lifelike characters in various environments, with wide applications in VR, storytelling, and content creation. While 2D diffusion-based methods have made significant progress, they struggle to generalize…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Liyuan Cui , Xiaogang Xu , Wenqi Dong , Zesong Yang , Hujun Bao , Zhaopeng Cui

We propose LaplacianFusion, a novel approach that reconstructs detailed and controllable 3D clothed-human body shapes from an input depth or 3D point cloud sequence. The key idea of our approach is to use Laplacian coordinates, well-known…

Graphics · Computer Science 2023-03-01 Hyomin Kim , Hyeonseo Nam , Jungeon Kim , Jaesik Park , Seungyong Lee

Despite significant progress in single image-based 3D human mesh recovery, accurately and smoothly recovering 3D human motion from a video remains challenging. Existing video-based methods generally recover human mesh by estimating the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Yingxuan You , Hong Liu , Ti Wang , Wenhao Li , Runwei Ding , Xia Li

Surface reconstruction from multiple, calibrated images is a challenging task - often requiring a large number of collected images with significant overlap. We look at the specific case of human foot reconstruction. As with previous…

Computer Vision and Pattern Recognition · Computer Science 2025-02-19 Oliver Boyne , Roberto Cipolla

Accurate 6D object pose estimation from images is a key problem in object-centric scene understanding, enabling applications in robotics, augmented reality, and scene reconstruction. Despite recent advances, existing methods often produce…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Martin Malenický , Martin Cífka , Médéric Fourmy , Louis Montaut , Justin Carpentier , Josef Sivic , Vladimir Petrik

We present a computational framework that transforms single images into 3D physical objects. The visual geometry of a physical object in an image is determined by three orthogonal attributes: mechanical properties, external forces, and…

Computer Vision and Pattern Recognition · Computer Science 2025-01-03 Minghao Guo , Bohan Wang , Pingchuan Ma , Tianyuan Zhang , Crystal Elaine Owens , Chuang Gan , Joshua B. Tenenbaum , Kaiming He , Wojciech Matusik

In this paper, we tackle the problem of scene-aware 3D human motion forecasting. A key challenge of this task is to predict future human motions that are consistent with the scene by modeling the human-scene interactions. While recent works…

Computer Vision and Pattern Recognition · Computer Science 2024-08-13 Chaoyue Xing , Wei Mao , Miaomiao Liu

Estimating human pose and shape from monocular images is a long-standing problem in computer vision. Since the release of statistical body models, 3D human mesh recovery has been drawing broader attention. With the same goal of obtaining…

Computer Vision and Pattern Recognition · Computer Science 2024-01-03 Yating Tian , Hongwen Zhang , Yebin Liu , Limin Wang

We introduce Intrinsic Image Fusion, a method that reconstructs high-quality physically based materials from multi-view images. Material reconstruction is highly underconstrained and typically relies on analysis-by-synthesis, which requires…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Peter Kocsis , Lukas Höllein , Matthias Nießner

We present "Humans and Structure from Motion" (HSfM), a method for jointly reconstructing multiple human meshes, scene point clouds, and camera parameters in a metric world coordinate system from a sparse set of uncalibrated multi-view…

Computer Vision and Pattern Recognition · Computer Science 2025-05-22 Lea Müller , Hongsuk Choi , Anthony Zhang , Brent Yi , Jitendra Malik , Angjoo Kanazawa