中文
相关论文

相关论文: Real-Time Human Frontal View Synthesis from a Sing…

200 篇论文

Photorealistic and animatable human avatars are a key enabler for virtual/augmented reality, telepresence, and digital entertainment. While recent advances in 3D Gaussian Splatting (3DGS) have greatly improved rendering quality and…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Cheng Peng , Jingxiang Sun , Yushuo Chen , Zhaoqi Su , Zhuo Su , Yebin Liu

Amidst the rapid advancement of camera-based autonomous driving technology, effectiveness is often prioritized with limited attention to computational efficiency. To address this issue, this paper introduces LRHPerception, a real-time…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Haixi Zhang , Aiyinsi Zuo , Zirui Li , Chunshu Wu , Tong Geng , Zhiyao Duan

Human Mesh Recovery (HMR) from an image is a challenging problem because of the inherent ambiguity of the task. Existing HMR methods utilized either temporal information or kinematic relationships to achieve higher accuracy, but there is no…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Hanbyel Cho , Jaesung Ahn , Yooshin Cho , Junmo Kim

The promise of unsupervised multi-view-stereo (MVS) is to leverage large unlabeled datasets, yet current methods underperform when training on difficult data, such as handheld smartphone videos of indoor scenes. Meanwhile, high-quality…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Alex Rich , Noah Stier , Pradeep Sen , Tobias Höllerer

Synthesizing realistic videos of humans using neural networks has been a popular alternative to the conventional graphics-based rendering pipeline due to its high efficiency. Existing works typically formulate this as an image-to-image…

Volumetric modeling and neural radiance field representations have revolutionized 3D face capture and photorealistic novel view synthesis. However, these methods often require hundreds of multi-view input images and are thus inapplicable to…

In this paper, we propose a novel learning approach for feed-forward one-shot 4D head avatar synthesis. Different from existing methods that often learn from reconstructing monocular videos guided by 3DMM, we employ pseudo multi-view videos…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Yu Deng , Duomin Wang , Baoyuan Wang

Recovering 3D human mesh from monocular images is a popular topic in computer vision and has a wide range of applications. This paper aims to estimate 3D mesh of multiple body parts (e.g., body, hands) with large-scale differences from a…

计算机视觉与模式识别 · 计算机科学 2020-10-28 Yu Sun , Qian Bao , Wu Liu , Wenpeng Gao , Yili Fu , Chuang Gan , Tao Mei

In this paper, we tackle the problem of generating a novel image from an arbitrary viewpoint given a single frame as input. While existing methods operating in this setup aim at predicting the target view depth map to guide the synthesis,…

计算机视觉与模式识别 · 计算机科学 2023-08-29 Giovanni Minelli , Matteo Poggi , Samuele Salti

The advancement in deep implicit modeling and articulated models has significantly enhanced the process of digitizing human figures in 3D from just a single image. While state-of-the-art methods have greatly improved geometric precision,…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Vishnu Mani Hema , Shubhra Aich , Christian Haene , Jean-Charles Bazin , Fernando de la Torre

This paper presents a novel technique for progressive online integration of uncalibrated image sequences with substantial geometric and/or photometric discrepancies into a single, geometrically and photometrically consistent image. Our…

图形学 · 计算机科学 2019-09-11 Markus Kluge , Tim Weyrich , Andreas Kolb

Reconstructing textured 3D human models from a single image is fundamental for AR/VR and digital human applications. However, existing methods mostly focus on single individuals and thus fail in multi-human scenes, where naive composition…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Gwanghyun Kim , Junghun James Kim , Suh Yoon Jeon , Jason Park , Se Young Chun

Accurate reconstruction of complex dynamic scenes from just a single viewpoint continues to be a challenging task in computer vision. Current dynamic novel view synthesis methods typically require videos from many different camera…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Basile Van Hoorick , Rundi Wu , Ege Ozguroglu , Kyle Sargent , Ruoshi Liu , Pavel Tokmakov , Achal Dave , Changxi Zheng , Carl Vondrick

To tackle the challeging problem of multi-person 3D pose estimation from a single image, we propose a multi-view matching (MVM) method in this work. The MVM method generates reliable 3D human poses from a large-scale video dataset, called…

计算机视觉与模式识别 · 计算机科学 2020-04-14 Yeji Shen , C. -C. Jay Kuo

We present an approach to infer a layer-structured 3D representation of a scene from a single input image. This allows us to infer not only the depth of the visible pixels, but also to capture the texture and depth for content in the scene…

计算机视觉与模式识别 · 计算机科学 2018-07-27 Shubham Tulsiani , Richard Tucker , Noah Snavely

Mirror reflections are common in everyday environments and can provide stereo information within a single capture, as the real and reflected virtual views are visible simultaneously. We exploit this property by treating the reflection as an…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Jing Wu , Zirui Wang , Iro Laina , Victor Adrian Prisacariu

While head-mounted displays (HMDs) for Virtual Reality (VR) have become widely available in the consumer market, they pose a considerable obstacle for a realistic face-to-face conversation in VR since HMDs hide a significant portion of the…

计算机视觉与模式识别 · 计算机科学 2024-09-20 Philipp Ladwig , Rene Ebertowski , Alexander Pech , Ralf Dörner , Christian Geiger

We present MagicMirror, a framework for generating identity-preserved videos with cinematic-level quality and dynamic motion. While recent advances in video diffusion models have shown impressive capabilities in text-to-video generation,…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Yuechen Zhang , Yaoyang Liu , Bin Xia , Bohao Peng , Zexin Yan , Eric Lo , Jiaya Jia

Novel view synthesis from a single image has recently attracted a lot of attention, and it has been primarily advanced by 3D deep learning and rendering techniques. However, most work is still limited by synthesizing new views within…

计算机视觉与模式识别 · 计算机科学 2022-03-18 Xuanchi Ren , Xiaolong Wang

A recent strand of work in view synthesis uses deep learning to generate multiplane images (a camera-centric, layered 3D representation) given two or more input images at known viewpoints. We apply this representation to single-view view…

计算机视觉与模式识别 · 计算机科学 2020-04-24 Richard Tucker , Noah Snavely