中文
相关论文

相关论文: Full Body Video-Based Self-Avatars for Mixed Reali…

200 篇论文

We present Neural Head Avatars, a novel neural representation that explicitly models the surface geometry and appearance of an animatable human avatar that can be used for teleconferencing in AR/VR or other applications in the movie or…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Philip-William Grassal , Malte Prinzler , Titus Leistner , Carsten Rother , Matthias Nießner , Justus Thies

Virtual Reality (VR) is an exciting new consumer technology which offers an immersive audio-visual experience to users through which they can navigate and interact with a digitally represented 3D space (i.e., a virtual world) using a…

密码学与安全 · 计算机科学 2024-04-16 Mohd Sabra , Nisha Vinayaga Sureshkanth , Ari Sharma , Anindya Maiti , Murtuza Jadliwala

We study the task of embodied visual active learning, where an agent is set to explore a 3d environment with the goal to acquire visual scene understanding by actively selecting views for which to request annotation. While accurate on some…

计算机视觉与模式识别 · 计算机科学 2020-12-18 David Nilsson , Aleksis Pirinen , Erik Gärtner , Cristian Sminchisescu

In this paper, we present an end-to-end pipeline for the creation of high-quality animatable volumetric video content of human performances. Going beyond the application of free-viewpoint volumetric video, we allow re-animation and…

计算机视觉与模式识别 · 计算机科学 2020-09-03 Anna Hilsmann , Philipp Fechteler , Wieland Morgenstern , Wolfgang Paier , Ingo Feldmann , Oliver Schreer , Peter Eisert

Facial expression and hand motions are necessary to express our emotions and interact with the world. Nevertheless, most of the 3D human avatars modeled from a casually captured video only support body motions without facial expressions and…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Gyeongsik Moon , Takaaki Shiratori , Shunsuke Saito

This study will investigate the user experience while interacting with highly photorealistic virtual job interviewer avatars in Virtual Reality (VR), Augmented Reality (AR), and on a 2D screen. Having a precise speech recognition mechanism,…

We introduce Visual Persona, a foundation model for text-to-image full-body human customization that, given a single in-the-wild human image, generates diverse images of the individual guided by text descriptions. Unlike prior methods that…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Jisu Nam , Soowon Son , Zhan Xu , Jing Shi , Difan Liu , Feng Liu , Aashish Misraa , Seungryong Kim , Yang Zhou

In this paper, we propose a novel pipeline for the 3D reconstruction of the full body from egocentric viewpoints. 3-D reconstruction of the human body from egocentric viewpoints is a challenging task as the view is skewed and the body parts…

计算机视觉与模式识别 · 计算机科学 2021-11-11 Shivam Grover , Kshitij Sidana , Vanita Jain

Egocentric body tracking, also known as inside-out body tracking (IOBT), is an essential technology for applications like gesture control and codec avatar in mixed reality (MR), including augmented reality (AR) and virtual reality (VR).…

图像与视频处理 · 电气工程与系统科学 2025-09-11 Yin Li , Sean Korphi , Sam Shiu , Yasuo Morimoto , Jiang Zhu , Rajalakshimi Nandakumar

This paper investigates the problem of understanding dynamic 3D scenes from egocentric observations, a key challenge in robotics and embodied AI. Unlike prior studies that explored this as long-form video understanding and utilized…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Yue Fan , Xiaojian Ma , Rongpeng Su , Jun Guo , Rujie Wu , Xi Chen , Qing Li

We present IntrinsicAvatar, a novel approach to recovering the intrinsic properties of clothed human avatars including geometry, albedo, material, and environment lighting from only monocular videos. Recent advancements in human-based…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Shaofei Wang , Božidar Antić , Andreas Geiger , Siyu Tang

Recently, self-supervised pre-training has advanced Vision Transformers on various tasks w.r.t. different data modalities, e.g., image and 3D point cloud data. In this paper, we explore this learning paradigm for 3D mesh data analysis based…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Yaqian Liang , Shanshan Zhao , Baosheng Yu , Jing Zhang , Fazhi He

Vision-language pre-training (VLP) on large-scale image-text pairs has achieved huge success for the cross-modal downstream tasks. The most existing pre-training methods mainly adopt a two-step training procedure, which firstly employs a…

计算机视觉与模式识别 · 计算机科学 2021-06-07 Haiyang Xu , Ming Yan , Chenliang Li , Bin Bi , Songfang Huang , Wenming Xiao , Fei Huang

The application and implementation of collaborative embodiment in virtual reality (VR) are a critical aspect of the computer science landscape, aiming to enhance multi-user interaction and teamwork in immersive environments. A notable and…

人机交互 · 计算机科学 2025-07-28 Hongyu Zhou , Yihao Dong , Masahiko Inami , Zhanna Sarsenbayeva , Anusha Withana

We introduce FlexAvatar, a method for creating high-quality and complete 3D head avatars from a single image. A core challenge lies in the limited availability of multi-view data and the tendency of monocular training to yield incomplete 3D…

计算机视觉与模式识别 · 计算机科学 2026-04-09 Tobias Kirschstein , Simon Giebenhain , Matthias Nießner

Modern medicine demands innovations in medical education, particularly in the learning of human anatomy, traditionally taught through textbooks, dissections, and lectures. Virtual Reality (VR) has emerged as a promising tool to address the…

人机交互 · 计算机科学 2024-11-11 Myint Zu Than , Kian Meng Yap

We present the first approach to render highly realistic free-viewpoint videos of a human actor in general apparel, from sparse multi-view recording to display, in real-time at an unprecedented 4K resolution. At inference, our method only…

计算机视觉与模式识别 · 计算机科学 2024-07-26 Ashwath Shetty , Marc Habermann , Guoxing Sun , Diogo Luvizon , Vladislav Golyanik , Christian Theobalt

With the recent explosive growth of interest and investment in virtual reality (VR) and the so-called "metaverse," public attention has rightly shifted toward the unique security and privacy threats that these platforms may pose. While it…

密码学与安全 · 计算机科学 2023-10-25 Vivek Nair , Wenbo Guo , Justus Mattern , Rui Wang , James F. O'Brien , Louis Rosenberg , Dawn Song

Current video-based Masked Autoencoders (MAEs) primarily focus on learning effective spatiotemporal representations from a visual perspective, which may lead the model to prioritize general spatial-temporal patterns but often overlook…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Shihab Aaqil Ahamed , Malitha Gunawardhana , Liel David , Michael Sidorov , Daniel Harari , Muhammad Haris Khan

We present a new solution to egocentric 3D body pose estimation from monocular images captured from a downward looking fish-eye camera installed on the rim of a head mounted virtual reality device. This unusual viewpoint, just 2 cm. away…

计算机视觉与模式识别 · 计算机科学 2019-07-24 Denis Tome , Patrick Peluse , Lourdes Agapito , Hernan Badino