English
Related papers

Related papers: Full Body Video-Based Self-Avatars for Mixed Reali…

200 papers

Recent progress in video-to-video (V2V) translation has enabled realistic resimulation of embodied AI demonstrations, a capability that allows pretrained robot policies to be transferable to new environments without additional data…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 George Eskandar , Fengyi Shen , Mohammad Altillawi , Dong Chen , Yang Bai , Liudi Yang , Ziyuan Liu

Providing a depth-rich Virtual Reality (VR) experience to users without causing discomfort remains to be a challenge with today's commercially available head-mounted displays (HMDs), which enforce strict measures on stereoscopic camera…

Graphics · Computer Science 2019-11-12 Emre Avan , Ufuk Celikcan , Tolga K. Capin , Hasmet Gurcay

In recent years, deep learning has made great progress in many fields such as image recognition, natural language processing, speech recognition and video super-resolution. In this survey, we comprehensively investigate 33 state-of-the-art…

Computer Vision and Pattern Recognition · Computer Science 2022-03-17 Hongying Liu , Zhubo Ruan , Peng Zhao , Chao Dong , Fanhua Shang , Yuanyuan Liu , Linlin Yang , Radu Timofte

Large-scale pre-training using egocentric human videos has proven effective for robot learning. However, the models pre-trained on such data can be suboptimal for robot learning due to the significant visual gap between human hands and…

Robotics · Computer Science 2026-03-17 Guangrun Li , Yaoxu Lyu , Zhuoyang Liu , Chengkai Hou , Jieyu Zhang , Shanghang Zhang

User embodiment is important for many virtual reality (VR) applications, for example, in the context of social interaction, therapy, training, or entertainment. However, there is no validated instrument to empirically measure the perception…

Human-Computer Interaction · Computer Science 2020-10-01 Daniel Roth , Marc Erich Latoschik

With the growing integration of human-computer interaction into everyday life, advances in machine learning have enabled systems to better perceive and respond to users' emotional states. Most existing affect recognition datasets focus on…

Machine Learning · Computer Science 2026-05-04 Karim Alghoul , Faisal Mohd , Fedwa Laamarti , Hussein Al Osman , Abdulmotaleb El Saddik

Virtual Reality (VR) becomes accessible to mimic a "real-like" world now. People who have a VR experience usually can be impressed by the immersive feeling, they might consider themselves are actually existed in the VR space.…

Human-Computer Interaction · Computer Science 2017-07-12 Peikun Xiong , Chen Sun , Dongsheng Cai

Photorealistic telepresence requires both high-fidelity body modeling and faithful driving to enable dynamically synthesized appearance that is indistinguishable from reality. In this work, we propose an end-to-end framework that addresses…

Computer Vision and Pattern Recognition · Computer Science 2022-07-21 Edoardo Remelli , Timur Bagautdinov , Shunsuke Saito , Tomas Simon , Chenglei Wu , Shih-En Wei , Kaiwen Guo , Zhe Cao , Fabian Prada , Jason Saragih , Yaser Sheikh

We introduce a deep appearance model for rendering the human face. Inspired by Active Appearance Models, we develop a data-driven rendering pipeline that learns a joint representation of facial geometry and appearance from a multiview…

Graphics · Computer Science 2018-08-02 Stephen Lombardi , Jason Saragih , Tomas Simon , Yaser Sheikh

We present SimXR, a method for controlling a simulated avatar from information (headset pose and cameras) obtained from AR / VR headsets. Due to the challenging viewpoint of head-mounted cameras, the human body is often clipped out of view,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-26 Zhengyi Luo , Jinkun Cao , Rawal Khirodkar , Alexander Winkler , Jing Huang , Kris Kitani , Weipeng Xu

In this paper, we propose a novel hybrid representation and end-to-end trainable network architecture to model fully editable and customizable neural avatars. At the core of our work lies a representation that combines the modeling power of…

Computer Vision and Pattern Recognition · Computer Science 2023-05-02 Hsuan-I Ho , Lixin Xue , Jie Song , Otmar Hilliges

Volumetric (4D) performance capture is fundamental for AR/VR content generation. Whereas previous work in 4D performance capture has shown impressive results in studio settings, the technology is still far from being accessible to a typical…

Recently, the advancement of self-supervised learning techniques, like masked autoencoders (MAE), has greatly influenced visual representation learning for images and videos. Nevertheless, it is worth noting that the predominant approaches…

Computer Vision and Pattern Recognition · Computer Science 2024-03-01 Gensheng Pei , Tao Chen , Xiruo Jiang , Huafeng Liu , Zeren Sun , Yazhou Yao

Consumer 3D scanners and depth cameras are increasingly being used to generate content and avatars for Virtual Reality (VR) environments and avoid the inconveniences of hand modeling; however, it is sometimes difficult to evaluate…

Human-Computer Interaction · Computer Science 2017-02-01 Jacob Thorn , Rodrigo Pizarro , Bernhard Spanlang , Pablo Bermell-Garcia , Mar Gonzalez-Franco

Despite recent progress in developing animatable full-body avatars, realistic modeling of clothing - one of the core aspects of human self-expression - remains an open challenge. State-of-the-art physical simulation methods can generate…

Mixed Reality (MR) technologies such as Virtual and Augmented Reality (VR, AR) are well established in medical practice, enhancing diagnostics, treatment, and education. However, there are still some limitations and challenges that may be…

Human-Computer Interaction · Computer Science 2025-07-28 Aliaksandr Marozau , Barbara Karpowicz , Tomasz Kowalewski , Pavlo Zinevych , Wiktor Stawski , Adam Kuzdraliński , Wiesław Kopeć

We present a system to create Mobile Realistic Fullbody (MoRF) avatars. MoRF avatars are rendered in real-time on mobile devices, learned from monocular videos, and have high realism. We use SMPL-X as a proxy geometry and render it with DNR…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Renat Bashirov , Alexey Larionov , Evgeniya Ustinova , Mikhail Sidorenko , David Svitov , Ilya Zakharkin , Victor Lempitsky

Recent advancements in Multi-modal Large Language Models (MLLMs) have opened new avenues for applications in Embodied AI. Building on previous work, EgoThink, we introduce VidEgoThink, a comprehensive benchmark for evaluating egocentric…

Computer Vision and Pattern Recognition · Computer Science 2024-10-16 Sijie Cheng , Kechen Fang , Yangyang Yu , Sicheng Zhou , Bohao Li , Ye Tian , Tingguang Li , Lei Han , Yang Liu

Teleconference or telepresence based on virtual reality (VR) headmount display (HMD) device is a very interesting and promising application since HMD can provide immersive feelings for users. However, in order to facilitate face-to-face…

Computer Vision and Pattern Recognition · Computer Science 2019-01-23 Guoxian Song , Jianfei Cai , Tat-Jen Cham , Jianmin Zheng , Juyong Zhang , Henry Fuchs

In contemporary biology and medicine, 3D microscopy is one of the most widely-used techniques for imaging and manipulation of various kinds of samples. Navigating such a micrometer-sized, 3-dimensional sample under the microscope -- e.g. to…

Human-Computer Interaction · Computer Science 2026-03-26 Jan Tiemann , Matthew McGinity , Ulrik Günther