English
Related papers

Related papers: ReliaAvatar: A Robust Real-Time Avatar Animator wi…

200 papers

Full-body egocentric pose estimation from head and hand poses alone has become an active area of research to power articulate avatar representations on headset-based platforms. However, existing methods over-rely on the indoor…

Computer Vision and Pattern Recognition · Computer Science 2024-09-09 Jiaxi Jiang , Paul Streli , Manuel Meier , Christian Holz

Action-conditioned video prediction models (often referred to as world models) have shown strong potential for robotics applications, but existing approaches are often slow and struggle to capture physically consistent interactions over…

Creating realistic, fully animatable whole-body avatars from a single portrait is challenging due to limitations in capturing subtle expressions, body movements, and dynamic backgrounds. Current evaluation datasets and metrics fall short in…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Chaoyi Wang , Yifan Yang , Jun Pei , Lijie Xia , Jianpo Liu , Xiaobing Yuan , Xinhan Di

In recent years, there has been significant interest in creating 3D avatars and motions, driven by their diverse applications in areas like film-making, video games, AR/VR, and human-robot interaction. However, current efforts primarily…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Zeyu Zhang , Yiran Wang , Biao Wu , Shuo Chen , Zhiyuan Zhang , Shiya Huang , Wenbo Zhang , Meng Fang , Ling Chen , Yang Zhao

To bridge the physical and virtual worlds for rapidly developed VR/AR applications, the ability to realistically drive 3D full-body avatars is of great significance. Although real-time body tracking with only the head-mounted displays…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Xiaozheng Zheng , Zhuo Su , Chao Wen , Zhou Xue , Xiaojie Jin

We present R3-Avatar, incorporating a temporal codebook, to overcome the inability of human avatars to be both animatable and of high-fidelity rendering quality. Existing video-based reconstruction of 3D human avatars either focuses solely…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Yifan Zhan , Wangze Xu , Qingtian Zhu , Muyao Niu , Mingze Ma , Yifei Liu , Zhihang Zhong , Xiao Sun , Yinqiang Zheng

We tackle the task of recovering an animatable 3D human avatar from a single or a sparse set of images. For this task, beyond a set of images, many prior state-of-the-art methods use accurate "ground-truth" camera poses and human poses as…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Jing Wen , Alexander G. Schwing , Shenlong Wang

The creation of 3D human avatars from multi-view videos is a significant yet challenging task in computer vision. However, existing techniques rely on high-quality, sharp images as input, which are often impractical to obtain in real-world…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Muyao Niu , Yifan Zhan , Qingtian Zhu , Zhuoxiao Li , Wei Wang , Zhihang Zhong , Xiao Sun , Yinqiang Zheng

The appearance of a human in clothing is driven not only by the pose but also by its temporal context, i.e., motion. However, such context has been largely neglected by existing monocular human modeling methods whose neural networks often…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Hansol Lee , Junuk Cha , Yunhoe Ku , Jae Shin Yoon , Seungryul Baek

Realistic hair motion is crucial for high-quality avatars, but it is often limited by the computational resources available for real-time applications. To address this challenge, we propose a novel neural approach to predict physically…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Tuur Stuyck , Gene Wei-Chin Lin , Egor Larionov , Hsiao-yu Chen , Aljaz Bozic , Nikolaos Sarafianos , Doug Roble

Building realistic and animatable avatars still requires minutes of multi-view or monocular self-rotating videos, and most methods lack precise control over gestures and expressions. To push this boundary, we address the challenge of…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Jun Xiang , Yudong Guo , Leipeng Hu , Boyang Guo , Yancheng Yuan , Juyong Zhang

Estimating robot pose from RGB images is a crucial problem in computer vision and robotics. While previous methods have achieved promising performance, most of them presume full knowledge of robot internal states, e.g. ground-truth robot…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Shikun Ban , Juling Fan , Xiaoxuan Ma , Wentao Zhu , Yu Qiao , Yizhou Wang

Controllability, generalizability and efficiency are the major objectives of constructing face avatars represented by neural implicit field. However, existing methods have not managed to accommodate the three requirements simultaneously.…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Zhiyuan Ma , Xiangyu Zhu , Guojun Qi , Zhen Lei , Lei Zhang

Vision-Language-Action (VLA) models trained via imitation learning suffer from significant performance degradation in data-scarce scenarios due to their reliance on large-scale demonstration datasets. Although reinforcement learning…

Robotics · Computer Science 2026-04-28 Junjin Xiao , Yandan Yang , Xinyuan Chang , Ronghan Chen , Feng Xiong , Mu Xu , Wei-Shi Zheng , Qing Zhang

Despite recent progress in developing animatable full-body avatars, realistic modeling of clothing - one of the core aspects of human self-expression - remains an open challenge. State-of-the-art physical simulation methods can generate…

In the realm of Virtual Reality (VR) and Human-Computer Interaction (HCI), real-time emotion recognition shows promise for supporting individuals with Autism Spectrum Disorder (ASD) in improving social skills. This task requires a strict…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Yarin Benyamin

In this paper, we take a significant step towards real-world applicability of monocular neural avatar reconstruction by contributing InstantAvatar, a system that can reconstruct human avatars from a monocular video within seconds, and these…

Computer Vision and Pattern Recognition · Computer Science 2022-12-21 Tianjian Jiang , Xu Chen , Jie Song , Otmar Hilliges

Vision-Language-Action (VLA) models offer a compelling framework for tackling complex robotic manipulation tasks, but they are often expensive to train. In this paper, we propose a novel VLA approach that leverages the competitive…

Robotics · Computer Science 2025-12-23 Max Argus , Jelena Bratulic , Houman Masnavi , Maxim Velikanov , Nick Heppert , Abhinav Valada , Thomas Brox

Photorealistic Codec Avatars (PCA), which generate high-fidelity human face renderings, are increasingly being used in Virtual Reality (VR) environments to enable immersive communication and interaction through deep learning-based…

Computer Vision and Pattern Recognition · Computer Science 2025-10-30 Mingzhi Zhu , Ding Shang , Sai Qian Zhang

Vision-Language-Action (VLA) models often suffer from performance degradation under distribution shifts, as they struggle to learn generalized behavior representations across varying environments. While existing approaches attempt to…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Bing Hu , Zaijing Li , Rui Shao , Junda Chen , April Hua Liu , Wei-Shi Zheng , Liqiang Nie
‹ Prev 1 4 5 6 7 8 10 Next ›