中文
相关论文

相关论文: ReliaAvatar: A Robust Real-Time Avatar Animator wi…

200 篇论文

We present FHAvatar, a novel framework for reconstructing 3D Gaussian avatars with composable face and hair components from an arbitrary number of views. Unlike previous approaches that couple facial and hair representations within a…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Yujie Sun , Zhuoqiang Cai , Chaoyue Niu , Jianchuan Chen , Zhiwen Chen , Chengfei Lv , Fan Wu

We introduce LIA-X, a novel interpretable portrait animator designed to transfer facial dynamics from a driving video to a source portrait with fine-grained control. LIA-X is an autoencoder that models motion transfer as a linear navigation…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Yaohui Wang , Di Yang , Xinyuan Chen , Francois Bremond , Yu Qiao , Antitza Dantcheva

A photorealistic and immersive human avatar experience demands capturing fine, person-specific details such as cloth and hair dynamics, subtle facial expressions, and characteristic motion patterns. Achieving this requires large,…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Michael Steiner , Zhang Chen , Alexander Richard , Vasu Agrawal , Markus Steinberger , Michael Zollhöfer

We study the problem of aligning a video that captures a local portion of an environment to the 2D LiDAR scan of the entire environment. We introduce a method (VioLA) that starts with building a semantic map of the local scene from the…

计算机视觉与模式识别 · 计算机科学 2023-11-09 Jun-Jee Chao , Selim Engin , Nikhil Chavan-Dafle , Bhoram Lee , Volkan Isler

Advancements in neural implicit representations and differentiable rendering have markedly improved the ability to learn animatable 3D avatars from sparse multi-view RGB videos. However, current methods that map observation space to…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Zichen Tang , Hongyu Yang , Hanchen Zhang , Jiaxin Chen , Di Huang

Recent years have witnessed great progress in creating vivid audio-driven portraits from monocular videos. However, how to seamlessly adapt the created video avatars to other scenarios with different backgrounds and lighting conditions…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Haonan Qiu , Zhaoxi Chen , Yuming Jiang , Hang Zhou , Xiangyu Fan , Lei Yang , Wayne Wu , Ziwei Liu

Existing Vision-Language-Action (VLA) models often suffer from feature collapse and low training efficiency because they entangle high-level perception with sparse, embodiment-specific action supervision. Since these models typically rely…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Haitao Lin , Hanyang Yu , Jingshun Huang , He Zhang , Yonggen Ling , Ping Tan , Xiangyang Xue , Yanwei Fu

A core challenge for an agent learning to interact with the world is to predict how its actions affect objects in its environment. Many existing methods for learning the dynamics of physical interactions require labeled object information.…

机器学习 · 计算机科学 2016-10-19 Chelsea Finn , Ian Goodfellow , Sergey Levine

Talking head generation creates lifelike avatars from static portraits for virtual communication and content creation. However, current models do not yet convey the feeling of truly interactive communication, often generating one-way…

机器学习 · 计算机科学 2026-01-05 Taekyung Ki , Sangwon Jang , Jaehyeong Jo , Jaehong Yoon , Sung Ju Hwang

3D animation of humans in action is quite challenging as it involves using a huge setup with several motion trackers all over the person's body to track the movements of every limb. This is time-consuming and may cause the person discomfort…

图形学 · 计算机科学 2020-02-10 Laxman Kumarapu , Prerana Mukherjee

Controllable character animation remains a challenging problem, particularly in handling rare poses, stylized characters, character-object interactions, complex illumination, and dynamic scenes. To tackle these issues, prior work has…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Jingkai Zhou , Yifan Wu , Shikai Li , Min Wei , Chao Fan , Weihua Chen , Wei Jiang , Fan Wang

While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A critical challenge in utilizing large-scale human video datasets lies in transforming…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Dujun Nie , Fengjiao Chen , Qi Lv , Jun Kuang , Xiaoyu Li , Xuezhi Cao , Xunliang Cai

Inferring full-body poses from Head Mounted Devices, which capture only 3-joint observations from the head and wrists, is a challenging task with wide AR/VR applications. Previous attempts focus on learning one-stage motion mapping and thus…

计算机视觉与模式识别 · 计算机科学 2025-05-13 Fangyu Du , Yang Yang , Xuehao Gao , Hongye Hou

Creating high-fidelity, animatable 3D avatars from a single image remains a formidable challenge. We identified three desirable attributes of avatar generation: 1) the method should be feed-forward, 2) model a 360{\deg} full-head, and 3)…

图形学 · 计算机科学 2026-02-13 Zehao Xia , Yiqun Wang , Zhengda Lu , Kai Liu , Jun Xiao , Peter Wonka

The ability to predict motion in real time is fundamental to many maneuvering activities in animals, particularly those critical for survival, such as attack and escape responses. Given its significance, it is no surprise that motion…

图像与视频处理 · 电气工程与系统科学 2025-09-16 Subhradip Chakraborty , Shay Snyder , Md Abdullah-Al Kaiser , Maryam Parsa , Gregory Schwartz , Akhilesh R. Jaiswal

We propose a real-time 3D human pose estimation and motion analysis method termed RePose for rehabilitation training. It is capable of real-time monitoring and evaluation of patients'motion during rehabilitation, providing immediate…

计算机视觉与模式识别 · 计算机科学 2026-01-05 Junxiao Xue , Pavel Smirnov , Ziao Li , Yunyun Shi , Shi Chen , Xinyi Yin , Xiaohan Yue , Lei Wang , Yiduo Wang , Feng Lin , Yijia Chen , Xiao Ma , Xiaoran Yan , Qing Zhang , Fengjian Xue , Xuecheng Wu

3D human avatar animation aims at transforming a human avatar from an arbitrary initial pose to a specified target pose using deformation algorithms. Existing approaches typically divide this task into two stages: canonical template…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Jian Shu , Nanjie Yao , Gangjian Zhang , Junlong Ren , Yu Feng , Hao Wang

There has been a continued trend towards minimizing instrumentation for full-body motion capture, going from specialized rooms and equipment, to arrays of worn sensors and recently sparse inertial pose capture methods. However, as these…

人机交互 · 计算机科学 2025-04-18 Vasco Xu , Chenfeng Gao , Henry Hoffmann , Karan Ahuja

The motion capture system that supports full-body virtual representation is of key significance for virtual reality. Compared to vision-based systems, full-body pose estimation from sparse tracking signals is not limited by environmental…

计算机视觉与模式识别 · 计算机科学 2025-05-09 Zunjie Zhu , Yan Zhao , Yihan Hu , Guoxiang Wang , Hai Qiu , Bolun Zheng , Chenggang Yan , Feng Xu

Reliable and robust user identification and authentication are important and often necessary requirements for many digital services. It becomes paramount in social virtual reality (VR) to ensure trust, specifically in digital encounters…

机器学习 · 计算机科学 2023-10-27 Christian Schell , Andreas Hotho , Marc Erich Latoschik
‹ 上一页 1 8 9 10 下一页 ›