中文
相关论文

相关论文: RITA: A Real-time Interactive Talking Avatars Fram…

200 篇论文

We present a novel approach for generating realistic speaking and talking faces by synthesizing a person's voice and facial movements from a static image, a voice profile, and a target text. The model encodes the prompt/driving text, the…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Aashish Chandra , Aashutosh A , Abhijit Das

We tackle the problem of quantifying the number of objects by a generative text-to-image model. Rather than retraining such a model for each new image domain of interest, which leads to high computational costs and limited scalability, we…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Wenfang Sun , Yingjun Du , Gaowen Liu , Yefeng Zheng , Cees G. M. Snoek

Timeseries analytics is of great importance in many real-world applications. Recently, the Transformer model, popular in natural language processing, has been leveraged to learn high quality feature embeddings from timeseries, core to the…

机器学习 · 计算机科学 2023-06-06 Jiaming Liang , Lei Cao , Samuel Madden , Zachary Ives , Guoliang Li

We propose RANA, a relightable and articulated neural avatar for the photorealistic synthesis of humans under arbitrary viewpoints, body poses, and lighting. We only require a short video clip of the person to create the avatar and assume…

计算机视觉与模式识别 · 计算机科学 2022-12-07 Umar Iqbal , Akin Caliskan , Koki Nagano , Sameh Khamis , Pavlo Molchanov , Jan Kautz

Recent advances have demonstrated compelling capabilities in synthesizing real individuals into generated videos, reflecting the growing demand for identity-aware content creation. Nevertheless, an openly accessible framework enabling…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Yingjie Chen , Shilun Lin , Cai Xing , Binxin Yang , Long Zhou , Qixin Yan , Wenjing Wang , Dingming Liu , Hao Liu , Chen Li , Jing Lyu

With the rising interest from the community in digital avatars coupled with the importance of expressions and gestures in communication, modeling natural avatar behavior remains an important challenge across many industries such as…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Kefan Chen , Sergiu Oprea , Justin Theiss , Sreyas Mohan , Srinath Sridhar , Aayush Prakash

Real-time synthesis of physically plausible human interactions remains a critical challenge for immersive VR/AR systems and humanoid robotics. While existing methods demonstrate progress in kinematic motion generation, they often fail to…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Kaiyang Ji , Ye Shi , Zichen Jin , Kangyi Chen , Lan Xu , Yuexin Ma , Jingyi Yu , Jingya Wang

We present an audio-driven real-time system for animating photorealistic 3D facial avatars with minimal latency, designed for social interactions in virtual reality for anyone. Central to our approach is an encoder model that transforms…

图形学 · 计算机科学 2025-11-04 Jiye Lee , Chenghui Li , Linh Tran , Shih-En Wei , Jason Saragih , Alexander Richard , Hanbyul Joo , Shaojie Bai

In this paper, we introduce FitMe, a facial reflectance model and a differentiable rendering optimization pipeline, that can be used to acquire high-fidelity renderable human avatars from single or multiple images. The model consists of a…

计算机视觉与模式识别 · 计算机科学 2023-05-17 Alexandros Lattas , Stylianos Moschoglou , Stylianos Ploumpis , Baris Gecer , Jiankang Deng , Stefanos Zafeiriou

Integrating generative AI such as Large Language Models into social robots has improved their ability to engage in natural, human-like communication. This study presents a method to examine their persuasive capabilities. We designed an…

机器人学 · 计算机科学 2026-02-10 Stephan Vonschallen , Larissa Julia Corina Finsler , Theresa Schmiedel , Friederike Eyssel

Foundation models have had a big impact in recent years and billions of dollars are being invested in them in the current AI boom. The more popular ones, such as Chat-GPT, are trained on large amounts of Internet data. However, it is…

人工智能 · 计算机科学 2024-08-07 Dionis Barcari , David Gamez , Aliya Grig

Controllability, generalizability and efficiency are the major objectives of constructing face avatars represented by neural implicit field. However, existing methods have not managed to accommodate the three requirements simultaneously.…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Zhiyuan Ma , Xiangyu Zhu , Guojun Qi , Zhen Lei , Lei Zhang

We introduce SimAvatar, a framework designed to generate simulation-ready clothed 3D human avatars from a text prompt. Current text-driven human avatar generation methods either model hair, clothing, and the human body using a unified…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Xueting Li , Ye Yuan , Shalini De Mello , Gilles Daviet , Jonathan Leaf , Miles Macklin , Jan Kautz , Umar Iqbal

As consumer adoption of immersive technologies grows, virtual avatars will play a prominent role in the future of social computing. However, as people begin to interact more frequently through virtual avatars, it is important to ensure that…

人机交互 · 计算机科学 2023-11-27 Tiffany D. Do , Steve Zelenty , Mar Gonzalez-Franco , Ryan P. McMahan

In this study, we propose a novel approach that supports requirements discussions in virtual environments by automatically generating personas from real-time speech-to-text data. In our pilot experiment, 18 participants (14 from…

人机交互 · 计算机科学 2026-02-05 Yi Wang , Zhengxin Zhang , Xiao Liu , Chetan Arora , John Grundy , Thuong Hoang

Human beings are social animals. How to equip 3D autonomous characters with similar social intelligence that can perceive, understand and interact with humans remains an open yet foundamental problem. In this paper, we introduce SOLAMI, the…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Jianping Jiang , Weiye Xiao , Zhengyu Lin , Huaizhong Zhang , Tianxiang Ren , Yang Gao , Zhiqian Lin , Zhongang Cai , Lei Yang , Ziwei Liu

We introduce PICA, a novel representation for high-fidelity animatable clothed human avatars with physics-accurate dynamics, even for loose clothing. Previous neural rendering-based representations of animatable clothed humans typically…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Bo Peng , Yunfan Tao , Haoyu Zhan , Yudong Guo , Juyong Zhang

With the emergence of increasingly powerful large language models, there is a burgeoning interest in leveraging these models for casual conversation and role-play applications. However, existing conversational and role-playing datasets…

计算与语言 · 计算机科学 2023-08-14 Tear Gosling , Alpin Dale , Yinhe Zheng

Vision-Language-Action (VLA) models enable embodied decision-making but rely heavily on imitation learning, leading to compounding errors and poor robustness under distribution shift. Reinforcement learning (RL) can mitigate these issues…

Incorporating personas information allows diverse and engaging responses in dialogue response generation. Unfortunately, prior works have primarily focused on self personas and have overlooked the value of partner personas. Moreover, in…

计算与语言 · 计算机科学 2021-11-30 Hongyuan Lu , Wai Lam , Hong Cheng , Helen M. Meng