中文
相关论文

相关论文: Audio-Driven Talking Face Generation with Blink Em…

200 篇论文

We present MultiNeRF, a 3D watermarking method that embeds multiple uniquely keyed watermarks within images rendered by a single Neural Radiance Field (NeRF) model, whilst maintaining high visual quality. Our approach extends the TensoRF…

计算机视觉与模式识别 · 计算机科学 2025-04-04 Yash Kulthe , Andrew Gilbert , John Collomosse

We devise a cascade GAN approach to generate talking face video, which is robust to different face shapes, view angles, facial characteristics, and noisy audio conditions. Instead of learning a direct mapping from audio to video frames, we…

计算机视觉与模式识别 · 计算机科学 2019-05-13 Lele Chen , Ross K. Maddox , Zhiyao Duan , Chenliang Xu

The generation of audio-driven talking head videos is a key challenge in computer vision and graphics, with applications in virtual avatars and digital media. Traditional approaches often struggle with capturing the complex interaction…

声音 · 计算机科学 2025-04-01 Ao Fu , Ziqi Ni , Yi Zhou

Video conferencing has caught much more attention recently. High fidelity and low bandwidth are two major objectives of video compression for video conferencing applications. Most pioneering methods rely on classic video compression codec…

计算机视觉与模式识别 · 计算机科学 2024-02-27 Yifei Li , Xiaohong Liu , Yicong Peng , Guangtao Zhai , Jun Zhou

Audio-guided face reenactment aims at generating photorealistic faces using audio information while maintaining the same facial movement as when speaking to a real person. However, existing methods can not generate vivid face images or only…

计算机视觉与模式识别 · 计算机科学 2020-05-01 Jiangning Zhang , Liang Liu , Zhucun Xue , Yong Liu

We propose a parametric model that maps free-view images into a vector space of coded facial shape, expression and appearance with a neural radiance field, namely Morphable Facial NeRF. Specifically, MoFaNeRF takes the coded facial shape,…

计算机视觉与模式识别 · 计算机科学 2022-07-25 Yiyu Zhuang , Hao Zhu , Xusen Sun , Xun Cao

This paper proposes a talking face generation method named "CP-EB" that takes an audio signal as input and a person image as reference, to synthesize a photo-realistic people talking video with head poses controlled by a short video clip…

计算机视觉与模式识别 · 计算机科学 2023-11-16 Jianzong Wang , Yimin Deng , Ziqi Liang , Xulong Zhang , Ning Cheng , Jing Xiao

Neural fields excel in computer vision and robotics due to their ability to understand the 3D visual world such as inferring semantics, geometry, and dynamics. Given the capabilities of neural fields in densely representing a 3D scene from…

计算机视觉与模式识别 · 计算机科学 2024-07-19 Muhammad Zubair Irshad , Sergey Zakharov , Vitor Guizilini , Adrien Gaidon , Zsolt Kira , Rares Ambrus

Audio-driven talking head generation requires precise synchronization between facial animations and audio signals. This paper introduces ATL-Diff, a novel approach addressing synchronization limitations while reducing noise and…

计算机视觉与模式识别 · 计算机科学 2025-07-18 Hoang-Son Vo , Quang-Vinh Nguyen , Seungwon Kim , Hyung-Jeong Yang , Soonja Yeom , Soo-Hyung Kim

We present dynamic neural radiance fields for modeling the appearance and dynamics of a human face. Digitally modeling and reconstructing a talking human is a key building-block for a variety of applications. Especially, for telepresence…

计算机视觉与模式识别 · 计算机科学 2020-12-08 Guy Gafni , Justus Thies , Michael Zollhöfer , Matthias Nießner

The advances in the Neural Radiance Fields (NeRF) research offer extensive applications in diverse domains, but protecting their copyrights has not yet been researched in depth. Recently, NeRF watermarking has been considered one of the…

计算机视觉与模式识别 · 计算机科学 2024-07-15 Youngdong Jang , Dong In Lee , MinHyuk Jang , Jong Wook Kim , Feng Yang , Sangpil Kim

In this paper, we propose a talking face generation method that takes an audio signal as input and a short target video clip as reference, and synthesizes a photo-realistic video of the target face with natural lip motions, head poses, and…

计算机视觉与模式识别 · 计算机科学 2021-08-19 Chenxu Zhang , Yifan Zhao , Yifei Huang , Ming Zeng , Saifeng Ni , Madhukar Budagavi , Xiaohu Guo

Generative Neural Radiance Fields (GNeRF)-based 3D-aware GANs have showcased remarkable prowess in crafting high-fidelity images while upholding robust 3D consistency, particularly face generation. However, specific existing models…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Jichao Zhang , Aliaksandr Siarohin , Yahui Liu , Hao Tang , Nicu Sebe , Wei Wang

The quality of three-dimensional reconstruction is a key factor affecting the effectiveness of its application in areas such as virtual reality (VR) and augmented reality (AR) technologies. Neural Radiance Fields (NeRF) can generate…

计算机视觉与模式识别 · 计算机科学 2023-06-09 Qianqiu Tan , Tao Liu , Yinling Xie , Shuwan Yu , Baohua Zhang

The problem of modeling an animatable 3D human head avatar under light-weight setups is of significant importance but has not been well solved. Existing 3D representations either perform well in the realism of portrait images synthesis or…

计算机视觉与模式识别 · 计算机科学 2023-10-02 Xiaochen Zhao , Lizhen Wang , Jingxiang Sun , Hongwen Zhang , Jinli Suo , Yebin Liu

Talking face synthesis has been widely studied in either appearance-based or warping-based methods. Previous works mostly utilize single face image as a source, and generate novel facial animations by merging other person's facial features.…

计算机视觉与模式识别 · 计算机科学 2019-11-22 Kuangxiao Gu , Yuqian Zhou , Thomas Huang

We introduce HyperFields, a method for generating text-conditioned Neural Radiance Fields (NeRFs) with a single forward pass and (optionally) some fine-tuning. Key to our approach are: (i) a dynamic hypernetwork, which learns a smooth…

计算机视觉与模式识别 · 计算机科学 2024-06-14 Sudarshan Babu , Richard Liu , Avery Zhou , Michael Maire , Greg Shakhnarovich , Rana Hanocka

Neural rendering techniques combining machine learning with geometric reasoning have arisen as one of the most promising approaches for synthesizing novel views of a scene from a sparse set of images. Among these, stands out the Neural…

计算机视觉与模式识别 · 计算机科学 2020-12-01 Albert Pumarola , Enric Corona , Gerard Pons-Moll , Francesc Moreno-Noguer

We present a novel semantic model for human head defined with neural radiance field. The 3D-consistent head model consist of a set of disentangled and interpretable bases, and can be driven by low-dimensional expression coefficients. Thanks…

图形学 · 计算机科学 2022-10-13 Xuan Gao , Chenglai Zhong , Jun Xiang , Yang Hong , Yudong Guo , Juyong Zhang

The demand for immersive and interactive communication has driven advancements in 3D video conferencing, yet achieving high-fidelity 3D talking face representation at low bitrates remains a challenge. Traditional 2D video compression…

计算机视觉与模式识别 · 计算机科学 2026-01-30 Jianglong Li , Jun Xu , Bingcong Lu , Zhengxue Cheng , Hongwei Hu , Ronghua Wu , Li Song