中文
相关论文

相关论文: FitDiff: Robust monocular 3D facial shape and refl…

200 篇论文

Diffusion models are a special type of generative model, capable of synthesising new data from a learnt distribution. We introduce DISPR, a diffusion-based model for solving the inverse problem of three-dimensional (3D) cell shape…

计算机视觉与模式识别 · 计算机科学 2023-03-15 Dominik J. E. Waibel , Ernst Röell , Bastian Rieck , Raja Giryes , Carsten Marr

Creating human avatars is a highly desirable yet challenging task. Recent advancements in radiance field rendering have achieved unprecedented photorealism and real-time performance for personalized dynamic human avatars. However, these…

图形学 · 计算机科学 2025-09-09 Dongliang Cao , Guoxing Sun , Marc Habermann , Florian Bernard

Over the last years, many face analysis tasks have accomplished astounding performance, with applications including face generation and 3D face reconstruction from a single "in-the-wild" image. Nevertheless, to the best of our knowledge,…

计算机视觉与模式识别 · 计算机科学 2021-12-14 Alexandros Lattas , Stylianos Moschoglou , Stylianos Ploumpis , Baris Gecer , Abhijeet Ghosh , Stefanos Zafeiriou

We present a novel framework for generating photorealistic 3D human head and subsequently manipulating and reposing them with remarkable flexibility. The proposed approach leverages an implicit function representation of 3D human heads,…

计算机视觉与模式识别 · 计算机科学 2023-12-22 Yushi Lan , Feitong Tan , Di Qiu , Qiangeng Xu , Kyle Genova , Zeng Huang , Sean Fanello , Rohit Pandey , Thomas Funkhouser , Chen Change Loy , Yinda Zhang

Face aging is the process of converting an individual's appearance to a younger or older version of themselves. Existing face aging techniques have been limited to 2D settings, which often weaken their applications as there is a growing…

计算机视觉与模式识别 · 计算机科学 2024-08-29 Junaid Wahid , Fangneng Zhan , Pramod Rao , Christian Theobalt

With the rapid development of imaging sensor technology in the field of remote sensing, multi-modal remote sensing data fusion has emerged as a crucial research direction for land cover classification tasks. While diffusion models have made…

计算机视觉与模式识别 · 计算机科学 2024-01-08 DaiXun Li , Weiying Xie , ZiXuan Wang , YiBing Lu , Yunsong Li , Leyuan Fang

Diffusion models have achieved significant success in both natural image and medical image domains, encompassing a wide range of applications. Previous investigations in medical images have often been constrained to specific anatomical…

计算机视觉与模式识别 · 计算机科学 2025-12-08 Yongrui Yu , Yannian Gu , Shaoting Zhang , Xiaofan Zhang

Facial attribute editing and style manipulation are crucial for applications like virtual avatars and photo editing. However, achieving precise control over facial attributes without altering unrelated features is challenging due to the…

计算机视觉与模式识别 · 计算机科学 2026-04-24 Wenmin Huang , Weiqi Luo , Xiaochun Cao , Jiwu Huang

Facial attribute editing aims to modify target attributes while preserving attribute-irrelevant content and overall image fidelity. Existing GAN-based methods provide favorable controllability, but often suffer from weak alignment between…

计算机视觉与模式识别 · 计算机科学 2026-04-24 Wenmin Huang , Weiqi Luo , Xiaochun Cao , Jiwu Huang

We present 3DiFACE, a novel method for personalized speech-driven 3D facial animation and editing. While existing methods deterministically predict facial animations from speech, they overlook the inherent one-to-many relationship between…

计算机视觉与模式识别 · 计算机科学 2023-12-05 Balamurugan Thambiraja , Sadegh Aliakbarian , Darren Cosker , Justus Thies

In this paper, we propose a diffusion-based face swapping framework for the first time, called DiffFace, composed of training ID conditional DDPM, sampling with facial guidance, and a target-preserving blending. In specific, in the training…

计算机视觉与模式识别 · 计算机科学 2022-12-29 Kihong Kim , Yunho Kim , Seokju Cho , Junyoung Seo , Jisu Nam , Kychul Lee , Seungryong Kim , KwangHee Lee

Deep MRI reconstruction is commonly performed with conditional models that de-alias undersampled acquisitions to recover images consistent with fully-sampled data. Since conditional models are trained with knowledge of the imaging operator,…

图像与视频处理 · 电气工程与系统科学 2023-09-19 Alper Güngör , Salman UH Dar , Şaban Öztürk , Yilmaz Korkmaz , Gokberk Elmas , Muzaffer Özbey , Tolga Çukur

Diffusion models have demonstrated impressive performance in face restoration. Yet, their multi-step inference process remains computationally intensive, limiting their applicability in real-world scenarios. Moreover, existing methods often…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Jingkai Wang , Jue Gong , Lin Zhang , Zheng Chen , Xing Liu , Hong Gu , Yutong Liu , Yulun Zhang , Xiaokang Yang

Talking head generation is a significant research topic that still faces numerous challenges. Previous works often adopt generative adversarial networks or regression models, which are plagued by generation quality and average facial shape…

计算机视觉与模式识别 · 计算机科学 2024-08-20 Ziyu Yao , Xuxin Cheng , Zhiqi Huang

Image generative models, particularly diffusion-based models, have surged in popularity due to their remarkable ability to synthesize highly realistic images. However, since these models are data-driven, they inherit biases from the…

机器学习 · 计算机科学 2025-03-18 Lin-Chun Huang , Ching Chieh Tsao , Fang-Yi Su , Jung-Hsien Chiang

With the increasing deployment of facial image data across a wide range of applications, efficient compression tailored to facial semantics has become critical for both storage and transmission. While recent learning-based face image…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Yimin Zhou , Yichong Xia , Bin Chen , Mingyao Hong , Jiawei Li , Zhi Wang , Yaowei Wang

Existing methods for image-to-3D avatar generation struggle to produce highly detailed, animation-ready avatars suitable for real-world applications. We introduce AdaHuman, a novel framework that generates high-fidelity animatable 3D…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Yangyi Huang , Ye Yuan , Xueting Li , Jan Kautz , Umar Iqbal

Speech-driven 3D facial animation seeks to produce lifelike facial expressions that are synchronized with the speech content and its emotional nuances, finding applications in various multimedia fields. However, previous methods often…

计算机视觉与模式识别 · 计算机科学 2025-03-17 Yixuan Zhang , Qing Chang , Yuxi Wang , Guang Chen , Zhaoxiang Zhang , Junran Peng

Different forms of customized 2D avatars are widely used in gaming applications, virtual communication, education, and content creation. However, existing approaches often fail to capture fine-grained facial expressions and struggle to…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Hao Yu , Rupayan Mallick , Margrit Betke , Sarah Adel Bargal

We present PrimDiffusion, the first diffusion-based framework for 3D human generation. Devising diffusion models for 3D human generation is difficult due to the intensive computational cost of 3D representations and the articulated topology…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Zhaoxi Chen , Fangzhou Hong , Haiyi Mei , Guangcong Wang , Lei Yang , Ziwei Liu