中文
相关论文

相关论文: Text2Face: A Multi-Modal 3D Face Model

200 篇论文

With the rapid advancement of intelligent transportation systems, text-driven image generation and editing techniques have demonstrated significant potential in providing rich, controllable visual scene data for applications such as traffic…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Feng Lv , Haoxuan Feng , Zilu Zhang , Chunlong Xia , Yanfeng Li

Face recognition now requires a large number of labelled masked face images in the era of this unprecedented COVID-19 pandemic. Unfortunately, the rapid spread of the virus has left us little time to prepare for such dataset in the wild. To…

计算机视觉与模式识别 · 计算机科学 2021-08-03 Je Hyeong Hong , Hanjo Kim , Minsoo Kim , Gi Pyo Nam , Junghyun Cho , Hyeong-Seok Ko , Ig-Jae Kim

While current talking head models are capable of generating photorealistic talking head videos, they provide limited pose controllability. Most methods require specific video sequences that should exactly contain the head pose desired,…

计算机视觉与模式识别 · 计算机科学 2023-05-19 Kwangho Lee , Patrick Kwon , Myung Ki Lee , Namhyuk Ahn , Junsoo Lee

Text-based video segmentation aims to segment the target object in a video based on a describing sentence. Incorporating motion information from optical flow maps with appearance and linguistic modalities is crucial yet has been largely…

计算机视觉与模式识别 · 计算机科学 2022-04-07 Wangbo Zhao , Kai Wang , Xiangxiang Chu , Fuzhao Xue , Xinchao Wang , Yang You

In this paper, we show how a 3D Morphable Model (i.e. a statistical model of the 3D shape of a class of objects such as faces) can be used to spatially transform input data as a module (a 3DMM-STN) within a convolutional neural network.…

计算机视觉与模式识别 · 计算机科学 2018-04-20 Anil Bas , Patrik Huber , William A. P. Smith , Muhammad Awais , Josef Kittler

3D Morphable Models are a class of generative models commonly used to model faces. They are typically applied to ill-posed problems such as 3D reconstruction from 2D data. Several ambiguities in this problem's image formation process have…

计算机视觉与模式识别 · 计算机科学 2021-09-30 Bernhard Egger , Skylar Sutherland , Safa C. Medin , Joshua Tenenbaum

Morphing attack detection has become an essential component of face recognition systems for ensuring a reliable verification scenario. In this paper, we present a multimodal learning approach that can provide a textual description of…

计算机视觉与模式识别 · 计算机科学 2025-08-15 Sushrut Patwardhan , Raghavendra Ramachandra , Sushma Venkatesh

3D Morphable Models (3DMMs) demonstrate great potential for reconstructing faithful and animatable 3D facial surfaces from a single image. The facial surface is influenced by the coarse shape, as well as the static detail (e,g.,…

计算机视觉与模式识别 · 计算机科学 2023-08-24 Zenghao Chai , Tianke Zhang , Tianyu He , Xu Tan , Tadas Baltrušaitis , HsiangTao Wu , Runnan Li , Sheng Zhao , Chun Yuan , Jiang Bian

Face analysis techniques have become a crucial component of human-machine interaction in the fields of assistive and humanoid robotics. However, the variations in head-pose that arise naturally in these environments are still a great…

计算机视觉与模式识别 · 计算机科学 2016-06-03 Michael Grupp , Philipp Kopp , Patrik Huber , Matthias Rätsch

Traditional methods for image-based 3D face reconstruction and facial motion retargeting fit a 3D morphable model (3DMM) to the face, which has limited modeling capacity and fail to generalize well to in-the-wild data. Use of deformation…

计算机视觉与模式识别 · 计算机科学 2020-07-21 Bindita Chaudhuri , Noranart Vesdapunt , Linda Shapiro , Baoyuan Wang

The ability to generate diverse 3D articulated head avatars is vital to a plethora of applications, including augmented reality, cinematography, and education. Recent work on text-guided 3D object generation has shown great promise in…

计算机视觉与模式识别 · 计算机科学 2023-07-12 Alexander W. Bergman , Wang Yifan , Gordon Wetzstein

We present ShapeCrafter, a neural network for recursive text-conditioned 3D shape generation. Existing methods to generate text-conditioned 3D shapes consume an entire text prompt to generate a 3D shape in a single step. However, humans…

计算机视觉与模式识别 · 计算机科学 2023-04-11 Rao Fu , Xiao Zhan , Yiwen Chen , Daniel Ritchie , Srinath Sridhar

Human motion naturally integrates body movements and facial expressions, forming a unified perception. If a virtual character's facial expression does not align well with its body movements, it may weaken the perception of the character as…

图形学 · 计算机科学 2025-11-19 Bokyung Jang , Eunho Jung , Yoonsang Lee

Text-to-Face (TTF) synthesis is a challenging task with great potential for diverse computer vision applications. Compared to Text-to-Image (TTI) synthesis tasks, the textual description of faces can be much more complicated and detailed…

计算机视觉与模式识别 · 计算机科学 2020-09-21 Tianren Wang , Teng Zhang , Brian Lovell

We present a new multi-modal face image generation method that converts a text prompt and a visual input, such as a semantic mask or scribble map, into a photo-realistic face image. To do this, we combine the strengths of Generative…

计算机视觉与模式识别 · 计算机科学 2024-05-08 Jihyun Kim , Changjae Oh , Hoseok Do , Soohyun Kim , Kwanghoon Sohn

3D facial animation is often produced by manipulating facial deformation models (or rigs), that are traditionally parameterized by expression controls. A key component that is usually overlooked is expression 'style', as in, how a…

计算机视觉与模式识别 · 计算机科学 2024-01-30 Lingchen Yang , Gaspard Zoss , Prashanth Chandran , Paulo Gotardo , Markus Gross , Barbara Solenthaler , Eftychios Sifakis , Derek Bradley

The ability to provide fine-grained control for generating and editing visual imagery has profound implications for computer vision and its applications. Previous works have explored extending controllability in two directions: instruction…

计算机视觉与模式识别 · 计算机科学 2024-10-18 Shufan Li , Harkanwar Singh , Aditya Grover

Embedding 3D morphable basis functions into deep neural networks opens great potential for models with better representation power. However, to faithfully learn those models from an image collection, it requires strong regularization to…

计算机视觉与模式识别 · 计算机科学 2019-04-11 Luan Tran , Feng Liu , Xiaoming Liu

In facial image generation, current text-to-image models often suffer from facial attribute leakage and insufficient physical consistency when responding to local semantic instructions. In this study, we propose Face-MakeUpV2, a facial…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Dawei Dai , Yinxiu Zhou , Chenghang Li , Guolai Jiang , Chengfang Zhang

We propose a method for constructing generative models of 3D objects from a single 3D mesh and improving them through unsupervised low-shot learning from 2D images. Our method produces a 3D morphable model that represents shape and albedo…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Skylar Sutherland , Bernhard Egger , Joshua Tenenbaum