中文
相关论文

相关论文: TEASER: Token Enhanced Spatial Modeling for Expres…

200 篇论文

Recently, token-based generation have demonstrated their effectiveness in image synthesis. As a representative example, non-autoregressive Transformers (NATs) can generate decent-quality images in a few steps. NATs perform generation in a…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Zanlin Ni , Yulin Wang , Renping Zhou , Yizeng Han , Jiayi Guo , Zhiyuan Liu , Yuan Yao , Gao Huang

We introduce a highly robust GAN-based framework for digitizing a normalized 3D avatar of a person from a single unconstrained photo. While the input image can be of a smiling person or taken in extreme lighting conditions, our method can…

计算机视觉与模式识别 · 计算机科学 2021-06-23 Huiwen Luo , Koki Nagano , Han-Wei Kung , Mclean Goldwhite , Qingguo Xu , Zejian Wang , Lingyu Wei , Liwen Hu , Hao Li

Drawing upon StyleGAN's expressivity and disentangled latent space, existing 2D approaches employ textual prompting to edit facial images with different attributes. In contrast, 3D-aware approaches that generate faces at different target…

计算机视觉与模式识别 · 计算机科学 2024-07-25 Amandeep Kumar , Muhammad Awais , Sanath Narayan , Hisham Cholakkal , Salman Khan , Rao Muhammad Anwer

In this paper, we present a large-scale detailed 3D face dataset, FaceScape, and the corresponding benchmark to evaluate single-view facial 3D reconstruction. By training on FaceScape data, a novel algorithm is proposed to predict elaborate…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Hao Zhu , Haotian Yang , Longwei Guo , Yidi Zhang , Yanru Wang , Mingkai Huang , Menghua Wu , Qiu Shen , Ruigang Yang , Xun Cao

Facial alignment involves finding a set of landmark points on an image with a known semantic meaning. However, this semantic meaning of landmark points is often lost in 2D approaches where landmarks are either moved to visible boundaries or…

计算机视觉与模式识别 · 计算机科学 2017-09-11 Chandrasekhar Bhagavatula , Chenchen Zhu , Khoa Luu , Marios Savvides

This paper proposes an encoder-decoder network to disentangle shape features during 3D face reconstruction from single 2D images, such that the tasks of reconstructing accurate 3D face shapes and learning discriminative shape features for…

计算机视觉与模式识别 · 计算机科学 2018-04-02 Feng Liu , Ronghang Zhu , Dan Zeng , Qijun Zhao , Xiaoming Liu

Recent advances in deep generative models have demonstrated impressive results in photo-realistic facial image synthesis and editing. Facial expressions are inherently the result of muscle movement. However, existing neural network-based…

计算机视觉与模式识别 · 计算机科学 2019-11-07 ShahRukh Athar , Zhixin Shu , Dimitris Samaras

Joint reconstruction of 3D human and object from a single image is an active research area, with pivotal applications in robotics and digital content creation. Despite recent advances, existing approaches suffer from two fundamental…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Hyeongjin Nam , Daniel Sungho Jung , Kyoung Mu Lee

Urban environments are continuously mapped and modeled by various data collection platforms, including satellites, unmanned aerial vehicles and street cameras. The growing availability of 3D geospatial data from multiple modalities has…

机器学习 · 计算机科学 2025-11-11 Bar Genossar , Sagi Dalyot , Roee Shraga , Avigdor Gal

In this paper, we present a novel approach to automatic 3D Facial Expression Recognition (FER) based on deep representation of facial 3D geometric and 2D photometric attributes. A 3D face is firstly represented by its geometric and…

计算机视觉与模式识别 · 计算机科学 2015-11-11 Huibin Li , Jian Sun , Dong Wang , Zongben Xu , Liming Chen

Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domains. In the 3D realm, prevailing approaches typically rely on a cascade pipeline: first…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Hongyu Yan , Kunming Luo , Weiyu Li , Kaiyi Zhang , Yixun Liang , Jingwei Huang , Chunchao Guo , Ping Tan

Meaningful facial parts can convey key cues for both facial action unit detection and expression prediction. Textured 3D face scan can provide both detailed 3D geometric shape and 2D texture appearance cues of the face which are beneficial…

计算机视觉与模式识别 · 计算机科学 2018-03-16 Asim Jan , Huaxiong Ding , Hongying Meng , Liming Chen , Huibin Li

Existing facial expression recognition (FER) methods typically fine-tune a pre-trained visual encoder using discrete labels. However, this form of supervision limits to specify the emotional concept of different facial expressions. In this…

计算机视觉与模式识别 · 计算机科学 2024-09-16 Hangyu Li , Yihan Xu , Jiangchao Yao , Nannan Wang , Xinbo Gao , Bo Han

We present THUNDR, a transformer-based deep neural network methodology to reconstruct the 3d pose and shape of people, given monocular RGB images. Key to our methodology is an intermediate 3d marker representation, where we aim to combine…

计算机视觉与模式识别 · 计算机科学 2021-06-18 Mihai Zanfir , Andrei Zanfir , Eduard Gabriel Bazavan , William T. Freeman , Rahul Sukthankar , Cristian Sminchisescu

Photo-realistic and controllable 3D avatars are crucial for various applications such as virtual and mixed reality (VR/MR), telepresence, gaming, and film production. Traditional methods for avatar creation often involve time-consuming…

Even though 3D face reconstruction has achieved impressive progress, most orthogonal projection-based face reconstruction methods can not achieve accurate and consistent reconstruction results when the face is very close to the camera due…

计算机视觉与模式识别 · 计算机科学 2022-08-16 Jia Guo , Jinke Yu , Alexandros Lattas , Jiankang Deng

We present Image2GS, a novel approach that addresses the challenging problem of reconstructing photorealistic 3D scenes from a single image by focusing specifically on the image-to-3D lifting component of the reconstruction process. By…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Tianshi Cao , Marie-Julie Rakotosaona , Ben Poole , Federico Tombari , Michael Niemeyer

Emotions play a central role in the social life of every human being, and their study, which represents a multidisciplinary subject, embraces a great variety of research fields. Especially concerning the latter, the analysis of facial…

计算机视觉与模式识别 · 计算机科学 2021-05-07 Fabio Valerio Massoli , Donato Cafarelli , Claudio Gennaro , Giuseppe Amato , Fabrizio Falchi

In autoregressive (AR) image generation, visual tokenizers compress images into compact discrete latent tokens, enabling efficient training of downstream autoregressive models for visual generation via next-token prediction. While scaling…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Tianwei Xiong , Jun Hao Liew , Zilong Huang , Jiashi Feng , Xihui Liu

Previous methods for dynamic facial expression recognition (DFER) in the wild are mainly based on Convolutional Neural Networks (CNNs), whose local operations ignore the long-range dependencies in videos. Transformer-based methods for DFER…

计算机视觉与模式识别 · 计算机科学 2023-05-08 Fuyan Ma , Bin Sun , Shutao Li