中文
相关论文

相关论文: TEASER: Token Enhanced Spatial Modeling for Expres…

200 篇论文

While existing methods for 3D face reconstruction from in-the-wild images excel at recovering the overall face shape, they commonly miss subtle, extreme, asymmetric, or rarely observed expressions. We improve upon these methods with SMIRK…

计算机视觉与模式识别 · 计算机科学 2025-03-14 George Retsinas , Panagiotis P. Filntisis , Radek Danecek , Victoria F. Abrevaya , Anastasios Roussos , Timo Bolkart , Petros Maragos

We address the problem of regressing 3D human pose and shape from a single image, with a focus on 3D accuracy. The current best methods leverage large datasets of 3D pseudo-ground-truth (p-GT) and 2D keypoints, leading to robust…

计算机视觉与模式识别 · 计算机科学 2024-04-26 Sai Kumar Dwivedi , Yu Sun , Priyanka Patel , Yao Feng , Michael J. Black

In this paper, a novel approach via embedded tensor manifold regularization for 2D+3D facial expression recognition (FERETMR) is proposed. Firstly, 3D tensors are constructed from 2D face images and 3D face shape models to keep the…

计算机视觉与模式识别 · 计算机科学 2022-02-01 Yunfang Fu , Qiuqi Ruan , Ziyan Luo , Gaoyun An , Yi Jin , Jun Wan

Facial expression editing methods can be mainly categorized into two types based on their architectures: 2D-based and 3D-based methods. The former lacks 3D face modeling capabilities, making it difficult to edit 3D factors effectively. The…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Yikang He , Jichao Zhang , Wei Wang , Nicu Sebe , Yao Zhao

In the field of medical images, although various works find Swin Transformer has promising effectiveness on pixelwise dense prediction, whether pre-training these models without using extra dataset can further boost the performance for the…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Xinrong Hu , Dewen Zeng , Yawen Wu , Xueyang Li , Yiyu Shi

High-fidelity head avatar reconstruction plays a crucial role in AR/VR, gaming, and multimedia content creation. Recent advances in 3D Gaussian Splatting (3DGS) have demonstrated effectiveness in modeling complex geometry with real-time…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Shikun Zhang , Cunjian Chen , Yiqun Wang , Qiuhong Ke , Yong Li

We introduce a self-supervised speech pre-training method called TERA, which stands for Transformer Encoder Representations from Alteration. Recent approaches often learn by using a single auxiliary task like contrastive prediction,…

音频与语音处理 · 电气工程与系统科学 2021-08-05 Andy T. Liu , Shang-Wen Li , Hung-yi Lee

This paper presents DENSER, an efficient and effective approach leveraging 3D Gaussian splatting (3DGS) for the reconstruction of dynamic urban environments. While several methods for photorealistic scene representations, both implicitly…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Mahmud A. Mohamad , Gamal Elghazaly , Arthur Hubert , Raphael Frank

Learned 3D representations of human faces are useful for computer vision problems such as 3D face tracking and reconstruction from images, as well as graphics applications such as character generation and animation. Traditional models learn…

计算机视觉与模式识别 · 计算机科学 2018-08-02 Anurag Ranjan , Timo Bolkart , Soubhik Sanyal , Michael J. Black

Micro-expression recognition can obtain the real emotion of the individual at the current moment. Although deep learning-based methods, especially Transformer-based methods, have achieved impressive results, these methods have high…

计算机视觉与模式识别 · 计算机科学 2026-04-10 Junbo Wang , Liangyu Fu , Yuke Li , Yining Zhu , Xuecheng Wu , Kun Hu

Three-dimensional medical image segmentation is a fundamental yet computationally demanding task due to the cubic growth of voxel processing and the redundant computation on homogeneous regions. To address these limitations, we propose…

计算机视觉与模式识别 · 计算机科学 2026-01-09 Sen Zeng , Hong Zhou , Zheng Zhu , Yang Liu

Visual generative models based on latent space have achieved great success, underscoring the significance of visual tokenization. Mapping images to latents boosts efficiency and enables multimodal alignment for scaling up in downstream…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Yunpeng Qu , Kaidong Zhang , Yukang Ding , Ying Chen , Jian Wang

This paper proposes a novel model fitting algorithm for 3D facial expression reconstruction from a single image. Face expression reconstruction from a single image is a challenging task in computer vision. Most state-of-the-art methods fit…

计算机视觉与模式识别 · 计算机科学 2018-08-20 Fanzi Wu , Songnan Li , Tianhao Zhao , King Ngi Ngan , Lv Sheng

Monocular facial performance capture in-the-wild is challenging due to varied capture conditions, face shapes, and expressions. Most current methods rely on linear 3D Morphable Models, which represent facial expressions independently of…

计算机视觉与模式识别 · 计算机科学 2025-07-14 Arthur Josi , Luiz Gustavo Hafemann , Abdallah Dib , Emeline Got , Rafael M. O. Cruz , Marc-Andre Carbonneau

In this work, we reveal the limitations of visual tokenizers and VAEs in preserving fine-grained features, and propose a benchmark to evaluate reconstruction performance for two challenging visual contents: text and face. Visual tokenizers…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Junfeng Wu , Dongliang Luo , Weizhi Zhao , Zhihao Xie , Yuanhao Wang , Junyi Li , Xudong Xie , Yuliang Liu , Xiang Bai

We present a novel end-to-end identity-agnostic face reenactment system, MaskRenderer, that can generate realistic, high fidelity frames in real-time. Although recent face reenactment works have shown promising results, there are still…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Tina Behrouzi , Atefeh Shahroudnejad , Payam Mousavi

In this paper, we propose $\tau$GAN a tensor-based method for modeling the latent space of generative models. The objective is to identify semantic directions in latent space. To this end, we propose to fit a multilinear tensor model on a…

计算机视觉与模式识别 · 计算机科学 2021-11-09 René Haas , Stella Graßhof , Sami Sebastian Brandt

Realistic, high-fidelity 3D facial animations are crucial for expressive avatar systems in human-computer interaction and accessibility. Although prior methods show promising quality, their reliance on the mesh domain limits their ability…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Alexandre Symeonidis-Herzig , Özge Mercanoğlu Sincan , Richard Bowden

We present SIDER(Single-Image neural optimization for facial geometric DEtail Recovery), a novel photometric optimization method that recovers detailed facial geometry from a single image in an unsupervised manner. Inspired by classical…

计算机视觉与模式识别 · 计算机科学 2021-08-13 Aggelina Chatziagapi , ShahRukh Athar , Francesc Moreno-Noguer , Dimitris Samaras

Recent multimodal models for instruction-based face editing enable semantic manipulation but still struggle with precise attribute control and identity preservation. Structural facial representations such as landmarks are effective for…

计算机视觉与模式识别 · 计算机科学 2026-01-30 Zhenghao Zhang , Ziying Zhang , Junchao Liao , Xiangyu Meng , Qiang Hu , Siyu Zhu , Xiaoyun Zhang , Long Qin , Weizhi Wang
‹ 上一页 1 2 3 10 下一页 ›