中文
相关论文

相关论文: LUCAS: Layered Universal Codec Avatars

200 篇论文

Generating realistic human 3D reconstructions using image or video data is essential for various communication and entertainment applications. While existing methods achieved impressive results for body and facial regions, realistic hair…

计算机视觉与模式识别 · 计算机科学 2023-06-13 Vanessa Sklyarova , Jenya Chelishev , Andreea Dogaru , Igor Medvedev , Victor Lempitsky , Egor Zakharov

Reconstructing an avatar from a portrait image has many applications in multimedia, but remains a challenging research problem. Extracting reflectance maps and geometry from one image is ill-posed: recovering geometry is a one-to-many…

计算机视觉与模式识别 · 计算机科学 2023-12-25 Abdallah Dib , Luiz Gustavo Hafemann , Emeline Got , Trevor Anderson , Amin Fadaeinejad , Rafael M. O. Cruz , Marc-Andre Carbonneau

There is a growing demand for the accessible creation of high-quality 3D avatars that are animatable and customizable. Although 3D morphable models provide intuitive control for editing and animation, and robustness for single-view face…

计算机视觉与模式识别 · 计算机科学 2023-05-05 Connor Z. Lin , Koki Nagano , Jan Kautz , Eric R. Chan , Umar Iqbal , Leonidas Guibas , Gordon Wetzstein , Sameh Khamis

Unified multimodal models (UMMs) unify visual understanding and generation within a single architecture. However, conventional training relies on image-text pairs (or sequences) whose captions are typically sparse and miss fine-grained…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Ji Xie , Trevor Darrell , Luke Zettlemoyer , XuDong Wang

Facial expression and hand motions are necessary to express our emotions and interact with the world. Nevertheless, most of the 3D human avatars modeled from a casually captured video only support body motions without facial expressions and…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Gyeongsik Moon , Takaaki Shiratori , Shunsuke Saito

In this paper, we explore a reconstruction and reenactment separated framework for 3D Gaussians head, which requires only a single portrait image as input to generate controllable avatar. Specifically, we developed a large-scale one-shot…

计算机视觉与模式识别 · 计算机科学 2025-09-18 Zhiling Ye , Cong Zhou , Xiubao Zhang , Haifeng Shen , Weihong Deng , Quan Lu

We propose a new "Unbiased through Textual Description (UTD)" video benchmark based on unbiased subsets of existing video classification and retrieval datasets to enable a more robust assessment of video understanding capabilities. Namely,…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Nina Shvetsova , Arsha Nagrani , Bernt Schiele , Hilde Kuehne , Christian Rupprecht

Creating high-fidelity, animatable 3D talking heads is crucial for immersive applications, yet often hindered by the prevalence of low-quality image or video sources, which yield poor 3D reconstructions. In this paper, we introduce…

计算机视觉与模式识别 · 计算机科学 2026-02-09 Ding-Jiun Huang , Yuanhao Wang , Shao-Ji Yuan , Albert Mosella-Montoro , Francisco Vicente Carrasco , Cheng Zhang , Fernando De la Torre

Learned 3D representations of human faces are useful for computer vision problems such as 3D face tracking and reconstruction from images, as well as graphics applications such as character generation and animation. Traditional models learn…

计算机视觉与模式识别 · 计算机科学 2018-08-02 Anurag Ranjan , Timo Bolkart , Soubhik Sanyal , Michael J. Black

The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing the users idiosyncrasies and detailed facial expressions. However, current 3D head avatar…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Jalees Nehvi , Timo Bolkart , Thabo Beeler , Justus Thies

We introduce a new hair modeling method that uses a dual representation of classical hair strands and 3D Gaussians to produce accurate and realistic strand-based reconstructions from multi-view data. In contrast to recent approaches that…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Egor Zakharov , Vanessa Sklyarova , Michael Black , Giljoo Nam , Justus Thies , Otmar Hilliges

Three-dimensional Morphable Models (3DMMs) are powerful statistical tools for representing the 3D shapes and textures of an object class. Here we present the most complete 3DMM of the human head to date that includes face, cranium, ears,…

VR telepresence consists of interacting with another human in a virtual space represented by an avatar. Today most avatars are cartoon-like, but soon the technology will allow video-realistic ones. This paper aims in this direction and…

计算机视觉与模式识别 · 计算机科学 2020-08-28 Hang Chu , Shugao Ma , Fernando De la Torre , Sanja Fidler , Yaser Sheikh

Recent studies have highlighted the interplay between diffusion models and representation learning. Intermediate representations from diffusion models can be leveraged for downstream visual tasks, while self-supervised vision models can…

计算机视觉与模式识别 · 计算机科学 2025-07-23 Xiangxiang Chu , Renda Li , Yong Wang

Photorealistic and animatable human avatars are a key enabler for virtual/augmented reality, telepresence, and digital entertainment. While recent advances in 3D Gaussian Splatting (3DGS) have greatly improved rendering quality and…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Cheng Peng , Jingxiang Sun , Yushuo Chen , Zhaoqi Su , Zhuo Su , Yebin Liu

The Facial Action Coding System (FACS) has been used by numerous studies to investigate the links between facial behavior and mental health. The laborious and costly process of FACS coding has motivated the development of machine learning…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Evangelos Sariyanidi , Lisa Yankowitz , Robert T. Schultz , John D. Herrington , Birkan Tunc , Jeffrey Cohn

Curvilinear structure segmentation (CSS) is essential in various domains, including medical imaging, landscape analysis, industrial surface inspection, and plant analysis. While existing methods achieve high performance within specific…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Kai Zhu , Li Chen , Dianshuo Li , Yunxiang Cao , Jun Cheng

We introduce 3D Gaussian blendshapes for modeling photorealistic head avatars. Taking a monocular video as input, we learn a base head model of neutral expression, along with a group of expression blendshapes, each of which corresponds to a…

图形学 · 计算机科学 2024-05-03 Shengjie Ma , Yanlin Weng , Tianjia Shao , Kun Zhou

Unprecedented visual details of biological structures are being revealed by subcellular-resolution whole-brain 3D microscopy data, enabled by recent advances in intact tissue processing and light-sheet fluorescence microscopy (LSFM). These…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Minyoung E. Kim , Dae Hee Yun , Aditi V. Patel , Madeline Hon , Webster Guan , Taegeon Lee , Brian Nguyen

Recent large reconstruction models have made notable progress in generating high-quality 3D objects from single images. However, current reconstruction methods often rely on explicit camera pose estimation or fixed viewpoints, restricting…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Hao He , Yixun Liang , Luozhou Wang , Yuanhao Cai , Xinli Xu , Hao-Xiang Guo , Xiang Wen , Yingcong Chen