English
Related papers

Related papers: FLARE: Fast Learning of Animatable and Relightable…

200 papers

We introduce FLARE, a family of vision language models (VLMs) with a fully vision-language alignment and integration paradigm. Unlike existing approaches that rely on single MLP projectors for modality alignment and defer cross-modal…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Zheng Liu , Mengjie Liu , Jingzhou Chen , Jingwei Xu , Bin Cui , Conghui He , Wentao Zhang

We introduce AvatarForge, a framework for generating animatable 3D human avatars from text or image inputs using AI-driven procedural generation. While diffusion-based methods have made strides in general 3D object generation, they struggle…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Xinhang Liu , Yu-Wing Tai , Chi-Keung Tang

Creating high-quality controllable 3D human models from multi-view RGB videos poses a significant challenge. Neural radiance fields (NeRFs) have demonstrated remarkable quality in reconstructing and free-viewpoint rendering of static as…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Paul Knoll , Wieland Morgenstern , Anna Hilsmann , Peter Eisert

We present HARP (HAnd Reconstruction and Personalization), a personalized hand avatar creation approach that takes a short monocular RGB video of a human hand as input and reconstructs a faithful hand avatar exhibiting a high-fidelity…

Computer Vision and Pattern Recognition · Computer Science 2023-07-06 Korrawe Karunratanakul , Sergey Prokudin , Otmar Hilliges , Siyu Tang

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

We present the Locally Adaptive Morphable Model (LAMM), a highly flexible Auto-Encoder (AE) framework for learning to generate and manipulate 3D meshes. We train our architecture following a simple self-supervised training scheme in which…

Computer Vision and Pattern Recognition · Computer Science 2024-01-08 Michail Tarasiou , Rolandos Alexandros Potamias , Eimear O'Sullivan , Stylianos Ploumpis , Stefanos Zafeiriou

Modeling animatable human avatars from monocular or multi-view videos has been widely studied, with recent approaches leveraging neural radiance fields (NeRFs) or 3D Gaussian Splatting (3DGS) achieving impressive results in novel-view and…

Computer Vision and Pattern Recognition · Computer Science 2025-04-03 Yahui Li , Zhi Zeng , Liming Pang , Guixuan Zhang , Shuwu Zhang

Surfaces are typically represented as meshes, which can be extracted from volumetric fields via meshing or optimized directly as surface parameterizations. Volumetric representations occupy 3D space and have a large effective receptive…

Graphics · Computer Science 2026-02-03 Ruiqi Zhang , Jiacheng Wu , Jie Chen

Recovering photorealistic and drivable full-body avatars is crucial for numerous applications, including virtual reality, 3D games, and tele-presence. Most methods, whether reconstruction or generation, require large numbers of human motion…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Yujiao Jiang , Qingmin Liao , Zhaolong Wang , Xiangru Lin , Zongqing Lu , Yuxi Zhao , Hanqing Wei , Jingrui Ye , Yu Zhang , Zhijing Shao

Over the last years, many face analysis tasks have accomplished astounding performance, with applications including face generation and 3D face reconstruction from a single "in-the-wild" image. Nevertheless, to the best of our knowledge,…

Computer Vision and Pattern Recognition · Computer Science 2021-12-14 Alexandros Lattas , Stylianos Moschoglou , Stylianos Ploumpis , Baris Gecer , Abhijeet Ghosh , Stefanos Zafeiriou

We present a system for realistic one-shot mesh-based human head avatars creation, ROME for short. Using a single photograph, our model estimates a person-specific head mesh and the associated neural texture, which encodes both local…

Computer Vision and Pattern Recognition · Computer Science 2022-06-17 Taras Khakhulin , Vanessa Sklyarova , Victor Lempitsky , Egor Zakharov

Fixed-length fingerprint representations, which map each fingerprint to a compact and fixed-size feature vector, are computationally efficient and well-suited for large-scale matching. However, designing a robust representation that…

Computer Vision and Pattern Recognition · Computer Science 2026-05-07 Zhiyu Pan , Xiongjun Guan , Yongjie Duan , Jianjiang Feng , Jie Zhou

This paper presents Neural Mesh Fusion (NMF), an efficient approach for joint optimization of polygon mesh from multi-view image observations and unsupervised 3D planar-surface parsing of the scene. In contrast to implicit neural…

Computer Vision and Pattern Recognition · Computer Science 2024-02-27 Farhad G. Zanjani , Hong Cai , Yinhao Zhu , Leyla Mirvakhabova , Fatih Porikli

High-quality textures are critical for realistic 3D content creation, yet existing generative methods are slow, rely on UV maps, and often fail to remain faithful to a reference image. To address these challenges, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2025-09-08 Arianna Rampini , Kanika Madan , Bruno Roy , AmirHossein Zamani , Derek Cheung

Visually exploring in a real-world 4D spatiotemporal space freely in VR has been a long-term quest. The task is especially appealing when only a few or even single RGB cameras are used for capturing the dynamic scene. To this end, we…

Computer Vision and Pattern Recognition · Computer Science 2023-02-21 Liangchen Song , Anpei Chen , Zhong Li , Zhang Chen , Lele Chen , Junsong Yuan , Yi Xu , Andreas Geiger

Traditional lecture videos offer flexibility but lack mechanisms for real-time clarification, forcing learners to search externally when confusion arises. Recent advances in large language models and neural avatars provide new opportunities…

Computer Vision and Pattern Recognition · Computer Science 2025-12-25 Md Zabirul Islam , Md Motaleb Hossen Manik , Ge Wang

3D characters are essential to modern creative industries, but making them animatable often demands extensive manual work in tasks like rigging and skinning. Existing automatic rigging tools face several limitations, including the necessity…

Graphics · Computer Science 2025-03-12 Zhiyang Guo , Jinxu Xiang , Kai Ma , Wengang Zhou , Houqiang Li , Ran Zhang

Creating high-fidelity 3D meshes with arbitrary topology, including open surfaces and complex interiors, remains a significant challenge. Existing implicit field methods often require costly and detail-degrading watertight conversion, while…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Xianglong He , Zi-Xin Zou , Chia-Hao Chen , Yuan-Chen Guo , Ding Liang , Chun Yuan , Wanli Ouyang , Yan-Pei Cao , Yangguang Li

With the booming of virtual reality (VR) technology, there is a growing need for customized 3D avatars. However, traditional methods for 3D avatar modeling are either time-consuming or fail to retain similarity to the person being modeled.…

Computer Vision and Pattern Recognition · Computer Science 2023-07-06 Chuanyu Pan , Guowei Yang , Taijiang Mu , Yu-Kun Lai

This paper propose iFlame, a novel transformer-based network architecture for mesh generation. While attention-based models have demonstrated remarkable performance in mesh generation, their quadratic computational complexity limits…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Hanxiao Wang , Biao Zhang , Weize Quan , Dong-Ming Yan , Peter Wonka