中文
相关论文

相关论文: InstantAvatar: Learning Avatars from Monocular Vid…

200 篇论文

Rendering moving human bodies at free viewpoints only from a monocular video is quite a challenging problem. The information is too sparse to model complicated human body structures and motions from both view and pose dimensions. Neural…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Taoran Yi , Jiemin Fang , Xinggang Wang , Wenyu Liu

Despite progress in human motion capture, existing multi-view methods often face challenges in estimating the 3D pose and shape of multiple closely interacting people. This difficulty arises from reliance on accurate 2D joint estimations,…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Feichi Lu , Zijian Dong , Jie Song , Otmar Hilliges

We present LiftAvatar, a new paradigm that completes sparse monocular observations in kinematic space (e.g., facial expressions and head pose) and uses the completed signals to drive high-fidelity avatar animation. LiftAvatar is a…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Hualiang Wei , Shunran Jia , Jialun Liu , Wenhui Li

Understanding articulated objects from monocular video is a crucial yet challenging task in robotics and digital twin creation. Existing methods often rely on complex multi-view setups, high-fidelity object scans, or fragile long-term point…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Arslan Artykov , Tom Ravaud , Corentin Sautier , Vincent Lepetit

We introduce a new method that generates photo-realistic humans under novel views and poses given a monocular video as input. Despite the significant progress recently on this topic, with several methods exploring shared canonical neural…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Tiantian Wang , Nikolaos Sarafianos , Ming-Hsuan Yang , Tony Tung

Creating photorealistic avatars for individuals traditionally involves extensive capture sessions with complex and expensive studio devices like the LightStage system. While recent strides in neural representations have enabled the…

计算机视觉与模式识别 · 计算机科学 2024-07-31 ShahRukh Athar , Shunsuke Saito , Zhengyu Yang , Stanislav Pidhorsky , Chen Cao

Recently, text-guided digital portrait editing has attracted more and more attentions. However, existing methods still struggle to maintain consistency across time, expression, and view or require specific data prerequisites. To solve these…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Haiyao Xiao , Chenglai Zhong , Xuan Gao , Yudong Guo , Juyong Zhang

This paper addresses the challenge of reconstructing photorealistic and animatable 3D human avatars from monocular videos. While existing methods rely on combining per-subject optimization with generic human priors, they often fail to…

计算机视觉与模式识别 · 计算机科学 2026-05-25 Gangjian Zhang , Jian Shu , Sicheng Yu , Wenhao Shen , Yu Feng , Hao Wang

SmartAvatar is a vision-language-agent-driven framework for generating fully rigged, animation-ready 3D human avatars from a single photo or textual prompt. While diffusion-based methods have made progress in general 3D object generation,…

计算机视觉与模式识别 · 计算机科学 2025-06-06 Alexander Huang-Menders , Xinhang Liu , Andy Xu , Yuyao Zhang , Chi-Keung Tang , Yu-Wing Tai

We present a method that reconstructs and animates a 3D head avatar from a single-view portrait image. Existing methods either involve time-consuming optimization for a specific person with multiple images, or they struggle to synthesize…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Xueting Li , Shalini De Mello , Sifei Liu , Koki Nagano , Umar Iqbal , Jan Kautz

Video inverse problems are fundamental to streaming, telepresence, and AR/VR, where high perceptual quality must coexist with tight latency constraints. Diffusion-based priors currently deliver state-of-the-art reconstructions, but existing…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Weimin Bai , Suzhe Xu , Yiwei Ren , Jinhua Hao , Ming Sun , Wenzheng Chen , He Sun

While there has been significant progress in the field of 3D avatar creation from visual observations, modeling physically plausible dynamics of humans with loose garments remains a challenging problem. Although a few existing works address…

图形学 · 计算机科学 2025-10-03 Changmin Lee , Jihyun Lee , Tae-Kyun Kim

We present HARP (HAnd Reconstruction and Personalization), a personalized hand avatar creation approach that takes a short monocular RGB video of a human hand as input and reconstructs a faithful hand avatar exhibiting a high-fidelity…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Korrawe Karunratanakul , Sergey Prokudin , Otmar Hilliges , Siyu Tang

Text-to-image diffusion models have significantly improved the seamless integration of visual text into diverse image contexts. Recent approaches further improve control over font styles through fine-tuning with predefined font…

计算机视觉与模式识别 · 计算机科学 2025-06-09 Myungkyu Koo , Subin Kim , Sangkyung Kwak , Jaehyun Nam , Seojin Kim , Jinwoo Shin

Creating a controllable and relightable digital avatar from multi-view video with fixed illumination is a very challenging problem since humans are highly articulated, creating pose-dependent appearance effects, and skin as well as clothing…

计算机视觉与模式识别 · 计算机科学 2024-07-29 Diogo Luvizon , Vladislav Golyanik , Adam Kortylewski , Marc Habermann , Christian Theobalt

Feedforward monocular face capture methods seek to reconstruct posed faces from a single image of a person. Current state of the art approaches have the ability to regress parametric 3D face models in real-time across a wide range of…

计算机视觉与模式识别 · 计算机科学 2024-09-13 Kelian Baert , Shrisha Bharadwaj , Fabien Castan , Benoit Maujean , Marc Christie , Victoria Abrevaya , Adnane Boukhayma

We present a novel method to learn Personalized Implicit Neural Avatars (PINA) from a short RGB-D sequence. This allows non-expert users to create a detailed and personalized virtual copy of themselves, which can be animated with realistic…

计算机视觉与模式识别 · 计算机科学 2022-04-11 Zijian Dong , Chen Guo , Jie Song , Xu Chen , Andreas Geiger , Otmar Hilliges

We present InstantMesh, a feed-forward framework for instant 3D mesh generation from a single image, featuring state-of-the-art generation quality and significant training scalability. By synergizing the strengths of an off-the-shelf…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Jiale Xu , Weihao Cheng , Yiming Gao , Xintao Wang , Shenghua Gao , Ying Shan

Recent works have shown that neural radiance fields (NeRFs) on top of parametric models have reached SOTA quality to build photorealistic head avatars from a monocular video. However, one major limitation of the NeRF-based avatars is the…

计算机视觉与模式识别 · 计算机科学 2024-11-08 Huan Wang , Feitong Tan , Ziqian Bai , Yinda Zhang , Shichen Liu , Qiangeng Xu , Menglei Chai , Anish Prabhu , Rohit Pandey , Sean Fanello , Zeng Huang , Yun Fu

Neural Radiance Field (NeRF) based 3D reconstruction is highly desirable for immersive Augmented and Virtual Reality (AR/VR) applications, but achieving instant (i.e., < 5 seconds) on-device NeRF training remains a challenge. In this work,…

硬件体系结构 · 计算机科学 2025-03-31 Sixu Li , Chaojian Li , Wenbo Zhu , Boyang Yu , Yang Zhao , Cheng Wan , Haoran You , Huihong Shi , Yingyan Celine Lin