中文
相关论文

相关论文: ID-Animator: Zero-Shot Identity-Preserving Human V…

200 篇论文

Subject-driven video generation (SDV-Gen) aims to produce videos of a specific subject by adapting a pretrained video model, enabling personalized and application-driven content creation. To achieve this goal, per-subject tuning methods…

计算机视觉与模式识别 · 计算机科学 2026-05-06 Daneul Kim , Jingxu Zhang , Wonjoon Jin , Sunghyun Cho , Qi Dai , Jaesik Park , Chong Luo

Human dance generation (HDG) aims to synthesize realistic videos from images and sequences of driving poses. Despite great success, existing methods are limited to generating videos of a single person with specific backgrounds, while the…

计算机视觉与模式识别 · 计算机科学 2024-01-25 Zhe Xu , Kun Wei , Xu Yang , Cheng Deng

Diffusion-based human animation aims to animate a human character based on a source human image as well as driving signals such as a sequence of poses. Leveraging the generative capacity of diffusion model, existing approaches are able to…

计算机视觉与模式识别 · 计算机科学 2024-12-30 Fa-Ting Hong , Zhan Xu , Haiyang Liu , Qinjie Lin , Luchuan Song , Zhixin Shu , Yang Zhou , Duygu Ceylan , Dan Xu

Recent advancements in personalized image generation using diffusion models have been noteworthy. However, existing methods suffer from inefficiencies due to the requirement for subject-specific fine-tuning. This computationally intensive…

计算机视觉与模式识别 · 计算机科学 2023-12-12 Xu Peng , Junwei Zhu , Boyuan Jiang , Ying Tai , Donghao Luo , Jiangning Zhang , Wei Lin , Taisong Jin , Chengjie Wang , Rongrong Ji

Recent advancements in text-to-image generation have enabled significant progress in zero-shot 3D shape generation. This is achieved by score distillation, a methodology that uses pre-trained text-to-image diffusion models to optimize the…

计算机视觉与模式识别 · 计算机科学 2023-05-29 Zhenzhen Weng , Zeyu Wang , Serena Yeung

Creating realistic 3D facial animation is crucial for various applications in the movie production and gaming industry, especially with the burgeoning demand in the metaverse. However, prevalent methods such as blendshape-based approaches…

计算机视觉与模式识别 · 计算机科学 2023-08-14 Haoyu Wang , Haozhe Wu , Junliang Xing , Jia Jia

For the last decades, the concern of producing convincing facial animation has garnered great interest, that has only been accelerating with the recent explosion of 3D content in both entertainment and professional activities. The use of…

图形学 · 计算机科学 2020-10-13 Eloïse Berson , Catherine Soladié , Nicolas Stoiber

We present X-Dancer, a novel zero-shot music-driven image animation pipeline that creates diverse and long-range lifelike human dance videos from a single static image. As its core, we introduce a unified transformer-diffusion framework,…

计算机视觉与模式识别 · 计算机科学 2025-07-14 Zeyuan Chen , Hongyi Xu , Guoxian Song , You Xie , Chenxu Zhang , Xin Chen , Chao Wang , Di Chang , Linjie Luo

This work proposes a novel method to generate realistic talking head videos using audio and visual streams. We animate a source image by transferring head motion from a driving video using a dense motion field generated using learnable…

计算机视觉与模式识别 · 计算机科学 2022-10-07 Madhav Agarwal , Rudrabha Mukhopadhyay , Vinay Namboodiri , C V Jawahar

The objective of face animation is to generate dynamic and expressive talking head videos from a single reference face, utilizing driving conditions derived from either video or audio inputs. Current approaches often require fine-tuning for…

计算机视觉与模式识别 · 计算机科学 2024-07-15 He Feng , Donglin Di , Yongjia Ma , Wei Chen , Tonghua Su

Personalized 3D avatars require an animatable representation of digital humans. Doing so instantly from monocular videos offers scalability to broad class of users and wide-scale applications. In this paper, we present a fast, simple, yet…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Pramish Paudel , Anubhav Khanal , Ajad Chhatkuli , Danda Pani Paudel , Jyoti Tandukar

Appearance editing according to user needs is a pivotal task in video editing. Existing text-guided methods often lead to ambiguities regarding user intentions and restrict fine-grained control over editing specific aspects of objects. To…

计算机视觉与模式识别 · 计算机科学 2025-05-30 Tongtong Su , Chengyu Wang , Jun Huang , Dongming Lu

Existing person video generation methods either lack the flexibility in controlling both the appearance and motion, or fail to preserve detailed appearance and temporal consistency. In this paper, we tackle the problem of motion transfer…

计算机视觉与模式识别 · 计算机科学 2019-08-13 Kun Cheng , Hao-Zhi Huang , Chun Yuan , Lingyiqing Zhou , Wei Liu

Person re-identification (ReID) suffers from a lack of large-scale high-quality training data due to challenges in data privacy and annotation costs. While previous approaches have explored pedestrian generation for data augmentation, they…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Changxiao Ma , Chao Yuan , Xincheng Shi , Yuzhuo Ma , Yongfei Zhang , Longkun Zhou , Yujia Zhang , Shangze Li , Yifan Xu

With the booming of virtual reality (VR) technology, there is a growing need for customized 3D avatars. However, traditional methods for 3D avatar modeling are either time-consuming or fail to retain similarity to the person being modeled.…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Chuanyu Pan , Guowei Yang , Taijiang Mu , Yu-Kun Lai

In this paper we present a new deep learning-driven approach to image-based synthesis of animations involving humanoid characters. Unlike previous deep approaches to image-based animation our method makes no assumptions on the type of…

图形学 · 计算机科学 2019-08-14 John Kanji , David I. W. Levin

We present a novel approach that enables photo-realistic re-animation of portrait videos using only an input video. In contrast to existing approaches that are restricted to manipulations of facial expressions only, we are the first to…

We present a deep learning-based framework for portrait reenactment from a single picture of a target (one-shot) and a video of a driving subject. Existing facial reenactment methods suffer from identity mismatch and produce inconsistent…

计算机视觉与模式识别 · 计算机科学 2020-04-28 Sitao Xiang , Yuming Gu , Pengda Xiang , Mingming He , Koki Nagano , Haiwei Chen , Hao Li

Video diffusion models substantially boost the productivity of artistic workflows with high-quality portrait video generative capacity. However, prevailing pipelines are primarily constrained to single-shot creation, while real-world…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Jiahao Wang , Hualian Sheng , Sijia Cai , Weizhan Zhang , Caixia Yan , Yachuang Feng , Bing Deng , Jieping Ye

Recent advances in large pretrained text-to-image models have shown unprecedented capabilities for high-quality human-centric generation, however, customizing face identity is still an intractable problem. Existing methods cannot ensure…

计算机视觉与模式识别 · 计算机科学 2024-01-30 Qinghe Wang , Xu Jia , Xiaomin Li , Taiqing Li , Liqian Ma , Yunzhi Zhuge , Huchuan Lu