English

Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation

Computer Vision and Pattern Recognition 2026-01-12 v1

Abstract

Generating high-quality 3D characters from single images remains a significant challenge in digital content creation, particularly due to complex body poses and self-occlusion. In this paper, we present RCM (Rotate your Character Model), an advanced image-to-video diffusion framework tailored for high-quality novel view synthesis (NVS) and 3D character generation. Compared to existing diffusion-based approaches, RCM offers several key advantages: (1) transferring characters with any complex poses into a canonical pose, enabling consistent novel view synthesis across the entire viewing orbit, (2) high-resolution orbital video generation at 1024x1024 resolution, (3) controllable observation positions given different initial camera poses, and (4) multi-view conditioning supporting up to 4 input images, accommodating diverse user scenarios. Extensive experiments demonstrate that RCM outperforms state-of-the-art methods in both novel view synthesis and 3D generation quality.

Keywords

Cite

@article{arxiv.2601.05722,
  title  = {Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation},
  author = {Jin Wang and Jianxiang Lu and Comi Chen and Guangzheng Xu and Haoyu Yang and Peng Chen and Na Zhang and Yifan Xu and Longhuang Wu and Shuai Shao and Qinglin Lu and Ping Luo},
  journal= {arXiv preprint arXiv:2601.05722},
  year   = {2026}
}

Comments

11 pages, 8 figures

R2 v1 2026-07-01T08:57:38.733Z