Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation
Abstract
Generating high-quality 3D characters from single images re-mains a significant challenge in digital content creation, particularly dueto complex body poses and self-occlusion. In this paper, we present RCM(Rotate your Character Model ), an advanced image-to-video diffusionframework tailored for high-quality novel view synthesis (NVS). Com-pared to existing diffusion-based approaches, RCM offers several keyadvantages: (1) transferring characters with any complex poses into acanonical pose, enabling consistent novel view synthesis across the entireviewing orbit, (2) high-resolution orbital video generation at 1024×1024resolution, (3) controllable observation positions given different initialcamera poses, and (4) multi-view conditioning supporting up to 4 inputimages, accommodating diverse user scenarios. Extensive experimentsdemonstrate that RCM outperforms state-of-the-art methods in novelview synthesis on character generations. The deliverables will be updatedhere.