Wan-Animate-2 pushes character animation beyond pose-driven pipelines: reference image + driving video in, cleaner character motion out. 🚀
🤖
modelscope.ai/models/Wan-AI/…
🌐
modelscope.ai/studios/Wan-AI…
🏆 Blind user study: preferred over Wan-Animate in 70%+ of overall-quality comparisons, with perceived quality comparable to Dreamina and Kling-MotionControl.
🎬 Cleaner animation: directly processes driving-video motion instead of relying on pose skeletons, reducing identity drift, artifacts, and shape distortion across different characters.
🎥 Text camera control: Viewpoint LoRA lets prompts steer the output perspective while keeping the original motion consistent.
⚡ Streaming path: Wan-Animate-2-Lite reports 24fps at 400x720, with stable long-sequence generation for live avatars and interactive virtual environments.
Built with dual-branch DiT, Time-Align RoPE, Sparse-Ref Attention, 100K+ video pairs, and 50K Unreal Engine multi-view samples.