Explore Alibaba's Wan family—from open video generation foundations to longer multimodal creation, native audio and precision editing.
Model overview
A generative model family spanning video, image, audio and editing workflows.
Wan is Alibaba's family of generative video and image models. The series covers text-to-video and image-to-video creation and has expanded into keyframe control, reference-to-video, character animation, native audio-video generation and instruction-led editing.
The newest official Wan experience brings generation and editing into a broader multimodal workflow. Capabilities, model IDs, deployment regions and output limits still vary by endpoint, so Panify verifies each workflow before making it available.
Creative system
Choose the Wan workflow around the source material and control your project needs.
Model timeline
Wan has grown from open visual generation into managed, multimodal production workflows.
Creation modes
Each workflow gives a different balance of speed, reference control and editability.
Creative applications
Use the family for both early creative exploration and more controlled production concepts.
Develop longer product and campaign narratives with controlled shot progression.
Plan solo or ensemble performance with authorized identity and voice references.
Previsualize movement, camera language and physically complex scenes.
Preserve important product details across connected shots and settings.
Build scenes around speech, ambience, rhythm and stereo sound.
Refine subjects, backgrounds, lighting or sound and extend supported footage.
Wan FAQ
Review Wan 3.0's official capability set and prepare your authorized references while Panify verifies production access.