AI Video Models by WAN
Leverage the latest AI video generation models to create cinematic high-fidelity videos from text and images.

WAN Video 3.0 Prime
Next-generation multimodal video generation model that creates cinematic videos up to 30 seconds at amazing speed

WAN Video 3.0
Next-generation AI video generation model with cinematic quality and advanced control capabilities

WAN Video 2.7
A next-generation video generation model that creates cinematic visuals with free control and overwhelming consistency

WAN Video 2.6
A next-generation video generation model with movie quality that supports multi-cuts up to 15 seconds and realistic audio synchronization
AI Image Models by WAN
Use industry-leading AI image generation models to turn prompts and reference images into stunning, production-ready visuals.

WAN Image 2.7 Pro
A professional-grade 4K image generation model with built-in thinking mode, enabling advanced composition and logical inference.

WAN Image 2.7
An innovative image generation and editing model equipped with a "thinking mode" and high-precision text rendering

WAN Image 2.6
Alibaba WANの最新画像生成モデルは、その圧倒的なディテールと光影表現で見る者を魅了しています
Models from WAN
Discover the core capabilities and technical highlights of each model.
| Model | Capability | Description |
|---|---|---|
| WAN Image 2.7 Pro | text to image 4K | A professional-grade 4K image generation model with built-in thinking mode, enabling advanced composition and logical inference. |
| WAN Image 2.7 | text to image 2K | An innovative image generation and editing model equipped with a "thinking mode" and high-precision text rendering |
| WAN Video 3.0 Prime | text to video 1080P | Next-generation multimodal video generation model that creates cinematic videos up to 30 seconds at amazing speed |
| WAN Video 3.0 | text to video 1080P | Next-generation AI video generation model with cinematic quality and advanced control capabilities |
| WAN Video 2.7 | text to video 1080P | A next-generation video generation model that creates cinematic visuals with free control and overwhelming consistency |
| WAN Image 2.6 | text to image 1K | Alibaba WANの最新画像生成モデルは、その圧倒的なディテールと光影表現で見る者を魅了しています |
| WAN Video 2.6 | text to video 1080P | A next-generation video generation model with movie quality that supports multi-cuts up to 15 seconds and realistic audio synchronization |
Use cases you can build with WAN models
Discover how WAN models power real-world applications and creative workflows. Ready for production — one API key for all use cases.
Commercial advertising poster production
Leveraging WAN Image 2.7 Pro's thinking mode and advanced text rendering capabilities, you can generate advertising visuals with complex compositions and text in up to 4K resolution.
Audio-Synced Short Video Production
Leveraging the audio-video synchronization features of WAN Video 3.0 and WAN Video 2.6, you can create cinema-quality promotional videos tailored to dialogue and BGM.
Narrative video production
By leveraging multi-shot coordination and camera work control in WAN Video 3.0, you can create cinematic previsualizations and video sequences with consistent storytelling.
Character Design and Consistent Rollout
Using WAN Image 2.7 Pro's feature of up to 12 reference images, you can generate multiple variations of assets while maintaining character consistency and tone.
FAQ
Explore other model providers
Explore 400+ production-ready models. One API key, one task flow.

HappyHorse
A generative AI provider boasting a bright, positive worldview and versatile expressive power. From live-action to anime and cyberpunk, it quickly brings vivid video content with audio to life.

Vidu
A provider specializing in high-quality video generation. Rapidly produces video featuring natural movement, lighting and shadows, and detailed character expressions from text or reference images, supporting commercial creative production.

MiniMax
MiniMax provides cutting-edge AI video generation models capable of expressing everything from high-quality live-action to anime styles. It supports commercial creatives with high-definition visuals, synchronized audio, and fast generation speeds.
