Ready to Generate
Configure your inputs and click run to generate an image preview.
WAN Video 3.0 Prime is a next-generation integrated video generation model that combines overwhelming generation speed and versatile multimodal inputs. While maintaining the extremely high rendering quality of the standard model, it significantly improves end-to-end inference speed and accelerates creative production workflows.
It supports highly flexible approaches, including precise text-to-video creation, 'image-to-video' that dynamically animates characters and compositions of still images while fully preserving them, and 'reference video generation' that integrates images, videos, and audio to craft consistent stories. It can generate seamless videos up to 30 seconds long in a single pass without cuts or splits, caters to a wide range of styles from live-action to anime, and works with various aspect ratios and high resolutions (up to 1080p). It also supports synchronous audio generation, delivering ready-to-use video experiences for commercial productions such as advertisements, films, and entertainment.
| Resolution | Credits Consumed |
|---|---|
| 480p(credits/s) | 6 |
| 720p(credits/s) | 12 |
| 1080p(credits/s) | 18 |
| Parameter | Specification |
|---|---|
| Core Capability | Text to video |
| Resolution | 480p,720p,1080p |
| Aspect Ratio | 16:9,4:3,1:1,3:4,9:16 |
| Duration | 2,3,4,5,6,7,8,9,10,11,12,13,14,15,16,17,18,19,20,21,22,23,24,25,26,27,28,29,30 |

BytePlus
A professional high-performance video generation model that produces cinema-quality videos up to 30 seconds long with synchronized audio and visuals

BytePlus
A lightweight video generation model that achieves both overwhelming generation speed and high quality

The latest model that seamlessly combines text and audio to quickly generate consistent, high-quality videos.

WAN
Next-generation AI video generation model with cinematic quality and advanced control capabilities

Google DeepMind's state-of-the-art AI video generation model. It produces cinematic-level visuals and synchronized sound.

Kling
An all-around AI video generation model that delivers cinematic image quality and native audio-visual synchronization