Ready to Generate
Configure your inputs and click run to generate an image preview.
WAN Video 2.6 is a next-generation AI video generation model renowned for its outstanding expressiveness and narrative capabilities. It can seamlessly generate complex multi-cut (multi-perspective) videos up to 15 seconds long at a high resolution of up to 1080p (24fps). The most distinctive feature of this model is the native generation of high-quality stereo audio that is fully synchronized with the visuals. It outputs videos with immersive sound effects, including natural lip-sync and the ability to differentiate dialogue among multiple characters. Furthermore, it excels at advanced physical simulations, fine facial expression control, and flexible camera work, while strictly maintaining the consistency of characters and the worldview. It supports a wide range of styles, from realistic live-action to anime, 3D, and cyberpunk, delivering commercial-grade value for creating movie storyboards, short advertising films, and professional-level content.
| Resolution | Credits Consumed |
|---|---|
| 720p(credits/s) | 4 |
| 1080p(credits/s) | 6 |
| Parameter | Specification |
|---|---|
| Core Capability | Text to video |
| Resolution | 720p,1080p |
| Aspect Ratio | 16:9,9:16,1:1 |
| Duration | 5,10,15 |

BytePlus
A professional high-performance video generation model that produces cinema-quality videos up to 30 seconds long with synchronized audio and visuals

BytePlus
A lightweight video generation model that achieves both overwhelming generation speed and high quality

WAN
Next-generation multimodal video generation model that creates cinematic videos up to 30 seconds at amazing speed

The latest model that seamlessly combines text and audio to quickly generate consistent, high-quality videos.

WAN
Next-generation AI video generation model with cinematic quality and advanced control capabilities

Google DeepMind's state-of-the-art AI video generation model. It produces cinematic-level visuals and synchronized sound.