Drag and drop media files from your computer, paste from clipboard (Ctrl/Cmd+V), or provide a URL. Accepted: .jpg, .jpeg, .png, .webp
Ready to Generate
Configure your inputs and click run to generate an image preview.
WAN Video 2.6 is a state-of-the-art video generation model that brings together cutting-edge multimodal technology. It breathes life into dynamic, coherent cinematic footage while preserving the details, characters, and worldview of the reference still images with extremely high accuracy. In addition to the realism of physical simulations, it reproduces fine-grained character expressions, eye movements, and professional camera work at a high level. Furthermore, it supports advanced audio synchronization (lip-syncing and natural conversations between multiple characters), enabling direction just like actual film shooting. In the creation of movie previews, high-quality commercial videos, and creative video content, it empowers creators to achieve commercial-level visual expression with overwhelming efficiency without limiting their imagination.
| Resolution | Credits Consumed |
|---|---|
| 720p(credits/s) | 4 |
| 1080p(credits/s) | 6 |
| Parameter | Specification |
|---|---|
| Core Capability | Reference image to video |
| Resolution | 720p,1080p |
| Aspect Ratio | 16:9,9:16,1:1,4:3,3:4 |
| Duration | 5,10 |

BytePlus
An innovative multimodal generative model that creates cinematic videos up to 30 seconds long from reference images.

BytePlus
A lightweight video generation model that quickly generates high-quality videos from reference images.

WAN
Next-generation model that transforms still images and audio into movie-quality videos with overwhelming generation speed

Google's fast multimodal video generation and editing model, which highly integrates audio and video

WAN
A next-generation integrated AI video generation model that achieves cinematic-level image quality and high consistency

Google's top-tier video generation model, crafting cinematic-level visuals and sound.