WAN Video 3.0 Prime

While maintaining high-speed inference speed, it generates high-quality videos that combine text, images, and audio, up to 30 seconds in length with synchronized audio.

Text Friendly
Audio Support
Commercial Use

Input

Reference image*

Drag and drop media files from your computer, paste from clipboard (Ctrl/Cmd+V), or provide a URL. Accepted: .jpg, .jpeg, .png, .webp

Result

Idle

Ready to Generate

Configure your inputs and click run to generate an image preview.

WAN Video 3.0 Prime Model Introduction

Model Overview

「WAN Video 3.0 Prime」 is a next-generation multimodal video generation model that combines overwhelming generation speed with excellent visual quality. It retains the delicate detail expression of the standard model while dramatically boosting inference speed to maximize creators' work efficiency. It supports text-to-video generation, as well as consistent motion rendering from still images (start and end frames can be specified), and advanced "Reference-to-Video (Ref-to-Video)" which precisely tracks and integrates character identity, visual style, and audio from multiple reference materials (images, videos, audio). It can output videos up to 30 seconds long at once in various aspect ratios and up to 1080p high resolution, and also generates perfectly synchronized high-quality audio tracks simultaneously. This is a powerful solution that enhances creative expression in professional scenarios like advertising, design, and video production.

Pricing

ResolutionCredits Consumed
480p(credits/s)6
720p(credits/s)12
1080p(credits/s)18

Technical Specifications

ParameterSpecification
Core CapabilityReference image to video
Resolution480p,720p,1080p
Aspect Ratio16:9,4:3,1:1,3:4,9:16
Duration2,3,4,5,6,7,8,9,10,11,12,13,14,15,16,17,18,19,20,21,22,23,24,25,26,27,28,29,30

Explore similar models