Exquisite facial expressions and subtle movements
Smoothly renders natural eye movements and lip motions, significantly enhancing the character's expressiveness.
A video generation model excelling in facial expression reproducibility, subtle movements, and stable camera work. It generates high-quality videos from still images or multiple reference images.
Each model has its own Playground, pricing, and field docs. Filter by capability to open the one you need.
Model id and a short note on what each capability is for.
| Model | ID | Description |
|---|---|---|
| Vidu Q2 Pro(refs image to video) | vidu-video-viduq2-pro-refs-image-to-video | Vidu’s next-generation professional model, boasting overwhelming character consistency and cinematic-level expressiveness |
| Vidu Q2 Pro(image to video) | vidu-video-viduq2-pro-image-to-video | Bring still images to life, a next-generation image and video generation model with cinematic quality |
Balancing natural character expression and stable camera work, it supports professional video production.
On the rooftop of the futuristic city, a cyber girl passionately performs street dance in the neon rain, all captured by a dynamic camera.
Prompt
On the rooftop of the futuristic city, a cyber girl in a glowing jacket is dancing street dance. Neon rainwater reflects on the ground, with high-speed passing hover cars and giant holographic advertisements in the backg…
Generated with Vidu Q3 Pro
Generated with Vidu Q2
Generated with Veo 3.1
An extremely realistic white wolf races across the mountain peak on a snowy night. The moonlight and the dancing snow dynamically chase after it.
Prompt
An extremely realistic wild white wolf is galloping through the snowy mountains at night. Its fur flutters in the wind, snow particles fly up, and the moonlight illuminates the lines of its muscles and the white breath i…
Generated with Vidu Q3 Pro
Generated with Vidu Q2
Generated with Veo 3.1
A powerful mage stands atop the shattered stone altar, raising his staff and unleashing an overwhelming burst of flame encased in Serpentine Blaze. This cinematic scene depicts a dramatic arc shot and a close-up shot of the moment of determination.
Prompt
A cinematic scene. Standing atop a shattered stone altar stands a powerful sorcerer. Above him, crimson clouds swirl ominously and violently. The camera rises from behind him, illuminating the ominous glowing runes at hi…
Generated with Vidu Q3 Pro
Generated with Vidu Q2
Generated with Veo 3.1
Typical jobs this series is used for.
Naturally renders subtle facial expressions such as eye movements and blinks, allowing you to create narrative-rich character cuts.
Generates video clips that preserve the subject's characteristics and worldview using up to 3 reference images.
Supports diverse styles such as photorealistic and 3D, enabling the creation of high-quality advertising materials in 1080P resolution.
Check production limits and starting credit cost before you pick another model for the same job.
| Model | Max clip duration (single run) | Multimodal reference inputs | Max resolution | Native audio | Sousaku starting price |
|---|---|---|---|---|---|
| Vidu Q2 Pro | Up to 10 seconds | Up to 3 images | Up to 1080P | ❌ | 2 Credits/s |
| Vidu Q3 | Up to 16 seconds | Up to 3 images | Up to 1080P | ✅ | 2 Credits/s |
| Vidu Q3 Pro | Up to 16 seconds | ❌ | Up to 1080P | ✅ | 2 Credits/s |
From sign-up to your first successful call. The same steps work for every model.
Sign up and verify so you can issue an API key and use Playground.
Open a card above to see that model's details, pricing, and field docs.
Confirm inputs and parameters in the form before you automate.
The field docs on the detail page list required parameters. Use the same names in create_task.
Call the generate API with your key, poll until the task succeeds, then download the result.
Track credit use on the model's pricing page and in the dashboard.
Other series on Sousaku. One API key, one task flow.