Gemini Omni Flash

A fast model that integrates text, images, and audio to generate consistent videos. Advanced physical simulations and natural audio synchronization help you quickly achieve professional-grade video production.

Text Friendly
Audio Support
Commercial Use

Input

Result

Idle

Ready to Generate

Configure your inputs and click run to generate an image preview.

Gemini Omni Flash Model Introduction

Model Overview

Gemini Omni Flash is a next-generation, fast multimodal video generation model developed by Google. It has a high-level understanding and reasoning capability of different materials such as text, images, audio and video, and can instantly generate consistent high-quality videos that naturally integrate these elements.

This model is equipped with excellent physics simulation technology, reproducing phenomena such as gravity and fluid movement just like in the real world. Furthermore, it has a practical video editing function that allows you to change the camera angle and environment according to natural language instructions while consistently maintaining the character's visuals. It also supports audio output that is fully synchronized with the video, generating vivid content of up to 10 seconds.

Since it can be operated intuitively without professional video editing skills, it delivers overwhelming productivity in various creative scenarios such as quickly creating mockups for ad creatives, short videos for SNS, and creating virtual avatars.

Pricing

ResolutionCredits Consumed
720p(credits/s)4

Technical Specifications

ParameterSpecification
Core CapabilityText to video
Resolution720p
Aspect Ratio16:9,9:16
Duration3,4,5,6,7,8,9,10

Explore similar models