Gemini Omni Flash
Google's fast, cost-efficient omni video model for text-to-video, image-to-video, and video-to-video with reference and frame conditioning.
Gemini Omni Flash is a cloud video-generation model built by Google. It produces short video clips from a text prompt. Clips run up to 10s. It also supports reference images. It runs through Runware using your own API key, from $1.50 per million input tokens.
- Pricing
- $1.50 / 1M in · $17.50 / 1M out
- Aspect ratios
- 16:9, 9:16, 1:1
- Clip length
- Up to 10s
Examples
Generated with Gemini Omni Flash via Runware. The same two prompts run against every video model in the catalog — one landscape, one portrait — so the only thing that changes between two models’ clips is the model.
Camera motion, scene coherence over time
Fluid motion, fine detail, temporal stability
Google and Google DeepMind build the Gemini family of multimodal models, the Imagen and Nano Banana image models, the Lyria music models, and the Veo video models.
deepmind.google ↗