Google Gemini Omni Flash
Google's fast, cost-efficient omni video model for text-to-video, image-to-video, and video-to-video with reference and frame conditioning.
Gemini Omni Flash is a cloud video-generation model built by Google. It produces short video clips from a text prompt. Clips run up to 10s. It also supports reference images. It runs through Runware using your own API key, from $1.50 per million input tokens.
- Released
- May 2026
- Pricing
- $1.50 / 1M in · $17.50 / 1M out
- Aspect ratios
- 16:9, 9:16, 1:1
- Clip length
- Up to 10s
Examples
Generated with Gemini Omni Flash via Runware. The same two prompts run against every video model in the catalog (one landscape, one portrait), so the only thing that changes between two models’ clips is the model.
Camera motion, scene coherence over time
Fluid motion, fine detail, temporal stability
Google and Google DeepMind build the Gemini family of multimodal models, the Imagen and Nano Banana image models, the Lyria music models, and the Veo video models.
deepmind.google ↗Google's most capable Flash model, with clear gains over 3.7 Flash in software engineering, agentic tasks, and multi-step reasoning, plus vision and a 1M-token context window.
Google's newest Flash workhorse, a more capable successor to 3.6 Flash with stronger software engineering, better document comprehension and more disciplined tool use, at half the price per token.
Google's most efficient multimodal Flash model: an updated reasoning stack with stronger coding and computer-use quality, at the same speed and scale as the rest of the Flash line.
Google's most capable multimodal model with deep reasoning and 1M token context.
Google's updated Nano Banana 2 image model, with sharper graphic composition, steadier subjects across edits and closer prompt following, from text or up to 14 reference images at 1K to 4K.
Ultra-fast image generation model optimised for speed and creative output.
One dies in 70 days, one was outranked by its own maker, one still ships. The matchup, re-run for real: six clips, first takes only.
Nine images, four clips, $3.17 in API bills: what ByteDance's Seedream and Seedance actually return on prompts we ran ourselves.
Six flagships in eight weeks, an open-weight model that tied GPT-5.5 on coding, and Meta walked away from Llama. The Feb–May 2026 LLM catalog.