← All models

Gemini Omni Flash

Video Preview Full comparison ↗

google-deepmind/gemini-omni-flash · by Google DeepMind · Diffusion transformer (DiT)

Pricing — 1 offering(s)

Standard (720p)

  • $0.10 / second Current 2026-06-30 → present

Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.

See how Gemini Omni Flash fits into a cost-aware routing setup

See how →

Capability profile

Prompt adherence strong
temporal consistency moderate
motion quality moderate
native audio weak
camera control unknown
character consistency moderate
Resolution ceiling weak
Inference speed strong
clip duration ceiling weak
image to video quality strong
Compositional accuracy unknown
cost efficiency strong
video editing moderate

Operator guidance

Default choice when editing capability and low latency are required and native audio is not needed. At $0.10/s (720p only), this is the cheapest option in the registry with API-level video editing. Now also independently confirmed as the #1-ranked T2V model on the Artificial Analysis Video Arena, Elo 1239 of 28 models tracked (2026-08-19, SCO-462 sweep) — genuinely strong on quality, not just cost/latency. For silent T2V/I2V at lower cost, use Veo 3.1 Lite ($0.05–0.08/s); for native audio or 4K, use Veo 3.1 Standard/Fast. The Interactions API's stateful session model (previous_interaction_id chains turns) enables multi-turn refinement — note that conversational session management is an API design pattern, not a product-only chat layer. The edit task (V2V) has an active limitation with video input processing as of 2026-07-02 — verify current status before committing to production editing workflows. Geo restriction: externally-uploaded video editing is unavailable in EEA, Switzerland, and UK.

Use cases

Limitations

Citations