Gemini Omni Flash

GA

Video generation and editing. Google describes it as a high-performance model for fast, conversational video generation and editing: text and images to video, refinement of generated video through natural-language conversation via the Interactions API, plus video extension, resolution upscaling and interpolation. Status ga on two first-party statements read 2026-09-01 — the model page Versions block reads 'Stable: gemini-omni-1.1-flash' (with 'Preview: gemini-omni-flash-preview' listed separately as the sibling endpoint), and the pricing page calls it 'now generally available to developers on the paid tier of the Gemini API'. Paid tier only: Google states 'Not available' on the free tier for both axes. Input $1.50 per 1M, one rate that the pricing page labels text / image / video / audio, although the model page spec table lists only Text, Image and Video (up to 10s for editing and extension) as supported inputs — that disagreement between two Google surfaces is recorded here rather than resolved. Two output rates: text and thinking tokens $9.00/1M, video tokens $17.50/1M — price_output_per_mtok holds the video rate, the model's product, the same convention gemini-3.1-flash-image uses for its image rate. Google states billing is on total output token consumption at 5,792 tokens per second of 720p video, an effective ~$0.10 per second under Standard pricing. Context window 1,048,576 tokens; output is video (3s-10s at 360p/720p/1080p/4K, 24 FPS) so there is no published max output token count and the field is null. No context caching rate is published. Released August 27, 2026 per Google's deprecations table, which states 'No shutdown date announced' and does NOT mark the row gray — Google is still asserting the endpoint runs. Added 2026-09-01: this row was missing from the catalog since launch, one of four Google models that check-google-prices.mjs could not report under 'on the page, not in the catalog' because every one of its price cells holds more than one rate (LEARNINGS #105).

Gemini Omni Flash by Google costs $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).

Last verified: 26 Sep 2026 · sourced from official provider documentation

Provider
Google
Status
GA
Input price
$1.50 / 1M tokens
Output price
$17.50 / 1M tokens
Cached input
—
Blended price
$5.50 / 1M tokens
Context window
1.048.576 tokens (1.0M)
Max output
—
Modality
text, image, video, video-out
Open weights
No — API only
Knowledge cutoff
—
Released
27 Aug 2026
API string
gemini-omni-1.1-flash

Source: Google official documentation ↗

FAQ

Gemini Omni Flash — questions & answers

How much does Gemini Omni Flash cost?

Gemini Omni Flash is priced at $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended at a 3:1 input-to-output ratio).

What is the context window of Gemini Omni Flash?

Gemini Omni Flash has a 1.048.576-token context window (1.0M).

Is Gemini Omni Flash deprecated?

No — Gemini Omni Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Gemini Omni Flash support image or vision input?

Yes — Gemini Omni Flash accepts image input. Listed modalities: text, image, video, video-out.

Can I self-host Gemini Omni Flash?

No — Gemini Omni Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.

What is the API model string for Gemini Omni Flash?

The API model identifier for Gemini Omni Flash is "gemini-omni-1.1-flash" when calling Google's API.