Gemini Omni Flash Preview

Deprecated

Preview endpoint of Gemini Omni Flash, Google's conversational video generation and editing model; the model page lists it under Versions as 'Preview: gemini-omni-flash-preview' alongside the stable gemini-omni-1.1-flash. Google's deprecations table, read 2026-09-01, gives release June 30, 2026, shutdown September 30, 2026 and gemini-omni-1.1-flash as the recommended replacement; the row is NOT gray, and gray is Google's own marker for an endpoint that is already off, so the status here is deprecated and not retired. Prices are identical to the stable endpoint and still published: input $1.50 per 1M labelled text / image / video / audio, and two output rates — text and thinking tokens $9.00/1M, video tokens $17.50/1M — with price_output_per_mtok holding the video rate. Google states billing is on total output token consumption at 5,792 tokens per second of 720p video, an effective ~$0.10 per second under Standard pricing. Paid tier only; no free tier and no context caching rate published. Context window 1,048,576 tokens; the output is video, so no max output token count exists. Added 2026-09-01 together with its replacement target, because a pointer to a row the catalog does not carry hard-fails check-refs (LEARNINGS #61b).

Gemini Omni Flash Preview by Google costs $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended), with a 1.0M (1.048.576-token) context window. It is deprecated, shutting down 30 Sep 2026; the recommended replacement is Gemini Omni Flash.

Last verified: 26 Sep 2026 · sourced from official provider documentation

Retires on 30 Sep 2026

Provider
Google
Status
Deprecated
Input price
$1.50 / 1M tokens
Output price
$17.50 / 1M tokens
Cached input
—
Blended price
$5.50 / 1M tokens
Context window
1.048.576 tokens (1.0M)
Max output
—
Modality
text, image, video, video-out
Open weights
No — API only
Knowledge cutoff
—
Released
30 Jun 2026
API string
gemini-omni-flash-preview

Source: Google official documentation ↗

FAQ

Gemini Omni Flash Preview — questions & answers

How much does Gemini Omni Flash Preview cost?

Gemini Omni Flash Preview is priced at $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended at a 3:1 input-to-output ratio).

What is the context window of Gemini Omni Flash Preview?

Gemini Omni Flash Preview has a 1.048.576-token context window (1.0M).

Is Gemini Omni Flash Preview being deprecated or retired?

Gemini Omni Flash Preview is deprecated. Its scheduled shutdown date is 30 Sep 2026. The recommended migration is Gemini Omni Flash.

Does Gemini Omni Flash Preview support image or vision input?

Yes — Gemini Omni Flash Preview accepts image input. Listed modalities: text, image, video, video-out.

Can I self-host Gemini Omni Flash Preview?

No — Gemini Omni Flash Preview is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.

What is the API model string for Gemini Omni Flash Preview?

The API model identifier for Gemini Omni Flash Preview is "gemini-omni-flash-preview" when calling Google's API.