Gemini Omni Flash Preview
DeprecatedPreview endpoint of Gemini Omni Flash, Google's conversational video generation and editing model; the model page lists it under Versions as 'Preview: gemini-omni-flash-preview' alongside the stable gemini-omni-1.1-flash. Google's deprecations table, read 2026-09-01, gives release June 30, 2026, shutdown September 30, 2026 and gemini-omni-1.1-flash as the recommended replacement; the row is NOT gray, and gray is Google's own marker for an endpoint that is already off, so the status here is deprecated and not retired. Prices are identical to the stable endpoint and still published: input $1.50 per 1M labelled text / image / video / audio, and two output rates — text and thinking tokens $9.00/1M, video tokens $17.50/1M — with price_output_per_mtok holding the video rate. Google states billing is on total output token consumption at 5,792 tokens per second of 720p video, an effective ~$0.10 per second under Standard pricing. Paid tier only; no free tier and no context caching rate published. Context window 1,048,576 tokens; the output is video, so no max output token count exists. Added 2026-09-01 together with its replacement target, because a pointer to a row the catalog does not carry hard-fails check-refs (LEARNINGS #61b).
Gemini Omni Flash Preview by Google costs $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended), with a 1.0M (1.048.576-token) context window. It is deprecated, shutting down 30 Sep 2026; the recommended replacement is Gemini Omni Flash.
Last verified: 26 Sep 2026 · sourced from official provider documentation
Retires on 30 Sep 2026
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
This model is deprecated. Recommended migration: Gemini Omni Flash.
Read the Gemini Omni Flash Preview migration guide → replacement & cheapest alternatives
Source: Google official documentation ↗
Gemini Omni Flash Preview — questions & answers
How much does Gemini Omni Flash Preview cost?
Gemini Omni Flash Preview is priced at $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended at a 3:1 input-to-output ratio).
What is the context window of Gemini Omni Flash Preview?
Gemini Omni Flash Preview has a 1.048.576-token context window (1.0M).
Is Gemini Omni Flash Preview being deprecated or retired?
Gemini Omni Flash Preview is deprecated. Its scheduled shutdown date is 30 Sep 2026. The recommended migration is Gemini Omni Flash.
Does Gemini Omni Flash Preview support image or vision input?
Yes — Gemini Omni Flash Preview accepts image input. Listed modalities: text, image, video, video-out.
Can I self-host Gemini Omni Flash Preview?
No — Gemini Omni Flash Preview is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.
What is the API model string for Gemini Omni Flash Preview?
The API model identifier for Gemini Omni Flash Preview is "gemini-omni-flash-preview" when calling Google's API.
Track Gemini Omni Flash Preview price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.