Gemini Omni Flash
GAVideo generation and editing. Google describes it as a high-performance model for fast, conversational video generation and editing: text and images to video, refinement of generated video through natural-language conversation via the Interactions API, plus video extension, resolution upscaling and interpolation. Status ga on two first-party statements read 2026-09-01 — the model page Versions block reads 'Stable: gemini-omni-1.1-flash' (with 'Preview: gemini-omni-flash-preview' listed separately as the sibling endpoint), and the pricing page calls it 'now generally available to developers on the paid tier of the Gemini API'. Paid tier only: Google states 'Not available' on the free tier for both axes. Input $1.50 per 1M, one rate that the pricing page labels text / image / video / audio, although the model page spec table lists only Text, Image and Video (up to 10s for editing and extension) as supported inputs — that disagreement between two Google surfaces is recorded here rather than resolved. Two output rates: text and thinking tokens $9.00/1M, video tokens $17.50/1M — price_output_per_mtok holds the video rate, the model's product, the same convention gemini-3.1-flash-image uses for its image rate. Google states billing is on total output token consumption at 5,792 tokens per second of 720p video, an effective ~$0.10 per second under Standard pricing. Context window 1,048,576 tokens; output is video (3s-10s at 360p/720p/1080p/4K, 24 FPS) so there is no published max output token count and the field is null. No context caching rate is published. Released August 27, 2026 per Google's deprecations table, which states 'No shutdown date announced' and does NOT mark the row gray — Google is still asserting the endpoint runs. Added 2026-09-01: this row was missing from the catalog since launch, one of four Google models that check-google-prices.mjs could not report under 'on the page, not in the catalog' because every one of its price cells holds more than one rate (LEARNINGS #105).
Gemini Omni Flash by Google costs $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).
Last verified: 26 Sep 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
Source: Google official documentation ↗
Gemini Omni Flash — questions & answers
How much does Gemini Omni Flash cost?
Gemini Omni Flash is priced at $1.50 per 1M input tokens and $17.50 per 1M output tokens ($5.50/1M blended at a 3:1 input-to-output ratio).
What is the context window of Gemini Omni Flash?
Gemini Omni Flash has a 1.048.576-token context window (1.0M).
Is Gemini Omni Flash deprecated?
No — Gemini Omni Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.
Does Gemini Omni Flash support image or vision input?
Yes — Gemini Omni Flash accepts image input. Listed modalities: text, image, video, video-out.
Can I self-host Gemini Omni Flash?
No — Gemini Omni Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.
What is the API model string for Gemini Omni Flash?
The API model identifier for Gemini Omni Flash is "gemini-omni-1.1-flash" when calling Google's API.
Track Gemini Omni Flash price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.