GPT-6 Luna

GA

Released 2026-09-22 per OpenAI's own API changelog ("Released GPT-6 Sol (gpt-6-sol) and GPT-6 Luna (gpt-6-luna). These reasoning models accept text and image inputs and generate text", tagged v1/responses and v1/chat/completions). The model page calls it "our most efficient model for focused, high-volume tasks". Text+image input, text output; 1,050,000-token context, 128,000 max output. reasoning.effort supports none/low/medium (default)/high/xhigh/max; use the Responses API for built-in tools and function calling - Chat Completions supports function calling only with reasoning_effort none. Cached input is 10% of uncached input and cache writes are 1.25x it. Prompts >272K input tokens are billed at 2x input and cache rates and 1.5x output for the full request; the price fields carry the <=272K band. Batch and Flex are 50% of Standard; Fast mode is 2x the applicable rates; regional processing adds a 10% premium where available, and EU data residency is available only with Standard processing. Standard tier per 1M: $0.10 in / $0.01 cached in / $0.125 cache write / $0.50 out, stated identically on the model page, the pricing page and the changelog entry. May 18 2026 knowledge cutoff. OpenAI states no relationship between this model and gpt-5.6-luna - no deprecation of that row, no migration pointer - so none is recorded. STATUS: recorded as ga. OpenAI's changelog says 'Released' and the model page carries no access qualifier - unlike GPT-6 Astra's launch page on 2026-09-03, which carried a staged Trusted Access rollout sentence and was held at preview for that reason. OpenAI does not use the words 'generally available' here either, so this is the same positive-content reading the Astra row settled on, not a quoted GA. Lifecycle fields are null: OpenAI states no deprecation for this model.

GPT-6 Luna by OpenAI costs $0.1 per 1M input tokens and $0.5 per 1M output tokens ($0.2/1M blended), with a 1.1M (1.050.000-token) context window. It is generally available (GA), with a knowledge cutoff of 18 May 2026.

Last verified: 26 Sep 2026 · sourced from official provider documentation

Provider
OpenAI
Status
GA
Input price
$0.1 / 1M tokens
Output price
$0.5 / 1M tokens
Cached input
$0.01 / 1M tokens
Blended price
$0.2 / 1M tokens
Context window
1.050.000 tokens (1.1M)
Max output
128.000 tokens
Modality
text, image
Open weights
No — API only
Knowledge cutoff
2026-05-18
Released
22 Sep 2026
API string
gpt-6-luna

Source: OpenAI official documentation ↗

FAQ

GPT-6 Luna — questions & answers

How much does GPT-6 Luna cost?

GPT-6 Luna is priced at $0.1 per 1M input tokens and $0.5 per 1M output tokens ($0.2/1M blended at a 3:1 input-to-output ratio), with cached input at $0.01 per 1M tokens.

What is the context window of GPT-6 Luna?

GPT-6 Luna has a 1.050.000-token context window (1.1M), with up to 128.000 output tokens per request.

Is GPT-6 Luna deprecated?

No — GPT-6 Luna is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does GPT-6 Luna support image or vision input?

Yes — GPT-6 Luna accepts image input. Listed modalities: text, image.

Can I self-host GPT-6 Luna?

No — GPT-6 Luna is proprietary. Its weights are not publicly released, so it is available only through OpenAI's API (or hosted partners), not for self-hosting.

What is the API model string for GPT-6 Luna?

The API model identifier for GPT-6 Luna is "gpt-6-luna" when calling OpenAI's API.

What is GPT-6 Luna's knowledge cutoff?

GPT-6 Luna's training knowledge cutoff is 18 May 2026.