Migrate off GPT-4.1 nano
DeprecatedGPT-4.1 nano is deprecated and shuts down on 23 Oct 2026. OpenAI's official replacement is GPT-5.6 Luna. The cheapest comparable model we track is Llama 4 Scout (17B-16E Instruct) (Meta) at $0.15 per 1M tokens blended.
Last verified: 26 Sep 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
GPT-4.1 nano → GPT-5.6 Luna
Moving to GPT-5.6 Luna means 2.6× more expensive ($0.175 → $0.45 /1M blended) and a larger context window (1.0M → 1.1M).
| Spec | GPT-4.1 nano | GPT-5.6 Luna |
|---|---|---|
| Status | Deprecated | GA |
| Blended price /1M | $0.175 | $0.45 |
| Input / output /1M | $0.1 / $0.4 | $0.2 / $1.20 |
| Context window | 1.0M | 1.1M |
| Max output | 32.768 | 128.000 |
| Modality | text, image | text, image |
| Knowledge cutoff | 1 Jun 2024 | 16 Feb 2026 |
API string was gpt-4.1-nano-2025-04-14 → switch to gpt-5.6-luna. See GPT-5.6 Luna's full spec sheet →
Budget-friendly replacements for GPT-4.1 nano
Generally-available models that match GPT-4.1 nano's capabilities (and at least half its 1.0M context), ranked by blended price. Every number is computed from official catalog prices.
| Model | Blended /1M | Input / output | Context |
|---|---|---|---|
| Llama 4 Scout (17B-16E Instruct) Meta | $0.15 | $0.1 / $0.3 | 10M |
| Qwen3.5-Flash Alibaba | $0.175 | $0.1 / $0.4 | 1M |
| GPT-6 Luna OpenAI | $0.2 | $0.1 / $0.5 | 1.1M |
Migrating off GPT-4.1 nano
What replaces GPT-4.1 nano?
OpenAI's recommended migration for GPT-4.1 nano is GPT-5.6 Luna. The cheapest comparable model in our catalog is Llama 4 Scout (17B-16E Instruct) (Meta) at $0.15/1M blended.
When is GPT-4.1 nano shut down?
GPT-4.1 nano is deprecated and scheduled to shut down on 23 Oct 2026 (deprecated 22 Apr 2026). Migrate before that date to avoid a 404.
What is the cheapest alternative to GPT-4.1 nano?
Llama 4 Scout (17B-16E Instruct) (Meta) is the cheapest generally-available alternative we track that matches GPT-4.1 nano's capabilities, at $0.1 per 1M input and $0.3 per 1M output tokens ($0.15/1M blended).
Browse the full deprecation watch → or see GPT-4.1 nano's full spec sheet →
Get a heads-up before your model is retired
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.