Signal log

Changelog

Every launch, price move, deprecation and retirement we track — newest first.

5 Aug 2026 · retirement · Anthropic

Anthropic's deprecations table now states Current state = Retired for claude-opus-4-1-20250805, with Deprecated June 5, 2026 and Retirement August 5, 2026 — exactly the date announced two months earlier. Calls to claude-opus-4-1-20250805 (and the claude-opus-4-1 alias) no longer serve; Anthropic's recommended replacement is claude-opus-4-8. Opus 4.1 was $15 in / $75 out per 1M tokens, 200K context, 32K max output. Caught by scripts/check-lifecycle-drift.mjs on 2026-08-06 — the provider's declared status flipped while our catalog still read 'deprecated'. Worth noting for anyone building on rate data: this event is invisible to a price diff, because the price never changed — the endpoint simply stopped existing.

source ↗
31 Jul 2026 · deprecation · Google

Google's deprecations table now names gemini-3.6-flash — not gemini-3.5-flash — as the recommended replacement for gemini-2.5-flash (shutdown 2026-10-16), gemini-2.5-flash-preview-05-20, gemini-2.5-flash-preview-09-25, and the whole gemini-2.0-flash line (gemini-2.0-flash and gemini-2.0-flash-001, both shut down 2026-06-01). The new target is CHEAPER on output than the old one: Gemini 3.6 Flash is $1.50 in / $7.50 out per 1M (context caching $0.15) against Gemini 3.5 Flash's $1.50 / $9.00 at the same input price, so the two Flash rows are not a straight ladder. Gemini 3.6 Flash is GA/stable, 1,048,576-token context, 65,536 max output, text+image+video+audio+PDF input; Google publishes no knowledge cutoff or release date for it. Catalog corrected and gemini-3.6-flash added the same day.

source ↗
30 Jul 2026 · launch · Google

Google released two new embodied-reasoning endpoints for robotics in public preview: gemini-robotics-er-2-preview (advanced spatial reasoning, agentic code execution, multi-step tool orchestration, video moment finding, progress classification, multi-robot coordination) and gemini-robotics-er-2-streaming-preview (real-time text streaming over the Live API for low-latency robot agents with bidirectional audio and video input). Both accept text, image, video and audio input, output text, and carry 131,072-token input / 65,536-token output limits. Paid-tier standard rates are $2.00/M in and $10.00/M out for both — 2x the input and 2x the output of Gemini Robotics-ER 1.6 ($1.00/$5.00). Only the non-streaming endpoint publishes a context-caching rate ($0.20/1M) and a batch tier ($1.00/$5.00). Google's own upgrade guidance is to replace model="gemini-robotics-er-1.6-preview" with either ER 2 endpoint; ER 1.6 is shut down 2026-08-31.

source ↗
21 Jul 2026 · launch · Google

Google shipped stable, production-ready Gemini 3.6 Flash and Gemini 3.5 Flash-Lite on the same day. Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite) is billed at $0.30/M in and $2.50/M out on the paid tier, context caching $0.03/M, batch $0.15/$1.25 — against gemini-3.1-flash-lite's $0.25/$1.50, i.e. the recommended migration target is more expensive on both sides, unusually for a Flash-Lite step. One offsetting simplification: 3.5 Flash-Lite publishes a SINGLE input rate covering text, image, video and audio, where 3.1 and 2.5 Flash-Lite both charged a higher audio-input tier. 1,048,576-token context, 65,536 max output, text/image/video/audio/PDF in, text out; computer use in preview. Google's deprecations table names it the recommended replacement for gemini-3.1-flash-lite, which shuts down 2027-05-07. No shutdown date announced for 3.5 Flash-Lite itself.

source ↗
9 Jul 2026 · launch · OpenAI

OpenAI's GPT-5.6 generation went GA on the API + Codex on 2026-07-09 (preview from 06-26), in three durable capability tiers: Sol (flagship) $5 in / $30 out, Terra (balanced) $2.50 / $15, Luna (fast, high-volume) $1 / $6 — all per 1M tokens, cached input $0.50 / $0.25 / $0.10. Every tier has a 1.05M-token context window, 128K max output, text + image input, and a 2026-02-16 knowledge cutoff. Sol matches GPT-5.5's price but is stronger on coding, knowledge work, cybersecurity and science. New caching model: explicit cache breakpoints, 30-min minimum cache life, cache writes billed 1.25x uncached input.

source ↗
8 Jul 2026 · launch · xAI

xAI's smartest model for chat, coding, agentic and knowledge work. Base pricing $2 input / $6 output per 1M tokens (cached input $0.50), tiered 2x above a 200K-token prompt. 500K context window (down from Grok 4.3's 1M), text + image input. Public API access opened 2026-07-09; aliases grok-4.5-latest / grok-build-latest.

source ↗
30 Jun 2026 · launch · Anthropic

Anthropic's most agentic Sonnet, with performance close to Opus 4.8 at lower cost. Standard pricing $3 input / $15 output per 1M tokens, with introductory pricing of $2/$10 in effect through 2026-08-31. 1M context window, 128k max output, Jan 2026 knowledge cutoff. Supersedes Claude Sonnet 4.6, now a legacy model.

source ↗
30 Jun 2026 · launch · Mistral

Mistral released Leanstral 1.5, a specialist model for Lean 4 formal proof engineering, automated theorem proving and autoformalization (not a general chat model). Sparse MoE 119B total / 6.5B active, 256K context, 128K max output, text + image input. Open weights under Apache 2.0. Offered FREE in Mistral's experimental Labs tier ($0/M in and out), with a stated retirement of 2026-09-30 — the catalog's first $0 model, recorded as preview (Labs/experimental) so it stays out of the GA 'cheapest' rankings.

source ↗
28 Jun 2026 · deprecation · Anthropic

Migrate to claude-mythos-5.

source ↗
28 Jun 2026 · deprecation · Google

Migrate to gemini-3.1-pro-preview.

source ↗
28 Jun 2026 · deprecation · Google

Migrate to gemini-3.5-flash.

source ↗
28 Jun 2026 · deprecation · Google

Migrate to gemini-3.1-flash-lite.

source ↗
28 Jun 2026 · retirement · Google

Migrate to gemini-3.5-flash.

source ↗
28 Jun 2026 · retirement · Google

Migrate to gemini-3.1-flash-lite.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-4.3.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-build-0.1.

source ↗
28 Jun 2026 · retirement · xAI

Migrate to grok-imagine-image-quality.

source ↗
28 Jun 2026 · deprecation · DeepSeek

Migrate to deepseek-v4-flash.

source ↗
28 Jun 2026 · deprecation · DeepSeek

Migrate to deepseek-v4-flash.

source ↗
28 Jun 2026 · deprecation · Amazon

Migrate to amazon.nova-2-lite-v1:0.

source ↗
28 Jun 2026 · deprecation · Alibaba

Migrate to qwen3.7-max.

source ↗
28 Jun 2026 · deprecation · Alibaba

Migrate to qwen3.7-max.

source ↗
28 Jun 2026 · deprecation · Alibaba

Migrate to qwen-flash.

source ↗
28 Jun 2026 · deprecation · Alibaba

Migrate to qwen3.7-plus.

source ↗
28 Jun 2026 · retirement · Moonshot

Migrate to kimi-k2.6.

source ↗
28 Jun 2026 · retirement · Moonshot

Migrate to kimi-k2.6.

source ↗
28 Jun 2026 · update
AI Model Watch goes live

Tracking 181 models across 12 providers.

12 Jun 2026 · launch · Moonshot

$0.95/$4 per 1M tokens · 262K context

source ↗
12 Jun 2026 · launch · Moonshot

$1.9/$8 per 1M tokens · 262K context

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.5.

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.4-mini.

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.4-nano.

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.5-pro.

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.5.

source ↗
11 Jun 2026 · deprecation · OpenAI

Migrate to gpt-5.5-pro.

source ↗
9 Jun 2026 · launch · Anthropic

$10/$50 per 1M tokens · 1000K context

source ↗
9 Jun 2026 · launch · Anthropic

$10/$50 per 1M tokens · 1000K context

source ↗
5 Jun 2026 · deprecation · Anthropic

Migrate to claude-opus-4-8.

source ↗
1 Jun 2026 · launch · Alibaba

$0.4/$1.6 per 1M tokens · 1000K context

source ↗
1 Jun 2026 · launch · Alibaba

$0.4/$1.6 per 1M tokens · 1000K context

source ↗
25 May 2026 · launch · Alibaba

$2.5/$7.5 per 1M tokens · 1000K context

source ↗
25 May 2026 · deprecation · Moonshot

Migrate to kimi-k2.6.

source ↗
25 May 2026 · deprecation · Moonshot

Migrate to kimi-k2.6.

source ↗
25 May 2026 · deprecation · Moonshot

Migrate to kimi-k2.6.

source ↗
25 May 2026 · deprecation · Moonshot

Migrate to kimi-k2.6.

source ↗
25 May 2026 · deprecation · Moonshot

Migrate to kimi-k2.6.

source ↗
22 May 2026 · deprecation · Mistral

Migrate to mistral-medium-latest.

source ↗
22 May 2026 · deprecation · Mistral

Migrate to ministral-3-8b-latest.

source ↗
22 May 2026 · deprecation · Mistral

Mistral deprecated its frontier reasoning model (magistral-medium-2509, $2/$5 per MTok, 128k context). Migrate to Mistral Medium 3.5 — Mistral is folding its specialist reasoning line back into its general-purpose models rather than shipping a Magistral successor.

source ↗
22 May 2026 · deprecation · Mistral

Mistral deprecated its frontier agentic-coding model (devstral-2512, $0.4/$2 per MTok, 256k context, open weights under a Modified MIT licence). Migrate to Mistral Medium 3.5 — the specialist coding line is being folded back into the general-purpose models.

source ↗
21 May 2026 · launch · Alibaba

$2.5/$7.5 per 1M tokens · 1000K context

source ↗
19 May 2026 · launch · Google

$1.5/$9 per 1M tokens · 1049K context

source ↗
7 May 2026 · launch · Google

$0.25/$1.5 per 1M tokens · 1049K context

source ↗
May 2026 · launch · Cohere

128K context

source ↗
30 Apr 2026 · launch · xAI

$1.25/$2.5 per 1M tokens · 1000K context

source ↗
30 Apr 2026 · deprecation · Mistral

Migrate to mistral-small-2603.

source ↗
30 Apr 2026 · deprecation · Mistral

Mistral deprecated its small reasoning model (magistral-small-2509, $0.5/$1.5 per MTok, 128k context, Apache-2.0 weights, 24B). Migrate to Mistral Small 4.

source ↗
24 Apr 2026 · launch · DeepSeek

$0.14/$0.28 per 1M tokens · 1000K context

source ↗
24 Apr 2026 · launch · DeepSeek

$0.435/$0.87 per 1M tokens · 1000K context

source ↗
23 Apr 2026 · launch · OpenAI

$5/$30 per 1M tokens · 1050K context

source ↗
23 Apr 2026 · launch · OpenAI

$30/$180 per 1M tokens · 1050K context

source ↗
22 Apr 2026 · deprecation · OpenAI

Migrate to gpt-5.5-pro.

source ↗
22 Apr 2026 · deprecation · OpenAI

Migrate to gpt-5.4-mini.

source ↗
22 Apr 2026 · deprecation · OpenAI

Migrate to gpt-5.5.

source ↗
21 Apr 2026 · launch · OpenAI

$8/$30 per 1M tokens

source ↗
20 Apr 2026 · launch · Moonshot

$0.95/$4 per 1M tokens · 262K context

source ↗
16 Apr 2026 · launch · Alibaba

$0.25/$1.5 per 1M tokens · 1000K context

source ↗
14 Apr 2026 · retirement · Anthropic

Migrate to claude-opus-4-8.

source ↗
14 Apr 2026 · retirement · Anthropic

Migrate to claude-sonnet-4-6.

source ↗
4 Apr 2026 · retirement · Cohere

Migrate to embed-v4.0.

source ↗
4 Apr 2026 · retirement · Cohere

Migrate to embed-v4.0.

source ↗
4 Apr 2026 · retirement · Cohere

Migrate to embed-v4.0.

source ↗
2 Apr 2026 · launch · Alibaba

$0.5/$3 per 1M tokens · 1000K context

source ↗
24 Mar 2026 · deprecation · OpenAI

Video generation (Videos API). Priced per second, not per token: sora-2 $0.10/s (720p, standard) / $0.05/s batch. Videos API + sora-2* announced for discontinua

source ↗
24 Mar 2026 · deprecation · OpenAI

Higher-quality video model. Per-second pricing: $0.30/s (720p) up to $0.70/s (1080p) standard; batch half. Discontinued with Videos API, shuts down 2026-09-24,

source ↗
27 Feb 2026 · deprecation · Mistral

Migrate to mistral-medium-latest.

source ↗
27 Feb 2026 · deprecation · Mistral

Migrate to mistral-medium-latest.

source ↗
27 Feb 2026 · deprecation · Mistral

Mistral deprecated its small agentic-coding model (labs-devstral-small-2512, $0.1/$0.3 per MTok, 256k context, Apache-2.0 weights, 24B). Migrate to Mistral Medium 3.5. Its stated 2026-03-31 retirement has passed, but Mistral still lists the model as deprecated rather than retired and still prices it publicly.

source ↗
19 Feb 2026 · retirement · Anthropic

Migrate to claude-haiku-4-5-20251001.

source ↗
19 Dec 2025 · retirement · Anthropic

Migrate to claude-haiku-4-5-20251001.

source ↗
15 Dec 2025 · retirement · Perplexity

Migrate to sonar-reasoning-pro.

source ↗
2 Dec 2025 · deprecation · Mistral

Migrate to ministral-3-14b-latest.

source ↗
2 Dec 2025 · deprecation · Mistral

Migrate to ministral-3-8b-latest.

source ↗
2 Dec 2025 · deprecation · Mistral

Migrate to ministral-3-3b-latest.

source ↗
14 Nov 2025 · deprecation · OpenAI

Migrate to gpt-image-2.

source ↗
14 Nov 2025 · retirement · OpenAI

Migrate to gpt-image-2.

source ↗
6 Nov 2025 · retirement · Mistral

Migrate to codestral-2508.

source ↗
28 Oct 2025 · retirement · Anthropic

Migrate to claude-sonnet-4-6.

source ↗
15 Sep 2025 · deprecation · Cohere

Migrate to command-r-plus-08-2024.

source ↗
15 Sep 2025 · deprecation · Cohere

Migrate to command-r-08-2024.

source ↗
15 Sep 2025 · deprecation · Cohere

Migrate to command-r-08-2024.

source ↗
15 Sep 2025 · deprecation · Cohere

Migrate to command-r-08-2024.

source ↗
15 Sep 2025 · retirement · Cohere

Migrate to Use the /chat endpoint.

source ↗
2 Dec 2024 · retirement · Cohere

Migrate to rerank-v3.5.

source ↗
2 Dec 2024 · retirement · Cohere

Migrate to rerank-v3.5.

source ↗
30 Nov 2024 · retirement · Mistral

Migrate to mistral-large-2512.

source ↗
30 Nov 2024 · retirement · Mistral

Migrate to mistral-small-2603.

source ↗
30 Nov 2024 · retirement · Mistral

Migrate to mistral-small-2603.

source ↗