Price changes, past and scheduled
Providers almost never announce a price rise loudly. It arrives as a footnote, as an introductory rate quietly expiring, or as a billing restructure. This page collects both what has already changed and what is already scheduled, so a cost model built today does not silently go stale.
Already scheduled
Already in effect
Retirements are repricing events too
When a model retires you are moved onto its replacement, which is usually priced differently. A shutdown date is therefore a scheduled repricing event for anyone still on the old model — and unlike a price change, it breaks outright rather than costing more. The next confirmed shutdowns:
| Date | Model | Provider | Replacement |
|---|---|---|---|
| 2026-08-31 | gemini-robotics-er-1.6-preview | gemini-robotics-er-2-preview | |
| 2026-09-24 | sora-2 | OpenAI | — |
| 2026-09-24 | sora-2-pro | OpenAI | — |
| 2026-09-28 | gpt-3.5-turbo-instruct | OpenAI | gpt-5.6-terra |
| 2026-09-28 | gpt-3.5-turbo-1106 | OpenAI | gpt-5.6-terra |
| 2026-09-28 | babbage-002 | OpenAI | gpt-5.6-terra |
| 2026-09-28 | davinci-002 | OpenAI | gpt-5.6-terra |
| 2026-10-23 | gpt-3.5-turbo-0125 | OpenAI | gpt-5.6-terra |
| 2026-10-23 | gpt-4-0613 | OpenAI | gpt-5.6-sol |
| 2026-10-23 | gpt-4-1106-preview | OpenAI | gpt-5.6-sol |
| 2026-10-23 | gpt-4-turbo | OpenAI | gpt-5.6-sol |
| 2026-10-23 | gpt-4.1-nano | OpenAI | gpt-5.6-luna |
Get told when this changes
Subscribe to /feed.xml in any
RSS reader. It carries every price change and every confirmed model shutdown as it is recorded —
no account, no email, nothing to unsubscribe from.
A price rise you find out about from your invoice has already cost you a month.
Frequently asked questions
Which LLM API prices are going up?
Gemini 3.6 Flash and Gemini 3.7 Flash double on 1 January 2027 — input goes from $0.75 to $1.50 and output from $3.75 to $7.50 per million tokens, per a footnote on Google’s own pricing page stating current rates are valid only through 31 December 2026. DeepSeek already raised V4 sharply on 16 August 2026 when it moved to peak/off-peak billing: V4 Flash output went from a flat $0.28 to $1.32 at peak, a 4.7x increase.
Did Claude Sonnet 5 get more expensive?
No — the opposite. The $2/$10 per million token rate was announced as introductory pricing through 31 August 2026, with an increase to $3/$15 scheduled for 1 September. Anthropic subsequently confirmed that increase will not happen and $2/$10 is now the standard price. Teams that budgeted for $3/$15 can revise downward.
How do I avoid being surprised by a price change?
Providers rarely announce increases prominently; they appear as footnotes on pricing pages, as introductory-rate expiry dates, or as billing restructures rather than headline rises. The practical defences are to re-read provider pricing pages on a schedule rather than at integration time, to record the date your cost model was built, and to treat any rate labelled introductory or promotional as temporary until proven otherwise.
Are model retirements a cost issue too?
Effectively yes. When a model retires you are forced onto its replacement, which is frequently priced differently, so a retirement date is a scheduled repricing event for anyone still on the old model. Retirements also break outright rather than degrading — a call to a retired model fails, it does not fall back.
Pricing verified 2026-08-18, retirements verified 2026-08-18, both against official provider documentation. This log records only changes a provider has published — we do not forecast.