Lab / Analyses/data as of 2026-10-08, refreshed daily
Last change to the index, recorded 8 October 2026: 1 new model (claude-haiku-5-5); 1 price cut (claude-sonnet-5-5 cache read $0.20 → $0.10).
What 108 AI models cost per million tokens
The AI Price Index tracks 130 models from 11 providers, with a source link on every price. 108 are on a price list today and 22 have been withdrawn. This page charts the 108, checks them against two other public trackers, and computes every figure from pinned snapshots that refresh daily.
6,000×
From the cheapest output price to the priciest: ministral-3-3b at $0.10 per million tokens, o1-pro at $600. That gap rests on one model: the next is $180, 1,800× the cheapest. o1-pro is a March 2025 model still on OpenAI's price list. Among the 67 models first recorded in 2026, the top price is $180, and $50 without a "-pro" suffix. All 8 models above $50 are "-pro" versions or were first recorded before 2026.
first recorded in 2026earliermiddle 80%
Grey dots were first recorded before 2026. 67 of the 108 are 2026 models, in blue.
The spread
Inside one price list: OpenAI spans 1,200×
One dot per model, by provider, on the same log scale; each gridline is ten times the last. The short vertical tick is each provider's median. OpenAI has the widest range, $0.50 to $600, a 1,200× spread inside one price list.
- $10 of output from ministral-3-3b
- 100M
- $10 of output at the median price
- 2.1M
- $10 of output from o1-pro
- 16,667
The extremes
The ten most expensive come from two providers
Current price per million output tokens. The top ten: OpenAI 7, Anthropic 3. None of the ten cheapest comes from OpenAI, Anthropic, Google and xAI.
Most expensive · $50 to $600 per 1M
Cheapest · $0.10 to $0.40 per 1M
On the most-expensive chart's scale, the longest of these bars would be under a pixel wide.
AnthropicOpenAIGoogleevery other providerThe month beside each model is when the index first recorded its price.
The output rule
Output costs 5× input, and Anthropic never deviates
Divide each model's output price by its input price and the median is 5×; 34 of 108 models sit exactly there. 19 of those 34 are Anthropic: every one of its 19 models charges exactly 5× input. xAI prices all 8 of its models at 2× or 3×. The six models that charge the same for output as input are all Ministral and Llama models. The three Llama prices come from Together AI, not from Meta, and the dataset flags them for review.
The fine print
The multipliers behind a bill
The headline price is rarely the whole bill. These are the rules the other price types follow, each one checked against every model that lists it.
A "-pro" version
6× to 12× the base model, on input and output alike
gpt-5.2-pro 12×, gpt-5.4-pro 12×, gpt-5.5-pro 6×, o1-pro 10×, o3-pro 10×
Long-context tier
varies on input; 1.5× on output (Google and OpenAI), 2× on output (xAI), 5× on output (Anthropic)
all 22 models that list one
Writing to a prompt cache
1.25× input; Anthropic's one-hour cache 2×
38 standard cache-write prices and 19 one-hour ones; none at or below the input price
Reading from a prompt cache
1/10 of input on 60 of 83 models; 1/50 to 1/2 overall
12 of the 60 are Alibaba rates derived from its published 10% rule; 4 OpenAI models charge half; deepseek-flash charges the least
| Provider | Models | Output ÷ input | Cache read | Cache write | Long context |
|---|---|---|---|---|---|
| OpenAI | 28 | 5× (4–8×) | 10% (5–50%) | 1.25× | in 2× · out 1.5× |
| Anthropic | 19 | 5× | 10% (2.5–10%) | 1.25× · 1 h 2× | in 5× · out 5× |
| Alibaba | 17 | 4.33× (3–8×) | 10% | 1.25× | — |
| 11 | 6× (4–8.33×) | 10% | — | in 2× · out 1.5× | |
| Mistral | 8 | 3× (1–5×) | 10% (10–10.3%) | — | — |
| xAI | 8 | 2× (2–3×) | 16% (15–25%) | — | in 2× · out 2× |
| Amazon | 6 | 4.5× (4–8.33×) | — | — | — |
| Meta | 5 | 1× (1–3.28×) | — | — | — |
| AI21 | 2 | 3× (2–4×) | — | — | — |
| Cohere | 2 | 4× | — | — | — |
| DeepSeek | 2 | 3.5× (3–4×) | 2.7% (2–3.3%) | — | — |
Each ratio is the median across the provider's current models that list that price, with the range in brackets where they differ. Cache read is a share of the input price. Cache write is a multiple of the input price; "1 h" is the one-hour cache. Long context is the long-prompt tier as a multiple of the standard input and output prices. A dash means the provider lists no such price.
Cache reads
Cache reads cost 1/10 of input on 60 of 83 models
Of the 83 models that list a cache-read price, 60 charge 1/10 of input, 12 of them Alibaba rates derived from its published 10% rule. Every cache price on a model first recorded before May 2025 is OpenAI's, and none is below 1/4. Of the models recorded since, 9 charge more than 1/10, and the seven that charge less were all first recorded from June 2026 on. These are today's prices: the index holds no cache-read price older than 31 March 2026, so the chart shows which models charge what now, not how cache prices moved.
Repricing
Ten repricings: seven cuts, three rises
The index records 26 changed price rows, but they belong to ten repricings, each one model on one date. OpenAI's change to gpt-5.6-sol on 22 August 2026 moved eight rates at once. Seven repricings were cuts and three were rises, all three in 2026 from Mistral and DeepSeek. The deepest cut is o3, from $40 to $8 per million output tokens. Four of the ten fall between 13 and 22 August 2026, the only month with more than one. The index first recorded 70 of its 130 models in 2026, so an older repricing is less likely to be on record. Three of the changes need context, and the note under each chart gives it.
price cut price riseEach panel is one model's output price, indexed so its first recorded price is 100, on its own time range. grok-4.5 and claude-sonnet-5-5 changed only other rates, so they have no panel.
New price recorded from 10 June 2025; tracked since 16 April 2025.
A promotion: the pricing page says $3.75 through 31 December 2026 and $7.50 from 1 January 2027.
New price recorded from 6 August 2024; tracked since 13 May 2024.
New price recorded from 22 August 2026; tracked since 9 July 2026.
New price recorded from 3 December 2024; tracked since 4 November 2024.
New price recorded from 24 June 2026; tracked since 16 March 2026.
The off-peak rate after DeepSeek split prices into peak and off-peak; peak is exactly double.
The off-peak rate after DeepSeek split prices into peak and off-peak; peak is exactly double. Renamed deepseek-flash on 10 September 2026, at $0.60.
Retirements
22 models have left the list
22 of the 130 models have left their price lists, 19 of them in 2026. The median one was listed for 444 days; gpt-4 lasted 1,263. Six of them end on the same day, 28 August 2026, because that is when the collector noticed they were gone, not when they went. Leaving the list is not the same as being switched off: gpt-4 and gpt-4-turbo are gone from OpenAI's price list, but the dataset's notes give both a scheduled retirement date of 23 October 2026.
Cross-check
Two other trackers agree on 269 of 276 shared prices
Three public datasets are pulled here. The AI Model Price Stability dataset reads prices from web-archive captures of the Anthropic, Google and OpenAI pricing pages, with a link to every capture. LLM Token Price History keeps a current table for 35 providers, resellers included. Where either states a price the index also holds, for the same model and price type, they match 269 times out of 276: 60 of 60 archive prices on the day of the capture, and 209 of 216 current prices across 83 models. Of the 83 shared models, 13 carry a non-current status in LLM Token Price History (8 deprecated, 4 retired, 1 preview) while the index still lists them; the prices agree anyway.
What only the archive caught
The archive also holds 19 price changes the index does not record, 11 of them on models the index tracks. Its authors call these candidates, not verified changes, so each one links to the two captures it rests on.
| Model | Price type | Before | After | Captured | In the index |
|---|---|---|---|---|---|
| claude-3-5-haiku | cache read | $0.03 | $0.025 ↓ | 2024-10-23 → 2024-10-29 | model tracked, change not recorded |
| claude-3-5-haiku | cache write | $0.30 | $0.3125 ↑ | 2024-10-23 → 2024-10-29 | model tracked, change not recorded |
| claude-3-5-haiku | cache read | $0.025 | $0.10 ↑ | 2024-10-29 → 2024-11-04 | model tracked, change not recorded |
| claude-3-5-haiku | cache write | $0.3125 | $1.25 ↑ | 2024-10-29 → 2024-11-04 | model tracked, change not recorded |
| claude-3-5-haikuThe earlier capture shows the model marked "Coming soon" at $0.25 input: a price listed before launch. | input | $0.25 | $1 ↑ | 2024-10-23 → 2024-11-04 | model tracked, change not recorded |
| claude-3-5-haiku | output | $1.25 | $5 ↑ | 2024-10-23 → 2024-11-04 | model tracked, change not recorded |
| claude-3-5-haiku | cache read | $0.10 | $0.08 ↓ | 2024-11-04 → 2024-12-03 | model tracked, change not recorded |
| claude-3-5-haiku | cache write | $1.25 | $1 ↓ | 2024-11-04 → 2024-12-03 | model tracked, change not recorded |
| o3 | cache read | $2.50 | $0.50 ↓ | 2025-04-17 → 2025-06-11 | model tracked, change not recorded |
| gemini-2-5-flash-preview | cache read | $0.0375 | $0.075 ↑ | 2025-05-11 → 2025-09-26 | model not tracked |
| gemini-2-5-flash-preview | input | $0.15 | $0.30 ↑ | 2025-04-18 → 2025-09-26 | model not tracked |
| gemini-2-5-flash | cache read | $0.075 | $0.03 ↓ | 2025-06-17 → 2025-10-09 | model tracked, change not recorded |
| gemini-2-5-flash-preview | cache read | $0.075 | $0.0375 ↓ | 2025-09-26 → 2025-10-09 | model not tracked |
| gemini-2-5-flash-preview | cache read | $0.0375 | $0.03 ↓ | 2025-10-09 → 2025-10-16 | model not tracked |
| gemini-2-5-flash-lite | cache read | $0.025 | $0.01 ↓ | 2025-07-23 → 2025-10-24 | model tracked, change not recorded |
| gemini-2-5-flash-lite-preview | cache read | $0.025 | $0.01 ↓ | 2025-06-17 → 2025-10-24 | model not tracked |
| gemini-3-1-flash-image-preview | input | $0.25 | $0.50 ↑ | 2026-02-26 → 2026-03-03 | model not tracked |
| gemini-robotics-er-preview | input | $0.30 | $1 ↑ | 2025-10-09 → 2026-04-30 | model not tracked |
| gemini-robotics-er-preview | output | $2.50 | $5 ↑ | 2025-10-09 → 2026-04-30 | model not tracked |
Each price links to the web-archive capture it was read from.
Limits
What this data cannot tell you
It is mostly recent.
70 of 130 models are first recorded in 2026, and dates are when a price was seen, not always when it changed. A model repriced before it was tracked shows no change here.
Not every price comes from the provider.
503 rows were read from a live pricing page, 36 from web archives and 2 from changelogs. The index has no first-party API price for Meta's models, so its 10 rows come from Together AI.
Some rows are inferred.
74 are marked that way: 24 are cache prices derived from a published multiplier, 10 are the Together AI rows, and 40 are unconfirmed for other reasons: approximate dates, uncorroborated prices and superseded records, which the notes give for all but 8 (those rows carry no note of their own). 435 are confirmed against the provider, 36 come from web archives, 2 from changelogs, and 4 are retained prices for delisted models.
List price is not what you pay.
Batch discounts, regional pricing and DeepSeek's peak hours all move the bill; the index records 12 kinds of price and this page uses the standard ones. If you run open weights yourself, the cost is memory rather than tokens: see the VRAM and KV-cache calculator.
Every price
All 108 current prices
Price per million tokens. Set the size of a call to see what it costs on each model. Which model is cheap depends on the shape of the call. From a long document (50,000 tokens in, 1,000 out) to a long answer (1,000 in, 8,000 out), 17 of 108 models move ten or more places in the order. llama-3.1-405b moves 35 places towards the cheapest, because it charges the same for output as for input. Its prices are Together AI's, flagged for review in the dataset.
| Model | Provider | First recorded | Input /M | Output /M | This call |
|---|---|---|---|---|---|
| o1-pro | OpenAI | Mar 2025 | $150 | $600 | $2.10 |
| gpt-5.5-pro | OpenAI | Apr 2026 | $30 | $180 | $0.48 |
| gpt-5.4-pro | OpenAI | Mar 2026 | $30 | $180 | $0.48 |
| gpt-5.2-pro | OpenAI | Jan 2026 | $21 | $168 | $0.38 |
| o3-pro | OpenAI | Jun 2025 | $20 | $80 | $0.28 |
| claude-opus-4-20250514 | Anthropic | May 2025 | $15 | $75 | $0.23 |
| claude-opus-4-1-20250805 | Anthropic | Aug 2025 | $15 | $75 | $0.23 |
| o1 | OpenAI | Dec 2024 | $15 | $60 | $0.21 |
| gpt-6-astra | OpenAI | Sep 2026 | $10 | $50 | $0.15 |
| claude-mythos-5-1 | Anthropic | Sep 2026 | $10 | $50 | $0.15 |
| claude-mythos-5 | Anthropic | Jun 2026 | $10 | $50 | $0.15 |
| claude-fable-5-1 | Anthropic | Sep 2026 | $10 | $50 | $0.15 |
| claude-fable-5 | Anthropic | Jun 2026 | $10 | $50 | $0.15 |
| gpt-5.5 | OpenAI | Apr 2026 | $5 | $30 | $0.08 |
| gpt-rosalind-research | OpenAI | Oct 2026 | $5 | $25 | $0.075 |
| claude-opus-5 | Anthropic | Jul 2026 | $5 | $25 | $0.075 |
| claude-opus-4-8 | Anthropic | May 2026 | $5 | $25 | $0.075 |
| claude-opus-4-7 | Anthropic | Apr 2026 | $5 | $25 | $0.075 |
| claude-opus-4-6 | Anthropic | Feb 2026 | $5 | $25 | $0.075 |
| claude-opus-4-5-20251101 | Anthropic | Nov 2025 | $5 | $25 | $0.075 |
| gpt-5.6-sol | OpenAI | Jul 2026 | $4 | $20 | $0.06 |
| claude-opus-5-5 | Anthropic | Sep 2026 | $4 | $20 | $0.06 |
| gpt-5.4 | OpenAI | Mar 2026 | $2.50 | $15 | $0.04 |
| claude-sonnet-4-6 | Anthropic | Feb 2026 | $3 | $15 | $0.045 |
| claude-sonnet-4-5-20250929 | Anthropic | Sep 2025 | $3 | $15 | $0.045 |
| claude-sonnet-4-20250514 | Anthropic | May 2025 | $3 | $15 | $0.045 |
| gpt-5.3-codex | OpenAI | Jul 2026 | $1.75 | $14 | $0.032 |
| gpt-5.2 | OpenAI | Jan 2026 | $1.75 | $14 | $0.032 |
| nova-premier | Amazon | Apr 2025 | $2.50 | $12.50 | $0.037 |
| gpt-5.6-terra | OpenAI | Jul 2026 | $2 | $12 | $0.032 |
| gemini-3.1-pro-preview | Feb 2026 | $2 | $12 | $0.032 | |
| gpt-6.1-sol | OpenAI | Sep 2026 | $2 | $10 | $0.03 |
| gpt-6-sol | OpenAI | Sep 2026 | $2 | $10 | $0.03 |
| gpt-5.1 | OpenAI | Nov 2025 | $1.25 | $10 | $0.022 |
| gpt-5 | OpenAI | Aug 2025 | $1.25 | $10 | $0.022 |
| gpt-4o | OpenAI | May 2024 | $2.50 | $10 | $0.035 |
| gemini-2.5-pro | Jun 2025 | $1.25 | $10 | $0.022 | |
| claude-sonnet-5-5 | Anthropic | Sep 2026 | $2 | $10 | $0.03 |
| claude-sonnet-5 | Anthropic | Jul 2026 | $2 | $10 | $0.03 |
| nova-2-pro | Amazon | Jun 2026 | $1.25 | $10 | $0.022 |
| gemini-3.5-flash | May 2026 | $1.50 | $9 | $0.024 | |
| o3 | OpenAI | Apr 2025 | $2 | $8 | $0.028 |
| gpt-4.1 | OpenAI | Apr 2025 | $2 | $8 | $0.028 |
| jamba-large | AI21 | Jul 2025 | $2 | $8 | $0.028 |
| mistral-medium-3.5 | Mistral | Apr 2026 | $1.50 | $7.50 | $0.022 |
| qwen3.7-max | Alibaba | May 2026 | $2.50 | $7.50 | $0.033 |
| qwen-max | Alibaba | Jan 2025 | $1.60 | $6.40 | $0.022 |
| grok-4.7 | xAI | Sep 2026 | $2 | $6 | $0.026 |
| grok-4.6 | xAI | Aug 2026 | $2 | $6 | $0.026 |
| grok-4.5 | xAI | Jul 2026 | $2 | $6 | $0.026 |
| qwen3.8-max | Alibaba | Aug 2026 | $2 | $6 | $0.026 |
| qwen3-max | Alibaba | Jan 2026 | $1.20 | $6 | $0.018 |
| claude-haiku-4-5-20251001 | Anthropic | Oct 2025 | $1 | $5 | $0.015 |
| qwen3-coder-plus | Alibaba | Jul 2025 | $1 | $5 | $0.015 |
| gpt-5.4-mini | OpenAI | Mar 2026 | $0.75 | $4.50 | $0.012 |
| o4-mini | OpenAI | Apr 2025 | $1.10 | $4.40 | $0.015 |
| o3-mini | OpenAI | Jan 2025 | $1.10 | $4.40 | $0.015 |
| gemini-3.8-flash | Sep 2026 | $0.75 | $3.75 | $0.011 | |
| gemini-3.7-flash | Aug 2026 | $0.75 | $3.75 | $0.011 | |
| gemini-3.6-flash | Jul 2026 | $0.75 | $3.75 | $0.011 | |
| llama-3.1-405b | Meta | Jun 2026 | $3.50 | $3.50 | $0.038 |
| nova-pro | Amazon | Dec 2024 | $0.80 | $3.20 | $0.011 |
| gemini-3-flash-preview | Dec 2025 | $0.50 | $3 | $0.008 | |
| qwen3.6-plus | Alibaba | Apr 2026 | $0.50 | $3 | $0.008 |
| grok-4.3 | xAI | May 2026 | $1.25 | $2.50 | $0.015 |
| grok-4.20-multi-agent-0309 | xAI | Mar 2026 | $1.25 | $2.50 | $0.015 |
| grok-4.20-0309-reasoning | xAI | Mar 2026 | $1.25 | $2.50 | $0.015 |
| grok-4.20-0309-non-reasoning | xAI | Jun 2026 | $1.25 | $2.50 | $0.015 |
| gemini-3.5-flash-lite | Jul 2026 | $0.30 | $2.50 | $0.0055 | |
| gemini-2.5-flash | Jun 2025 | $0.30 | $2.50 | $0.0055 | |
| nova-2-lite | Amazon | Dec 2025 | $0.30 | $2.50 | $0.0055 |
| qwen3.5-plus | Alibaba | Feb 2026 | $0.40 | $2.40 | $0.0064 |
| mistral-large-4 | Mistral | Oct 2026 | $0.68 | $2.09 | $0.0089 |
| grok-build-0.1 | xAI | Jun 2026 | $1 | $2 | $0.012 |
| deepseek-v4-pro | DeepSeek | Jun 2026 | $0.66 | $1.98 | $0.0086 |
| qwen3.7-plus | Alibaba | Jun 2026 | $0.40 | $1.60 | $0.0056 |
| mistral-large-3 | Mistral | Dec 2025 | $0.50 | $1.50 | $0.0065 |
| gemini-3.1-flash-lite | May 2026 | $0.25 | $1.50 | $0.004 | |
| qwen3.6-flash | Alibaba | Apr 2026 | $0.25 | $1.50 | $0.004 |
| qwen3-coder-next | Alibaba | Aug 2026 | $0.30 | $1.50 | $0.0045 |
| qwen3-coder-flash | Alibaba | Aug 2026 | $0.30 | $1.50 | $0.0045 |
| gpt-5.4-nano | OpenAI | Mar 2026 | $0.20 | $1.25 | $0.0032 |
| gpt-5.6-luna | OpenAI | Jul 2026 | $0.20 | $1.20 | $0.0032 |
| qwen-plus | Alibaba | Dec 2025 | $0.40 | $1.20 | $0.0052 |
| llama-3.3-70b | Meta | Jun 2026 | $1.04 | $1.04 | $0.011 |
| codestral-25.08 | Mistral | Jul 2025 | $0.30 | $0.90 | $0.0039 |
| llama-4-maverick | Meta | Apr 2025 | $0.27 | $0.85 | $0.0036 |
| gpt-4o-mini | OpenAI | Jul 2024 | $0.15 | $0.60 | $0.0021 |
| mistral-small-4 | Mistral | Mar 2026 | $0.15 | $0.60 | $0.0021 |
| deepseek-flash | DeepSeek | Sep 2026 | $0.15 | $0.60 | $0.0021 |
| command-r | Cohere | Aug 2024 | $0.15 | $0.60 | $0.0021 |
| llama-4-scout | Meta | Apr 2025 | $0.18 | $0.59 | $0.0024 |
| gpt-6-luna | OpenAI | Sep 2026 | $0.10 | $0.50 | $0.0015 |
| claude-haiku-5-5 | Anthropic | Oct 2026 | $0.10 | $0.50 | $0.0015 |
| qwen3.8-flash | Alibaba | Aug 2026 | $0.15 | $0.47 | $0.002 |
| gemini-2.5-flash-lite | Jun 2025 | $0.10 | $0.40 | $0.0014 | |
| qwen3.5-flash | Alibaba | Feb 2026 | $0.10 | $0.40 | $0.0014 |
| qwen-flash | Alibaba | Jul 2025 | $0.05 | $0.40 | $0.0009 |
| jamba-mini | AI21 | Jan 2026 | $0.20 | $0.40 | $0.0024 |
| nova-lite | Amazon | Dec 2024 | $0.06 | $0.24 | $0.00084 |
| ministral-3-14b | Mistral | Dec 2025 | $0.20 | $0.20 | $0.0022 |
| qwen-turbo | Alibaba | Apr 2025 | $0.05 | $0.20 | $0.0007 |
| llama-3.1-8b | Meta | Jun 2026 | $0.18 | $0.18 | $0.002 |
| ministral-3-8b | Mistral | Dec 2025 | $0.15 | $0.15 | $0.0016 |
| command-r7b | Cohere | Dec 2024 | $0.0375 | $0.15 | $0.00052 |
| nova-micro | Amazon | Dec 2024 | $0.035 | $0.14 | $0.00049 |
| qwen3.7-flash | Alibaba | Jul 2026 | $0.03 | $0.13 | $0.00043 |
| ministral-3-3b | Mistral | Dec 2025 | $0.10 | $0.10 | $0.0011 |
The bar is the output price on the charts' log scale, $0.10 to $1,000; blue is first recorded in 2026.
All 26 changed price rows (10 repricings)
| Model | Price type | Before | After | Recorded from |
|---|---|---|---|---|
| gpt-4o | input | $5 | $2.50 ↓ | 2024-08-06 |
| gpt-4o | output | $15 | $10 ↓ | 2024-08-06 |
| claude-3-5-haiku-20241022 | input | $1 | $0.80 ↓ | 2024-12-03 |
| claude-3-5-haiku-20241022 | output | $5 | $4 ↓ | 2024-12-03 |
| o3 | input | $10 | $2 ↓ | 2025-06-10 |
| o3 | output | $40 | $8 ↓ | 2025-06-10 |
| mistral-small-4 | input | $0.10 | $0.15 ↑ | 2026-06-24 |
| mistral-small-4 | output | $0.30 | $0.60 ↑ | 2026-06-24 |
| grok-4.5 | cache read | $0.50 | $0.30 ↓ | 2026-07-19 |
| gemini-3.6-flash | input | $1.50 | $0.75 ↓ | 2026-08-13 |
| gemini-3.6-flash | output | $7.50 | $3.75 ↓ | 2026-08-13 |
| deepseek-v4-flash | cache read | $0.0028 | $0.007 ↑ | 2026-08-16 |
| deepseek-v4-flash | input | $0.14 | $0.22 ↑ | 2026-08-16 |
| deepseek-v4-flash | output | $0.28 | $0.66 ↑ | 2026-08-16 |
| deepseek-v4-pro | cache read | $0.0036 | $0.022 ↑ | 2026-08-16 |
| deepseek-v4-pro | input | $0.435 | $0.66 ↑ | 2026-08-16 |
| deepseek-v4-pro | output | $0.87 | $1.98 ↑ | 2026-08-16 |
| gpt-5.6-sol | cache read | $0.50 | $0.40 ↓ | 2026-08-22 |
| gpt-5.6-sol | cache write | $6.25 | $5 ↓ | 2026-08-22 |
| gpt-5.6-sol | input | $5 | $4 ↓ | 2026-08-22 |
| gpt-5.6-sol | output | $30 | $20 ↓ | 2026-08-22 |
| gpt-5.6-sol | tier2 cache read | $1 | $0.80 ↓ | 2026-08-22 |
| gpt-5.6-sol | tier2 cache write | $12.50 | $10 ↓ | 2026-08-22 |
| gpt-5.6-sol | tier2 input | $10 | $8 ↓ | 2026-08-22 |
| gpt-5.6-sol | tier2 output | $45 | $30 ↓ | 2026-08-22 |
| claude-sonnet-5-5 | cache read | $0.20 | $0.10 ↓ | 2026-10-08 |
Method
Method and source
Prices are providers' list prices as recorded in the dataset snapshots of 8 October 2026, and providers change them without notice. None of them is a retail price, and nothing on this page is an affiliate link.
- Dataset
- RoninForge/ai-price-index at revision
02fc932d333a, retrieved 2026-10-08; sha256620ad4ed8137…AI Price Index by RoninForge (https://roninforge.org/data/ai-price-index/), CC BY 4.0. - Cross-check
- CurtisPyke/ai-model-price-stability at revision
cd23c3223863, retrieved 2026-09-24; sha25697badbe5cf88…AI Model Price Stability by Curtis Pyke / Kingy AI (DOI 10.5281/zenodo.22288610), CC BY 4.0. Its own limitations, which its licence asks users to preserve: archive coverage is incomplete, detected transitions are candidates rather than verified vendor price changes, and vendor scope and tiers may be ambiguous. - Cross-check
- modelpricewatch/llm-token-price-history at revision
2e95adfec8e7, retrieved 2026-10-04; sha256e77f0a452e74…LLM Token Price History by ModelPriceWatch (https://modelpricewatch.com), CC BY 4.0. - Refresh
- Every day each dataset is checked for a new revision. A new one is validated (columns, units, row counts) and shipped only if this page still builds and its checks pass; the figures above are always computed from the committed snapshot.
- Current price
- A row with no end date, as the dataset defines it. Medians and percentiles are type-7 quantiles, the NumPy and R default.
- First recorded
- The earliest date the index holds a model's input price. For older models it usually comes from an archived pricing page near launch; it is not a published release date.
- Repricing
- Consecutive rows for one model and one price type whose price differs, grouped by model and recorded-from date.