Skip to main content

Lab / Analyses/data as of 2026-10-08, refreshed daily

Last change to the index, recorded 8 October 2026: 1 new model (claude-haiku-5-5); 1 price cut (claude-sonnet-5-5 cache read $0.20 → $0.10).

What 108 AI models cost per million tokens

The AI Price Index tracks 130 models from 11 providers, with a source link on every price. 108 are on a price list today and 22 have been withdrawn. This page charts the 108, checks them against two other public trackers, and computes every figure from pinned snapshots that refresh daily.

6,000×

From the cheapest output price to the priciest: ministral-3-3b at $0.10 per million tokens, o1-pro at $600. That gap rests on one model: the next is $180, 1,800× the cheapest. o1-pro is a March 2025 model still on OpenAI's price list. Among the 67 models first recorded in 2026, the top price is $180, and $50 without a "-pro" suffix. All 8 models above $50 are "-pro" versions or were first recorded before 2026.

first recorded in 2026earliermiddle 80%

middle 80%: $0.40 to $50$50: five 2026 models, the newest top end$0.10$1$10$100$1,000median $4.75ministral-3-3b $0.10o1-pro $600$0.10$1$10$100$1,000median $4.75ministral-3-3b $0.10o1-pro $600

Grey dots were first recorded before 2026. 67 of the 108 are 2026 models, in blue.

The spread

Inside one price list: OpenAI spans 1,200×

One dot per model, by provider, on the same log scale; each gridline is ten times the last. The short vertical tick is each provider's median. OpenAI has the widest range, $0.50 to $600, a 1,200× spread inside one price list.

$10 of output from ministral-3-3b
100M
$10 of output at the median price
2.1M
$10 of output from o1-pro
16,667
$0.10$1$10$100$1,000Cohere2 · med $0.375Mistral8 · med $0.75Meta5 · med $0.85DeepSeek2 · med $1.29Alibaba17 · med $1.50xAI8 · med $2.50Amazon6 · med $2.85Google11 · med $3.75AI212 · med $4.20OpenAI28 · med $11Anthropic19 · med $25$0.10$1$10$100$1,000CohereMistralMetaDeepSeekAlibabaxAIAmazonGoogleAI21OpenAIAnthropic
AnthropicOpenAIGoogleevery other providerprovider median
Each dot is one model. The cloud above a provider's dots is a smoothed density of its prices (a Gaussian kernel 0.15 decades wide), drawn only for providers with five or more models. Providers are ordered by median, cheapest first; median input across all 108: $1.25, median output: $4.75.

The extremes

The ten most expensive come from two providers

Current price per million output tokens. The top ten: OpenAI 7, Anthropic 3. None of the ten cheapest comes from OpenAI, Anthropic, Google and xAI.

Most expensive · $50 to $600 per 1M

$0$200$400$600Mar 2025o1-pro$600Apr 2026gpt-5.5-pro$180Mar 2026gpt-5.4-pro$180Jan 2026gpt-5.2-pro$168Jun 2025o3-pro$80May 2025claude-opus-4$75Aug 2025claude-opus-4-1$75Dec 2024o1$60Sep 2026gpt-6-astra$50Sep 2026claude-mythos-5-1$50$0$200$400$600o1-pro · Mar 2025$600gpt-5.5-pro · Apr 2026$180gpt-5.4-pro · Mar 2026$180gpt-5.2-pro · Jan 2026$168o3-pro · Jun 2025$80claude-opus-4 · May 2025$75claude-opus-4-1 · Aug 2025$75o1 · Dec 2024$60gpt-6-astra · Sep 2026$50claude-mythos-5-1 · Sep 2026$50

Cheapest · $0.10 to $0.40 per 1M

$0$0.10$0.20$0.30$0.40Dec 2025ministral-3-3b$0.10Jul 2026qwen3.7-flash$0.13Dec 2024nova-micro$0.14Dec 2024command-r7b$0.15Dec 2025ministral-3-8b$0.15Jun 2026llama-3.1-8b$0.18Apr 2025qwen-turbo$0.20Dec 2025ministral-3-14b$0.20Dec 2024nova-lite$0.24Jan 2026jamba-mini$0.40$0$0.10$0.20$0.30$0.40ministral-3-3b · Dec 2025$0.10qwen3.7-flash · Jul 2026$0.13nova-micro · Dec 2024$0.14command-r7b · Dec 2024$0.15ministral-3-8b · Dec 2025$0.15llama-3.1-8b · Jun 2026$0.18qwen-turbo · Apr 2025$0.20ministral-3-14b · Dec 2025$0.20nova-lite · Dec 2024$0.24jamba-mini · Jan 2026$0.40

On the most-expensive chart's scale, the longest of these bars would be under a pixel wide.

AnthropicOpenAIGoogleevery other providerThe month beside each model is when the index first recorded its price.

The output rule

Output costs 5× input, and Anthropic never deviates

Divide each model's output price by its input price and the median is 5×; 34 of 108 models sit exactly there. 19 of those 34 are Anthropic: every one of its 19 models charges exactly 5× input. xAI prices all 8 of its models at 2× or 3×. The six models that charge the same for output as input are all Ministral and Llama models. The three Llama prices come from Together AI, not from Meta, and the dataset flags them for review.

$0.01$0.10$1$10$100$1,000$0.10$1$10$100$1,000input price per 1M →↑ output price per 1M1×2×4×8×5×: 34 models, 19 of them Anthropic's$0.01$1$100$0.10$1$10$100$1,0005×
AnthropicOpenAIGoogleevery other provider
Each mark is a model at its input price (across) and output price (up); marks grow where several models share both prices. A diagonal is a fixed multiple: output = k × input (faint lines mark 3× and 6×). Models per multiple: 1× 6 · 2× 6 · 3× 9 · 4× 22 · 5× 34 · 6× 14 · 8× 8 · other 9.

The fine print

The multipliers behind a bill

The headline price is rarely the whole bill. These are the rules the other price types follow, each one checked against every model that lists it.

A "-pro" version

6× to 12× the base model, on input and output alike

gpt-5.2-pro 12×, gpt-5.4-pro 12×, gpt-5.5-pro 6×, o1-pro 10×, o3-pro 10×

Long-context tier

varies on input; 1.5× on output (Google and OpenAI), 2× on output (xAI), 5× on output (Anthropic)

all 22 models that list one

Writing to a prompt cache

1.25× input; Anthropic's one-hour cache 2×

38 standard cache-write prices and 19 one-hour ones; none at or below the input price

Reading from a prompt cache

1/10 of input on 60 of 83 models; 1/50 to 1/2 overall

12 of the 60 are Alibaba rates derived from its published 10% rule; 4 OpenAI models charge half; deepseek-flash charges the least

Multipliers by provider: median across each provider's current models, range in brackets
ProviderModelsOutput ÷ inputCache readCache writeLong context
OpenAI285× (4–8×)10% (5–50%)1.25×in 2× · out 1.5×
Anthropic195×10% (2.5–10%)1.25× · 1 h 2×in 5× · out 5×
Alibaba174.33× (3–8×)10%1.25×—
Google116× (4–8.33×)10%—in 2× · out 1.5×
Mistral83× (1–5×)10% (10–10.3%)——
xAI82× (2–3×)16% (15–25%)—in 2× · out 2×
Amazon64.5× (4–8.33×)———
Meta51× (1–3.28×)———
AI2123× (2–4×)———
Cohere24×———
DeepSeek23.5× (3–4×)2.7% (2–3.3%)——

Each ratio is the median across the provider's current models that list that price, with the range in brackets where they differ. Cache read is a share of the input price. Cache write is a multiple of the input price; "1 h" is the one-hour cache. Long context is the long-prompt tier as a multiple of the standard input and output prices. A dash means the provider lists no such price.

Cache reads

Cache reads cost 1/10 of input on 60 of 83 models

Of the 83 models that list a cache-read price, 60 charge 1/10 of input, 12 of them Alibaba rates derived from its published 10% rule. Every cache price on a model first recorded before May 2025 is OpenAI's, and none is below 1/4. Of the models recorded since, 9 charge more than 1/10, and the seven that charge less were all first recorded from June 2026 on. These are today's prices: the index holds no cache-read price older than 31 March 2026, so the chart shows which models charge what now, not how cache prices moved.

1/21/41/101/201/50202420252026OpenAI, 4 modelsOpenAI, 3 modelsxAIMistraldeepseek-v4-pro 1/30claude-fable-5-1, claude-mythos-5-1 1/40deepseek-flash 1/50claude-opus-5-5 1/20claude-sonnet-5-5 1/20gpt-6.1-sol 1/201/21/41/101/201/50202420252026OpenAI, 4 modelsOpenAI, 3 modelsxAIMistraldeepseek-flash 1/50
AnthropicOpenAIGoogleevery other providerhollow: Alibaba's rates, derived from its published 10% rule
Across: the date the index first recorded the model's input price, not a launch date and not the date of its cache price. Up: today's cache-read price ÷ input price, on a log scale. The 12 hollow dots share one observation date, 7 September 2026, and each sits at its model's date. Amazon, Cohere, Meta and AI21 list no cache-read price, so they are not on this chart.

Repricing

Ten repricings: seven cuts, three rises

The index records 26 changed price rows, but they belong to ten repricings, each one model on one date. OpenAI's change to gpt-5.6-sol on 22 August 2026 moved eight rates at once. Seven repricings were cuts and three were rises, all three in 2026 from Mistral and DeepSeek. The deepest cut is o3, from $40 to $8 per million output tokens. Four of the ten fall between 13 and 22 August 2026, the only month with more than one. The index first recorded 70 of its 130 models in 2026, so an older repricing is less likely to be on record. Three of the changes need context, and the note under each chart gives it.

20252026Aug 2024Oct 2026Four between 13 and 22 August20252026Aug 2024Oct 2026Four between 13 and 22 August
cutrise
One mark per repricing, in the month it was recorded. Hover or focus a mark for the model and the rates it moved.

price cut price riseEach panel is one model's output price, indexed so its first recorded price is 100, on its own time range. grok-4.5 and claude-sonnet-5-5 changed only other rates, so they have no panel.

o3−80% · $40 → $8

New price recorded from 10 June 2025; tracked since 16 April 2025.

0100200Apr 2025Oct 2026
gemini-3.6-flash−50% · $7.50 → $3.75

A promotion: the pricing page says $3.75 through 31 December 2026 and $7.50 from 1 January 2027.

0100200Jul 2026Oct 2026
gpt-4o−33% · $15 → $10

New price recorded from 6 August 2024; tracked since 13 May 2024.

0100200May 2024Oct 2026
gpt-5.6-sol−33% · $30 → $20

New price recorded from 22 August 2026; tracked since 9 July 2026.

0100200Jul 2026Oct 2026
claude-3-5-haiku−20% · $5 → $4 · retired

New price recorded from 3 December 2024; tracked since 4 November 2024.

0100200Nov 2024Feb 2026
mistral-small-4+100% · $0.30 → $0.60

New price recorded from 24 June 2026; tracked since 16 March 2026.

0100200Mar 2026Oct 2026
deepseek-v4-pro+128% · $0.87 → $1.98

The off-peak rate after DeepSeek split prices into peak and off-peak; peak is exactly double.

0100200Jun 2026Oct 2026
deepseek-v4-flash+136% · $0.28 → $0.66

The off-peak rate after DeepSeek split prices into peak and off-peak; peak is exactly double. Renamed deepseek-flash on 10 September 2026, at $0.60.

0100200Apr 2026Sep 2026

Retirements

22 models have left the list

22 of the 130 models have left their price lists, 19 of them in 2026. The median one was listed for 444 days; gpt-4 lasted 1,263. Six of them end on the same day, 28 August 2026, because that is when the collector noticed they were gone, not when they went. Leaving the list is not the same as being switched off: gpt-4 and gpt-4-turbo are gone from OpenAI's price list, but the dataset's notes give both a scheduled retirement date of 23 October 2026.

202320242025202628 August 2026: six found gone at oncegpt-4gpt-4-turboclaude-3-haikuclaude-3-opusclaude-3-sonnetclaude-3-5-sonnetgemini-1.5-flashcommand-r-pluso1-minigemini-1.5-proclaude-3-5-haikugemini-2.0-flashclaude-3-7-sonnetmagistral-mediummagistral-smallo3-deep-researcho4-mini-deep-researchdevstral-2devstral-small-2deepseek-v4-flashcommand-a-plusdeepseek-v4-flash-vision-exp202320242025202628 August 2026: six found gone at oncegpt-4gpt-4-turboclaude-3-haikuclaude-3-opusclaude-3-sonnetclaude-3-5-sonnetgemini-1.5-flashcommand-r-pluso1-minigemini-1.5-proclaude-3-5-haikugemini-2.0-flashclaude-3-7-sonnetmagistral-mediummagistral-smallo3-deep-researcho4-mini-deep-researchdevstral-2devstral-small-2deepseek-v4-flashcommand-a-plusdeepseek-v4-flash-vision-exp
AnthropicOpenAIGoogleevery other provideroutline: renamed, not withdrawn
From the first recorded price to the last day the model was seen on its provider's price list. Oldest first.

Cross-check

Two other trackers agree on 269 of 276 shared prices

Three public datasets are pulled here. The AI Model Price Stability dataset reads prices from web-archive captures of the Anthropic, Google and OpenAI pricing pages, with a link to every capture. LLM Token Price History keeps a current table for 35 providers, resellers included. Where either states a price the index also holds, for the same model and price type, they match 269 times out of 276: 60 of 60 archive prices on the day of the capture, and 209 of 216 current prices across 83 models. Of the 83 shared models, 13 carry a non-current status in LLM Token Price History (8 deprecated, 4 retired, 1 preview) while the index still lists them; the prices agree anyway.

Index vs AI Model Price Stability (archive captures, same day)60 matchIndex vs ModelPriceWatch (current prices)209 match · 7 differIndex vs AI Model Price Stability (archive captures, same day)60 matchIndex vs ModelPriceWatch (current prices)209 match · 7 differ
same price in bothdifferent price
One square per price that both datasets state for the same model and price type (60 of the archive's 147 comparable states share a date with the index). Archive prices are compared with the index's price on the day of the capture; ModelPriceWatch's with the index's current price. Hover a square for the model and both prices.

What only the archive caught

The archive also holds 19 price changes the index does not record, 11 of them on models the index tracks. Its authors call these candidates, not verified changes, so each one links to the two captures it rests on.

ModelPrice typeBeforeAfterCapturedIn the index
claude-3-5-haikucache read$0.03$0.025 ↓2024-10-23 → 2024-10-29model tracked, change not recorded
claude-3-5-haikucache write$0.30$0.3125 ↑2024-10-23 → 2024-10-29model tracked, change not recorded
claude-3-5-haikucache read$0.025$0.10 ↑2024-10-29 → 2024-11-04model tracked, change not recorded
claude-3-5-haikucache write$0.3125$1.25 ↑2024-10-29 → 2024-11-04model tracked, change not recorded
claude-3-5-haikuThe earlier capture shows the model marked "Coming soon" at $0.25 input: a price listed before launch.input$0.25$1 ↑2024-10-23 → 2024-11-04model tracked, change not recorded
claude-3-5-haikuoutput$1.25$5 ↑2024-10-23 → 2024-11-04model tracked, change not recorded
claude-3-5-haikucache read$0.10$0.08 ↓2024-11-04 → 2024-12-03model tracked, change not recorded
claude-3-5-haikucache write$1.25$1 ↓2024-11-04 → 2024-12-03model tracked, change not recorded
o3cache read$2.50$0.50 ↓2025-04-17 → 2025-06-11model tracked, change not recorded
gemini-2-5-flash-previewcache read$0.0375$0.075 ↑2025-05-11 → 2025-09-26model not tracked
gemini-2-5-flash-previewinput$0.15$0.30 ↑2025-04-18 → 2025-09-26model not tracked
gemini-2-5-flashcache read$0.075$0.03 ↓2025-06-17 → 2025-10-09model tracked, change not recorded
gemini-2-5-flash-previewcache read$0.075$0.0375 ↓2025-09-26 → 2025-10-09model not tracked
gemini-2-5-flash-previewcache read$0.0375$0.03 ↓2025-10-09 → 2025-10-16model not tracked
gemini-2-5-flash-litecache read$0.025$0.01 ↓2025-07-23 → 2025-10-24model tracked, change not recorded
gemini-2-5-flash-lite-previewcache read$0.025$0.01 ↓2025-06-17 → 2025-10-24model not tracked
gemini-3-1-flash-image-previewinput$0.25$0.50 ↑2026-02-26 → 2026-03-03model not tracked
gemini-robotics-er-previewinput$0.30$1 ↑2025-10-09 → 2026-04-30model not tracked
gemini-robotics-er-previewoutput$2.50$5 ↑2025-10-09 → 2026-04-30model not tracked

Each price links to the web-archive capture it was read from.

Limits

What this data cannot tell you

It is mostly recent.

70 of 130 models are first recorded in 2026, and dates are when a price was seen, not always when it changed. A model repriced before it was tracked shows no change here.

Not every price comes from the provider.

503 rows were read from a live pricing page, 36 from web archives and 2 from changelogs. The index has no first-party API price for Meta's models, so its 10 rows come from Together AI.

Some rows are inferred.

74 are marked that way: 24 are cache prices derived from a published multiplier, 10 are the Together AI rows, and 40 are unconfirmed for other reasons: approximate dates, uncorroborated prices and superseded records, which the notes give for all but 8 (those rows carry no note of their own). 435 are confirmed against the provider, 36 come from web archives, 2 from changelogs, and 4 are retained prices for delisted models.

List price is not what you pay.

Batch discounts, regional pricing and DeepSeek's peak hours all move the bill; the index records 12 kinds of price and this page uses the standard ones. If you run open weights yourself, the cost is memory rather than tokens: see the VRAM and KV-cache calculator.

Every price

All 108 current prices

Price per million tokens. Set the size of a call to see what it costs on each model. Which model is cheap depends on the shape of the call. From a long document (50,000 tokens in, 1,000 out) to a long answer (1,000 in, 8,000 out), 17 of 108 models move ten or more places in the order. llama-3.1-405b moves 35 places towards the cheapest, because it charges the same for output as for input. Its prices are Together AI's, flagged for review in the dataset.

Long documentLong answercheapestpriciestllama-3.1-8b ↑12ministral-3-14b ↑12llama-3.3-70b ↑23grok-build-0.1 ↑11grok-4.20-0309-non-reasoning ↑11grok-4.20-0309-reasoning ↑11grok-4.20-multi-agent-0309 ↑11grok-4.3 ↑11llama-3.1-405b ↑35gpt-5.4-mini ↓10qwen3.7-max ↑15gemini-2.5-pro ↓12gpt-5 ↓12gpt-5.1 ↓12nova-2-pro ↓12gpt-5.2 ↓17gpt-5.3-codex ↓17Long documentLong answercheapestpriciestllama-3.1-8b ↑12llama-3.3-70b ↑23grok-build-0.1 ↑11grok-4.20-non-… ↑11llama-3.1-405b ↑35qwen3.7-max ↑15gemini-2.5-pro ↓12gpt-5.2 ↓17
moves ten or more places towards the cheapestmoves ten or more places away
Each line is a model: its cost rank for a 50,000-in, 1,000-out call on the left, and for a 1,000-in, 8,000-out call on the right; cheapest at the top. Try other shapes in the table below.
Two-model crossover:vsgemini-3.5-flash is cheaper while a call returns fewer than 364 output tokens per 1,000 input tokens; above that, llama-3.1-405b is.
ModelProviderFirst recordedInput /MOutput /MThis call
o1-proOpenAIMar 2025$150$600$2.10
gpt-5.5-proOpenAIApr 2026$30$180$0.48
gpt-5.4-proOpenAIMar 2026$30$180$0.48
gpt-5.2-proOpenAIJan 2026$21$168$0.38
o3-proOpenAIJun 2025$20$80$0.28
claude-opus-4-20250514AnthropicMay 2025$15$75$0.23
claude-opus-4-1-20250805AnthropicAug 2025$15$75$0.23
o1OpenAIDec 2024$15$60$0.21
gpt-6-astraOpenAISep 2026$10$50$0.15
claude-mythos-5-1AnthropicSep 2026$10$50$0.15
claude-mythos-5AnthropicJun 2026$10$50$0.15
claude-fable-5-1AnthropicSep 2026$10$50$0.15
claude-fable-5AnthropicJun 2026$10$50$0.15
gpt-5.5OpenAIApr 2026$5$30$0.08
gpt-rosalind-researchOpenAIOct 2026$5$25$0.075
claude-opus-5AnthropicJul 2026$5$25$0.075
claude-opus-4-8AnthropicMay 2026$5$25$0.075
claude-opus-4-7AnthropicApr 2026$5$25$0.075
claude-opus-4-6AnthropicFeb 2026$5$25$0.075
claude-opus-4-5-20251101AnthropicNov 2025$5$25$0.075
gpt-5.6-solOpenAIJul 2026$4$20$0.06
claude-opus-5-5AnthropicSep 2026$4$20$0.06
gpt-5.4OpenAIMar 2026$2.50$15$0.04
claude-sonnet-4-6AnthropicFeb 2026$3$15$0.045
claude-sonnet-4-5-20250929AnthropicSep 2025$3$15$0.045
claude-sonnet-4-20250514AnthropicMay 2025$3$15$0.045
gpt-5.3-codexOpenAIJul 2026$1.75$14$0.032
gpt-5.2OpenAIJan 2026$1.75$14$0.032
nova-premierAmazonApr 2025$2.50$12.50$0.037
gpt-5.6-terraOpenAIJul 2026$2$12$0.032
gemini-3.1-pro-previewGoogleFeb 2026$2$12$0.032
gpt-6.1-solOpenAISep 2026$2$10$0.03
gpt-6-solOpenAISep 2026$2$10$0.03
gpt-5.1OpenAINov 2025$1.25$10$0.022
gpt-5OpenAIAug 2025$1.25$10$0.022
gpt-4oOpenAIMay 2024$2.50$10$0.035
gemini-2.5-proGoogleJun 2025$1.25$10$0.022
claude-sonnet-5-5AnthropicSep 2026$2$10$0.03
claude-sonnet-5AnthropicJul 2026$2$10$0.03
nova-2-proAmazonJun 2026$1.25$10$0.022
gemini-3.5-flashGoogleMay 2026$1.50$9$0.024
o3OpenAIApr 2025$2$8$0.028
gpt-4.1OpenAIApr 2025$2$8$0.028
jamba-largeAI21Jul 2025$2$8$0.028
mistral-medium-3.5MistralApr 2026$1.50$7.50$0.022
qwen3.7-maxAlibabaMay 2026$2.50$7.50$0.033
qwen-maxAlibabaJan 2025$1.60$6.40$0.022
grok-4.7xAISep 2026$2$6$0.026
grok-4.6xAIAug 2026$2$6$0.026
grok-4.5xAIJul 2026$2$6$0.026
qwen3.8-maxAlibabaAug 2026$2$6$0.026
qwen3-maxAlibabaJan 2026$1.20$6$0.018
claude-haiku-4-5-20251001AnthropicOct 2025$1$5$0.015
qwen3-coder-plusAlibabaJul 2025$1$5$0.015
gpt-5.4-miniOpenAIMar 2026$0.75$4.50$0.012
o4-miniOpenAIApr 2025$1.10$4.40$0.015
o3-miniOpenAIJan 2025$1.10$4.40$0.015
gemini-3.8-flashGoogleSep 2026$0.75$3.75$0.011
gemini-3.7-flashGoogleAug 2026$0.75$3.75$0.011
gemini-3.6-flashGoogleJul 2026$0.75$3.75$0.011
llama-3.1-405bMetaJun 2026$3.50$3.50$0.038
nova-proAmazonDec 2024$0.80$3.20$0.011
gemini-3-flash-previewGoogleDec 2025$0.50$3$0.008
qwen3.6-plusAlibabaApr 2026$0.50$3$0.008
grok-4.3xAIMay 2026$1.25$2.50$0.015
grok-4.20-multi-agent-0309xAIMar 2026$1.25$2.50$0.015
grok-4.20-0309-reasoningxAIMar 2026$1.25$2.50$0.015
grok-4.20-0309-non-reasoningxAIJun 2026$1.25$2.50$0.015
gemini-3.5-flash-liteGoogleJul 2026$0.30$2.50$0.0055
gemini-2.5-flashGoogleJun 2025$0.30$2.50$0.0055
nova-2-liteAmazonDec 2025$0.30$2.50$0.0055
qwen3.5-plusAlibabaFeb 2026$0.40$2.40$0.0064
mistral-large-4MistralOct 2026$0.68$2.09$0.0089
grok-build-0.1xAIJun 2026$1$2$0.012
deepseek-v4-proDeepSeekJun 2026$0.66$1.98$0.0086
qwen3.7-plusAlibabaJun 2026$0.40$1.60$0.0056
mistral-large-3MistralDec 2025$0.50$1.50$0.0065
gemini-3.1-flash-liteGoogleMay 2026$0.25$1.50$0.004
qwen3.6-flashAlibabaApr 2026$0.25$1.50$0.004
qwen3-coder-nextAlibabaAug 2026$0.30$1.50$0.0045
qwen3-coder-flashAlibabaAug 2026$0.30$1.50$0.0045
gpt-5.4-nanoOpenAIMar 2026$0.20$1.25$0.0032
gpt-5.6-lunaOpenAIJul 2026$0.20$1.20$0.0032
qwen-plusAlibabaDec 2025$0.40$1.20$0.0052
llama-3.3-70bMetaJun 2026$1.04$1.04$0.011
codestral-25.08MistralJul 2025$0.30$0.90$0.0039
llama-4-maverickMetaApr 2025$0.27$0.85$0.0036
gpt-4o-miniOpenAIJul 2024$0.15$0.60$0.0021
mistral-small-4MistralMar 2026$0.15$0.60$0.0021
deepseek-flashDeepSeekSep 2026$0.15$0.60$0.0021
command-rCohereAug 2024$0.15$0.60$0.0021
llama-4-scoutMetaApr 2025$0.18$0.59$0.0024
gpt-6-lunaOpenAISep 2026$0.10$0.50$0.0015
claude-haiku-5-5AnthropicOct 2026$0.10$0.50$0.0015
qwen3.8-flashAlibabaAug 2026$0.15$0.47$0.002
gemini-2.5-flash-liteGoogleJun 2025$0.10$0.40$0.0014
qwen3.5-flashAlibabaFeb 2026$0.10$0.40$0.0014
qwen-flashAlibabaJul 2025$0.05$0.40$0.0009
jamba-miniAI21Jan 2026$0.20$0.40$0.0024
nova-liteAmazonDec 2024$0.06$0.24$0.00084
ministral-3-14bMistralDec 2025$0.20$0.20$0.0022
qwen-turboAlibabaApr 2025$0.05$0.20$0.0007
llama-3.1-8bMetaJun 2026$0.18$0.18$0.002
ministral-3-8bMistralDec 2025$0.15$0.15$0.0016
command-r7bCohereDec 2024$0.0375$0.15$0.00052
nova-microAmazonDec 2024$0.035$0.14$0.00049
qwen3.7-flashAlibabaJul 2026$0.03$0.13$0.00043
ministral-3-3bMistralDec 2025$0.10$0.10$0.0011

The bar is the output price on the charts' log scale, $0.10 to $1,000; blue is first recorded in 2026.

All 26 changed price rows (10 repricings)
ModelPrice typeBeforeAfterRecorded from
gpt-4oinput$5$2.50 ↓2024-08-06
gpt-4ooutput$15$10 ↓2024-08-06
claude-3-5-haiku-20241022input$1$0.80 ↓2024-12-03
claude-3-5-haiku-20241022output$5$4 ↓2024-12-03
o3input$10$2 ↓2025-06-10
o3output$40$8 ↓2025-06-10
mistral-small-4input$0.10$0.15 ↑2026-06-24
mistral-small-4output$0.30$0.60 ↑2026-06-24
grok-4.5cache read$0.50$0.30 ↓2026-07-19
gemini-3.6-flashinput$1.50$0.75 ↓2026-08-13
gemini-3.6-flashoutput$7.50$3.75 ↓2026-08-13
deepseek-v4-flashcache read$0.0028$0.007 ↑2026-08-16
deepseek-v4-flashinput$0.14$0.22 ↑2026-08-16
deepseek-v4-flashoutput$0.28$0.66 ↑2026-08-16
deepseek-v4-procache read$0.0036$0.022 ↑2026-08-16
deepseek-v4-proinput$0.435$0.66 ↑2026-08-16
deepseek-v4-prooutput$0.87$1.98 ↑2026-08-16
gpt-5.6-solcache read$0.50$0.40 ↓2026-08-22
gpt-5.6-solcache write$6.25$5 ↓2026-08-22
gpt-5.6-solinput$5$4 ↓2026-08-22
gpt-5.6-soloutput$30$20 ↓2026-08-22
gpt-5.6-soltier2 cache read$1$0.80 ↓2026-08-22
gpt-5.6-soltier2 cache write$12.50$10 ↓2026-08-22
gpt-5.6-soltier2 input$10$8 ↓2026-08-22
gpt-5.6-soltier2 output$45$30 ↓2026-08-22
claude-sonnet-5-5cache read$0.20$0.10 ↓2026-10-08

Method

Method and source

Prices are providers' list prices as recorded in the dataset snapshots of 8 October 2026, and providers change them without notice. None of them is a retail price, and nothing on this page is an affiliate link.

Dataset
RoninForge/ai-price-index at revision 02fc932d333a, retrieved 2026-10-08; sha256 620ad4ed8137…AI Price Index by RoninForge (https://roninforge.org/data/ai-price-index/), CC BY 4.0.
Cross-check
CurtisPyke/ai-model-price-stability at revision cd23c3223863, retrieved 2026-09-24; sha256 97badbe5cf88…AI Model Price Stability by Curtis Pyke / Kingy AI (DOI 10.5281/zenodo.22288610), CC BY 4.0. Its own limitations, which its licence asks users to preserve: archive coverage is incomplete, detected transitions are candidates rather than verified vendor price changes, and vendor scope and tiers may be ambiguous.
Cross-check
modelpricewatch/llm-token-price-history at revision 2e95adfec8e7, retrieved 2026-10-04; sha256 e77f0a452e74…LLM Token Price History by ModelPriceWatch (https://modelpricewatch.com), CC BY 4.0.
Refresh
Every day each dataset is checked for a new revision. A new one is validated (columns, units, row counts) and shipped only if this page still builds and its checks pass; the figures above are always computed from the committed snapshot.
Current price
A row with no end date, as the dataset defines it. Medians and percentiles are type-7 quantiles, the NumPy and R default.
First recorded
The earliest date the index holds a model's input price. For older models it usually comes from an archived pricing page near launch; it is not a published release date.
Repricing
Consecutive rows for one model and one price type whose price differs, grouped by model and recorded-from date.