LLMPrice.io

LLMPrice indices · 2026-08

What the same AI workload costs today versus July 2024 = 100: below 100 means cheaper, above means more expensive.

Four price indices for AI compute, each pricing the same twelve models on a different workload shape. Base period 2024-07 = 100.

Prices last synced 2026-08-01. Latest issue; canonical points at the dated release.

This issue carries a correction to previously published figures. Read what changed and why.

Priced on the models a team already runs, AI compute has barely moved in nineteen months. The agent-shaped workload is down 42.1% since 2024-07, but only 14.4% of that came after January 2025.

The frontier did get cheaper over the same period, and by a lot. It got cheaper by replacement: providers ship new models at lower rates far more often than they cut the price of a model already in service. Only 10 of the 25 monthly links in this series carry a price change on a constituent present in both periods. A matched-model index measures the price of what you are already buying, so a saving that arrives as a new model name is invisible to it by construction, and invisible on your bill too until you migrate.

Which is the point of the split below: the output-heavy profile now reads 90.52, below its base period, while the cache-heavy agent profile reads 57.87. Whether a team captured any of the decline depended on the shape of its workload, not on which vendor it picked.

Retrieval Index
71.56
▼ 28.4% since 2024-07  ·  $2.38 per 1M tokens
Flagship premium 2.79×
Chat Index
87.83
▼ 12.2% since 2024-07  ·  $5.07 per 1M tokens
Flagship premium 2.8×
Content Index
90.52
▼ 9.5% since 2024-07  ·  $7.34 per 1M tokens
Flagship premium 2.8×
Agent Index
57.87
▼ 42.1% since 2024-07  ·  $1.39 per 1M tokens
Flagship premium 2.76×

Which of these is your workload

The four shapes are not categories of business, they are ratios of text in to text out. Pick the one that sounds like what you run and price it on today's rates.

Downloads and citation

Full series, JSON  ·  Full series, CSV

These are the living files: they lengthen by one reading each month. To cite a figure, use the dated issue, whose own downloads are frozen with it.

LLMPrice.io, “LLMPrice indices, 2026-08”. Base 2024-07 = 100. Retrieved 2026-08-01. https://llmprice.io/index/2026-08

Showing a reading on your own page rather than quoting one: the badge carries any of the four values and updates when the monthly reading does.

Corrections to this issue

Restatement notice

The previous issue, 2026-07, corrected the basket-selection rule and restated the full back series; every reading before it differs from its first publication, and the superseded figures are recorded on that issue. This issue's readings chain from the corrected series.

How to read these indices

Each index prices a fixed basket, the LLMPrice 12: two models per provider, the newest member of the provider's own flagship line and of its value line. An index moves only when one of those twelve is repriced. It does not track every model on the market, and it does not track the wider repricing record on the site's ticker, which includes models outside the basket.

A new model entering the basket cannot move the index. Substitutions are chain-linked, so only a price change on a model present in consecutive readings registers. This is deliberate: a cheaper replacement reaches your bill only if you migrate to it, so it must not appear here as though prices fell on their own.

Providers rarely reprice a model in service. Only 10 of the 25 monthly links in this series carry any change at all, and the long flat stretches in the chart are the finding, not a gap in it: the price of what a team already runs barely moves, while the market's headline prices fall by replacement.

That gap is what these indices exist to measure. A busy repricing ticker beside a flat index is the market working exactly as observed: individual models change price occasionally, new models arrive cheaper constantly, and neither reaches an existing bill without a decision. The basket, the line assignments and their sources are on the methodology page.

Deflation by cacheable share of input

A standing section, reported every issue. The four profiles differ in how much of their input can be served from cache. Ordering the profiles by that share shows whether price movement has tracked it.

ProfileCacheable inputIndexSince 2024-07
Agent Index85%57.87-42.1%
Retrieval Index70%71.56-28.4%
Chat Index30%87.83-12.2%
Content Index20%90.52-9.5%

Movement is ordered by cacheable share of input this issue: the more of a workload's input can be cached, the further its index has fallen. This is a correlation across four profiles, not a demonstration of intent. It is reported each issue whether or not the ordering holds.

History

The series began 2024-07 and lengthens by one reading each month. 26 monthly readings so far, chat profile shown. It is short, and will stay short for a while.

The LLMPrice 12, as priced this issue

ConstituentProviderLineCache rateChat cost
GPT-5.6
gpt-5.6
OpenAIflagship$0.500$16.82
GPT-5.4 Mini
gpt-5.4-mini
OpenAIvalue$0.075$2.52
Claude Opus 5
claude-opus-5
Anthropicflagship$0.500$14.32
Claude Sonnet 5
claude-sonnet-5
Anthropicvalue$0.200$5.73
Gemini 3.1 Pro Preview
gemini/gemini-3.1-pro-preview
Googleflagship$0.200$6.73
Gemini 3.6 Flash
gemini/gemini-3.6-flash
Googlevalue$0.150$4.30
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeekflagship$0.004$0.588
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeekvalue$0.003$0.189
Grok 4.5
xai/grok-4.5
xAIflagship$0.500$3.77
Grok 3 Mini
xai/grok-3-mini
xAIvalue$0.075$0.366
Mistral Large (2025-12)
mistral/mistral-large-2512
Mistralflagshipnone published$1.00
Mistral Medium (2026-04)
mistral/mistral-medium-2604
Mistralvaluenone published$4.50

Substitutions

Every basket change since the series began, with the linking factor applied so the change does not appear as a price movement.

DateInOutLink factor
2024-08-01GPT-4o Mini, Mistral Large (2024-07)GPT-3.5 Turbo, Mistral Large (2024-02)1.0999
2024-11-01Claude Sonnet 3.5 (2024-10)Claude Sonnet 3.5 (2024-06)0.9460
2025-02-01DeepSeek Reasoner, DeepSeek Chat, Grok 2, Mistral Large (2024-11)Mistral Large (2024-07)1.0000
2025-03-01Gemini 2.0 FlashGemini 1.5 Flash1.0053
2025-05-01GPT-4.1, GPT-4.1 MiniGPT-4o, GPT-4o Mini1.0000
2025-06-01Claude Opus 4 (2025-05), Claude Sonnet 4 (2025-05), Grok 3, Mistral Medium (2025-05)Claude Opus 3 (2024-02), Claude Sonnet 3.5 (2024-10), Grok 2, Mistral Medium (2023-12)1.0000
2025-07-01Gemini 2.5 Pro, Gemini 2.5 Flash, DeepSeek R1, Grok 3 MiniGemini 1.5 Pro, Gemini 2.0 Flash, DeepSeek Reasoner1.0000
2025-08-01Grok 4Grok 30.9814
2025-09-01GPT-5, GPT-5 Mini, Claude Opus 4.1GPT-4.1, GPT-4.1 Mini, Claude Opus 4 (2025-05)1.0000
2025-10-01Claude Sonnet 4.5Claude Sonnet 4 (2025-05)1.0000
2025-12-01GPT-5.1, Gemini 3 Pro PreviewGPT-5, Gemini 2.5 Pro0.9993
2026-01-01GPT-5.2, Claude Opus 4.5GPT-5.1, Claude Opus 4.11.0000
2026-03-01Claude Opus 4.6, Claude Sonnet 4.6, Gemini 3.1 Pro PreviewClaude Opus 4.5, Claude Sonnet 4.5, Gemini 3 Pro Preview1.0000
2026-04-01GPT-5.4, GPT-5.4 Mini, Mistral Large (2025-12)GPT-5.2, GPT-5 Mini, Mistral Large (2024-11)1.0000
2026-05-01GPT-5.5, Claude Opus 4.7GPT-5.4, Claude Opus 4.61.0000
2026-06-01Claude Opus 4.8, Gemini 3.5 Flash, Grok 4.3Claude Opus 4.7, Gemini 2.5 Flash, Grok 41.0000
2026-07-01Claude Sonnet 5, DeepSeek V4 Pro, DeepSeek V4 Flash, Mistral Medium (2026-04)Claude Sonnet 4.6, DeepSeek R1, DeepSeek Chat, Mistral Medium (2025-05)1.0000
2026-08-01GPT-5.6, Claude Opus 5, Gemini 3.6 Flash, Grok 4.5GPT-5.5, Claude Opus 4.8, Gemini 3.5 Flash, Grok 4.30.8919

Full methodology  ·  All issues