LLMPrice indices · 2026-08
What the same AI workload costs today versus July 2024 = 100: below 100 means cheaper, above means more expensive.
Four price indices for AI compute, each pricing the same twelve models on a different workload shape. Base period 2024-07 = 100.
Prices last synced 2026-08-01. Latest issue; canonical points at the dated release.
This issue carries a correction to previously published figures. Read what changed and why.
Priced on the models a team already runs, AI compute has barely moved in nineteen months. The agent-shaped workload is down 42.1% since 2024-07, but only 14.4% of that came after January 2025.
The frontier did get cheaper over the same period, and by a lot. It got cheaper by replacement: providers ship new models at lower rates far more often than they cut the price of a model already in service. Only 10 of the 25 monthly links in this series carry a price change on a constituent present in both periods. A matched-model index measures the price of what you are already buying, so a saving that arrives as a new model name is invisible to it by construction, and invisible on your bill too until you migrate.
Which is the point of the split below: the output-heavy profile now reads 90.52, below its base period, while the cache-heavy agent profile reads 57.87. Whether a team captured any of the decline depended on the shape of its workload, not on which vendor it picked.
Which of these is your workload
The four shapes are not categories of business, they are ratios of text in to text out. Pick the one that sounds like what you run and price it on today's rates.
- Answering questions from your documents
Long context read repeatedly, short answers out. The Retrieval shape, and the one that benefits most from caching.
- A customer support chatbot
Roughly as much text out as in, back and forth. The Chat shape.
- Writing content, reports or summaries
Short brief in, a great deal out. The Content shape, and the most exposed to output rates.
- An agent working through a task
Many steps over a stable prefix, little written back each time. The Agent shape.
Downloads and citation
Full series, JSON · Full series, CSV
These are the living files: they lengthen by one reading each month. To cite a figure, use the dated issue, whose own downloads are frozen with it.
Showing a reading on your own page rather than quoting one: the badge carries any of the four values and updates when the monthly reading does.
Corrections to this issue
The previous issue, 2026-07, corrected the basket-selection rule and restated the full back series; every reading before it differs from its first publication, and the superseded figures are recorded on that issue. This issue's readings chain from the corrected series.
How to read these indices
Each index prices a fixed basket, the LLMPrice 12: two models per provider, the newest member of the provider's own flagship line and of its value line. An index moves only when one of those twelve is repriced. It does not track every model on the market, and it does not track the wider repricing record on the site's ticker, which includes models outside the basket.
A new model entering the basket cannot move the index. Substitutions are chain-linked, so only a price change on a model present in consecutive readings registers. This is deliberate: a cheaper replacement reaches your bill only if you migrate to it, so it must not appear here as though prices fell on their own.
Providers rarely reprice a model in service. Only 10 of the 25 monthly links in this series carry any change at all, and the long flat stretches in the chart are the finding, not a gap in it: the price of what a team already runs barely moves, while the market's headline prices fall by replacement.
That gap is what these indices exist to measure. A busy repricing ticker beside a flat index is the market working exactly as observed: individual models change price occasionally, new models arrive cheaper constantly, and neither reaches an existing bill without a decision. The basket, the line assignments and their sources are on the methodology page.
Deflation by cacheable share of input
A standing section, reported every issue. The four profiles differ in how much of their input can be served from cache. Ordering the profiles by that share shows whether price movement has tracked it.
| Profile | Cacheable input | Index | Since 2024-07 |
|---|---|---|---|
| Agent Index | 85% | 57.87 | -42.1% |
| Retrieval Index | 70% | 71.56 | -28.4% |
| Chat Index | 30% | 87.83 | -12.2% |
| Content Index | 20% | 90.52 | -9.5% |
Movement is ordered by cacheable share of input this issue: the more of a workload's input can be cached, the further its index has fallen. This is a correlation across four profiles, not a demonstration of intent. It is reported each issue whether or not the ordering holds.
History
The series began 2024-07 and lengthens by one reading each month. 26 monthly readings so far, chat profile shown. It is short, and will stay short for a while.
The LLMPrice 12, as priced this issue
| Constituent | Provider | Line | Cache rate | Chat cost |
|---|---|---|---|---|
| GPT-5.6 gpt-5.6 | OpenAI | flagship | $0.500 | $16.82 |
| GPT-5.4 Mini gpt-5.4-mini | OpenAI | value | $0.075 | $2.52 |
| Claude Opus 5 claude-opus-5 | Anthropic | flagship | $0.500 | $14.32 |
| Claude Sonnet 5 claude-sonnet-5 | Anthropic | value | $0.200 | $5.73 |
| Gemini 3.1 Pro Preview gemini/gemini-3.1-pro-preview | flagship | $0.200 | $6.73 | |
| Gemini 3.6 Flash gemini/gemini-3.6-flash | value | $0.150 | $4.30 | |
| DeepSeek V4 Pro deepseek-v4-pro | DeepSeek | flagship | $0.004 | $0.588 |
| DeepSeek V4 Flash deepseek-v4-flash | DeepSeek | value | $0.003 | $0.189 |
| Grok 4.5 xai/grok-4.5 | xAI | flagship | $0.500 | $3.77 |
| Grok 3 Mini xai/grok-3-mini | xAI | value | $0.075 | $0.366 |
| Mistral Large (2025-12) mistral/mistral-large-2512 | Mistral | flagship | none published | $1.00 |
| Mistral Medium (2026-04) mistral/mistral-medium-2604 | Mistral | value | none published | $4.50 |
Substitutions
Every basket change since the series began, with the linking factor applied so the change does not appear as a price movement.
| Date | In | Out | Link factor |
|---|---|---|---|
| 2024-08-01 | GPT-4o Mini, Mistral Large (2024-07) | GPT-3.5 Turbo, Mistral Large (2024-02) | 1.0999 |
| 2024-11-01 | Claude Sonnet 3.5 (2024-10) | Claude Sonnet 3.5 (2024-06) | 0.9460 |
| 2025-02-01 | DeepSeek Reasoner, DeepSeek Chat, Grok 2, Mistral Large (2024-11) | Mistral Large (2024-07) | 1.0000 |
| 2025-03-01 | Gemini 2.0 Flash | Gemini 1.5 Flash | 1.0053 |
| 2025-05-01 | GPT-4.1, GPT-4.1 Mini | GPT-4o, GPT-4o Mini | 1.0000 |
| 2025-06-01 | Claude Opus 4 (2025-05), Claude Sonnet 4 (2025-05), Grok 3, Mistral Medium (2025-05) | Claude Opus 3 (2024-02), Claude Sonnet 3.5 (2024-10), Grok 2, Mistral Medium (2023-12) | 1.0000 |
| 2025-07-01 | Gemini 2.5 Pro, Gemini 2.5 Flash, DeepSeek R1, Grok 3 Mini | Gemini 1.5 Pro, Gemini 2.0 Flash, DeepSeek Reasoner | 1.0000 |
| 2025-08-01 | Grok 4 | Grok 3 | 0.9814 |
| 2025-09-01 | GPT-5, GPT-5 Mini, Claude Opus 4.1 | GPT-4.1, GPT-4.1 Mini, Claude Opus 4 (2025-05) | 1.0000 |
| 2025-10-01 | Claude Sonnet 4.5 | Claude Sonnet 4 (2025-05) | 1.0000 |
| 2025-12-01 | GPT-5.1, Gemini 3 Pro Preview | GPT-5, Gemini 2.5 Pro | 0.9993 |
| 2026-01-01 | GPT-5.2, Claude Opus 4.5 | GPT-5.1, Claude Opus 4.1 | 1.0000 |
| 2026-03-01 | Claude Opus 4.6, Claude Sonnet 4.6, Gemini 3.1 Pro Preview | Claude Opus 4.5, Claude Sonnet 4.5, Gemini 3 Pro Preview | 1.0000 |
| 2026-04-01 | GPT-5.4, GPT-5.4 Mini, Mistral Large (2025-12) | GPT-5.2, GPT-5 Mini, Mistral Large (2024-11) | 1.0000 |
| 2026-05-01 | GPT-5.5, Claude Opus 4.7 | GPT-5.4, Claude Opus 4.6 | 1.0000 |
| 2026-06-01 | Claude Opus 4.8, Gemini 3.5 Flash, Grok 4.3 | Claude Opus 4.7, Gemini 2.5 Flash, Grok 4 | 1.0000 |
| 2026-07-01 | Claude Sonnet 5, DeepSeek V4 Pro, DeepSeek V4 Flash, Mistral Medium (2026-04) | Claude Sonnet 4.6, DeepSeek R1, DeepSeek Chat, Mistral Medium (2025-05) | 1.0000 |
| 2026-08-01 | GPT-5.6, Claude Opus 5, Gemini 3.6 Flash, Grok 4.5 | GPT-5.5, Claude Opus 4.8, Gemini 3.5 Flash, Grok 4.3 | 0.8919 |