LLM API pricing, ranked by cost
Every current first-party text model, priced on the workload you actually run. Cheapest depends on the shape of the work: a model that wins on a retrieval workload can lose badly on a content one. Rates captured 2026-08-10 from public pricing data and committed to our archive.
| # | ||||||
|---|---|---|---|---|---|---|
| 1 | Mistral Small 3.2 2506 | Mistral | $0.060 | $0.180 | — | $0.084 |
| 2 | Gemini 2.0 Flash Lite 001 | $0.075 | $0.300 | $0.019 | $0.088 | |
| 3 | DeepSeek V4 Flash | DeepSeek | $0.140 | $0.280 | $0.003 | $0.091 |
| 4 | GPT 5 Nano | OpenAI | $0.050 | $0.400 | $0.005 | $0.095 |
| 5 | Ministral 3 3B 2512 | Mistral | $0.100 | $0.100 | — | $0.100 |
| 6 | Gemini 2.5 Flash Lite | $0.100 | $0.400 | $0.010 | $0.110 | |
| 7 | Devstral Small 2512 | Mistral | $0.100 | $0.300 | — | $0.140 |
| 8 | Ministral 3 8B 2512 | Mistral | $0.150 | $0.150 | — | $0.150 |
| 9 | Grok 4.1 Fast | xAI | $0.200 | $0.500 | $0.050 | $0.176 |
| 10 | Ministral 3 14B 2512 | Mistral | $0.200 | $0.200 | — | $0.200 |
| 11 | Grok 3 Mini | xAI | $0.300 | $0.500 | $0.075 | $0.214 |
| 12 | DeepSeek V4 Pro | DeepSeek | $0.435 | $0.870 | $0.004 | $0.280 |
| 13 | GPT 5.6 Luna | OpenAI | $0.200 | $1.20 | $0.020 | $0.299 |
| 14 | GPT 5.4 Nano | OpenAI | $0.200 | $1.25 | $0.020 | $0.309 |
| 15 | Grok Code Fast | xAI | $0.200 | $1.50 | $0.020 | $0.359 |
| 16 | Gemini 3.1 Flash Lite | $0.250 | $1.50 | $0.025 | $0.374 | |
| 17 | Codestral 2508 | Mistral | $0.300 | $0.900 | — | $0.420 |
| 18 | GPT 5 Mini | OpenAI | $0.250 | $2.00 | $0.025 | $0.474 |
| 19 | Gemini 3.5 Flash Lite | $0.300 | $2.50 | $0.030 | $0.589 | |
| 20 | Mistral Large 2512 | Mistral | $0.500 | $1.50 | — | $0.700 |
| 21 | Mistral Medium 2508 | Mistral | $0.400 | $2.00 | — | $0.720 |
| 22 | Grok 4.3 | xAI | $1.25 | $2.50 | $0.200 | $0.912 |
| 23 | Sonar | Perplexity | $1.00 | $1.00 | — | $1.00 |
| 24 | Grok 3 Mini Fast | xAI | $0.600 | $4.00 | $0.150 | $1.03 |
| 25 | GPT 5.4 Mini | OpenAI | $0.750 | $4.50 | $0.075 | $1.12 |
| 26 | Claude Haiku 4.5 | Anthropic | $1.00 | $5.00 | $0.100 | $1.30 |
| 27 | Sonar Reasoning | Perplexity | $1.00 | $5.00 | — | $1.80 |
| 28 | Gemini 3.6 Flash | $1.50 | $7.50 | $0.150 | $1.94 | |
| 29 | Grok 4.5 | xAI | $2.00 | $6.00 | $0.500 | $1.96 |
| 30 | Gemini 3.5 Flash | $1.50 | $9.00 | $0.150 | $2.24 | |
| 31 | Gemini 2.5 Pro | $1.25 | $10.00 | $0.125 | $2.37 | |
| 32 | GPT 5 Chat | OpenAI | $1.25 | $10.00 | $0.125 | $2.37 |
| 33 | Claude Sonnet 5 | Anthropic | $2.00 | $10.00 | $0.200 | $2.59 |
| 34 | Magistral Medium 1.2 2509 | Mistral | $2.00 | $5.00 | — | $2.60 |
| 35 | GPT 5.6 Terra | OpenAI | $2.00 | $12.00 | $0.200 | $2.99 |
| 36 | Sonar Reasoning Pro | Perplexity | $2.00 | $8.00 | — | $3.20 |
| 37 | GPT 5.3 | OpenAI | $1.75 | $14.00 | $0.175 | $3.32 |
| 38 | Sonar Pro | Perplexity | $3.00 | $15.00 | — | $5.40 |
| 39 | Grok 4 | xAI | $3.00 | $15.00 | — | $5.40 |
| 40 | Claude Opus 5 | Anthropic | $5.00 | $25.00 | $0.500 | $6.48 |
| 41 | GPT 5.6 | OpenAI | $5.00 | $30.00 | $0.500 | $7.48 |
| 42 | Claude Fable 5 | Anthropic | $10.00 | $50.00 | $1.00 | $12.96 |
42 models priced on the retrieval profile. Models without a published cache-read rate are billed at the standard input rate for the whole input, and their cache column reads em dash.
If the model you use is not on this list
This table shows only models a provider will sell you today. If the one you are running is missing, it has been retired or replaced, and a migration is coming whether or not you have planned for one. We keep its final rates and its full recorded history, so you can compare what you are paying now against what is actually on sale, and the ticker at the top of every page carries its last observed change.
Using one of these figures
Rates on this page were captured 2026-08-10 from each provider's published pricing and committed to our archive, so this reading does not change when a provider edits a page. Quote the date with the number.
The same rates as JSON · how the archive works · questions about these numbers