LLM API pricing, ranked by cost
Every current first-party text model, priced on the workload you actually run. Cheapest depends on the shape of the work: a model that wins on a retrieval workload can lose badly on a content one. Rates captured 2026-09-21 from public pricing data and committed to our archive.
| # | ||||||
|---|---|---|---|---|---|---|
| 1 | Ministral 3 3B 2512 | Mistral | $0.100 | $0.100 | $0.010 | $0.050 |
| 2 | Ministral 3 8B 2512 | Mistral | $0.150 | $0.150 | $0.015 | $0.074 |
| 3 | Gemini 2.0 Flash Lite 001 | $0.075 | $0.300 | $0.019 | $0.088 | |
| 4 | GPT 5 Nano | OpenAI | $0.050 | $0.400 | $0.005 | $0.095 |
| 5 | Ministral 3 14B 2512 | Mistral | $0.200 | $0.200 | $0.020 | $0.099 |
| 6 | Gemini 2.5 Flash Lite | $0.100 | $0.400 | $0.010 | $0.110 | |
| 7 | Devstral Small 2512 | Mistral | $0.100 | $0.300 | — | $0.140 |
| 8 | Mistral vibe cli Fast | Mistral | $0.150 | $0.600 | $0.015 | $0.164 |
| 9 | Codestral 2508 | Mistral | $0.300 | $0.900 | $0.030 | $0.269 |
| 10 | GPT 5.6 Luna | OpenAI | $0.200 | $1.20 | $0.020 | $0.299 |
| 11 | GPT 5.4 Nano | OpenAI | $0.200 | $1.25 | $0.020 | $0.309 |
| 12 | DeepSeek Flash | DeepSeek | $0.300 | $1.20 | $0.006 | $0.315 |
| 13 | Gemini 3.1 Flash Lite | $0.250 | $1.50 | $0.025 | $0.374 | |
| 14 | Mistral Large 2512 | Mistral | $0.500 | $1.50 | $0.050 | $0.448 |
| 15 | GPT 5 Mini | OpenAI | $0.250 | $2.00 | $0.025 | $0.474 |
| 16 | Gemini 3.5 Flash Lite | $0.300 | $2.50 | $0.030 | $0.589 | |
| 17 | Devstral 2512 | Mistral | $0.400 | $2.00 | — | $0.720 |
| 18 | Grok Code Fast | xAI | $1.00 | $2.00 | $0.200 | $0.752 |
| 19 | Grok 4.3 | xAI | $1.25 | $2.50 | $0.200 | $0.912 |
| 20 | Gemini 3.8 Flash | $0.750 | $3.75 | $0.075 | $0.972 | |
| 21 | Sonar | Perplexity | $1.00 | $1.00 | — | $1.00 |
| 22 | GPT 5.4 Mini | OpenAI | $0.750 | $4.50 | $0.075 | $1.12 |
| 23 | DeepSeek V4 Pro | DeepSeek | $1.32 | $3.96 | $0.044 | $1.13 |
| 24 | Claude Haiku 4.5 | Anthropic | $1.00 | $5.00 | $0.100 | $1.30 |
| 25 | Sonar Reasoning | Perplexity | $1.00 | $5.00 | — | $1.80 |
| 26 | Mistral Medium 3.5 | Mistral | $1.50 | $7.50 | $0.150 | $1.94 |
| 27 | Grok 4.6 | xAI | $2.00 | $6.00 | $0.500 | $1.96 |
| 28 | Gemini 2.5 Pro | $1.25 | $10.00 | $0.125 | $2.37 | |
| 29 | GPT 5 Chat | OpenAI | $1.25 | $10.00 | $0.125 | $2.37 |
| 30 | Claude Sonnet 5 | Anthropic | $2.00 | $10.00 | $0.200 | $2.59 |
| 31 | Magistral Medium 1.2 2509 | Mistral | $2.00 | $5.00 | — | $2.60 |
| 32 | GPT 5.6 Terra | OpenAI | $2.00 | $12.00 | $0.200 | $2.99 |
| 33 | Gemini omni 1.1 Flash | $1.50 | $9.00 | — | $3.00 | |
| 34 | Sonar Reasoning Pro | Perplexity | $2.00 | $8.00 | — | $3.20 |
| 35 | GPT 5.3 | OpenAI | $1.75 | $14.00 | $0.175 | $3.32 |
| 36 | GPT 5.6 | OpenAI | $4.00 | $20.00 | $0.400 | $5.18 |
| 37 | Sonar Pro | Perplexity | $3.00 | $15.00 | — | $5.40 |
| 38 | Claude Opus 5 | Anthropic | $5.00 | $25.00 | $0.500 | $6.48 |
| 39 | Claude Fable 5.1 | Anthropic | $10.00 | $50.00 | $0.250 | $12.54 |
| 40 | GPT 6 astra | OpenAI | $10.00 | $50.00 | $1.00 | $12.96 |
40 models priced on the retrieval profile. Models without a published cache-read rate are billed at the standard input rate for the whole input, and their cache column reads em dash.
If the model you use is not on this list
This table shows only models a provider will sell you today. If the one you are running is missing, it has been retired or replaced, and a migration is coming whether or not you have planned for one. We keep its final rates and its full recorded history, so you can compare what you are paying now against what is actually on sale, and the ticker at the top of every page carries its last observed change.
Using one of these figures
Rates on this page were captured 2026-09-21 from each provider's published pricing and committed to our archive, so this reading does not change when a provider edits a page. Quote the date with the number.
The same rates as JSON · how the archive works · questions about these numbers