LLMPrice.io

Mistral Medium 3 API pricing

Published on-demand rates for Mistral's mistral/mistral-medium-3, captured 2026-08-29 from public pricing data and committed to our archive. Context window 262k tokens; first observed in our archive 2026-08-28.

MeterUSD per 1M tokens
Input$1.50
Output$7.50
Cached inputno published cache-read rate; the whole input bills at the standard rate

Standard on-demand rates. Excludes batch endpoints and negotiated pricing; confirm on Mistral's pricing page before committing.

What it costs on a real workload

Per 1M tokens processed, at the rates above. How the four profiles are defined.

ProfileIn / out mixCached shareCost
Retrieval800k / 200k70%$2.70
Chat500k / 500k30%$4.50
Content200k / 800k20%$6.30
Agent900k / 100k85%$2.10

Compare every model on these profiles →

Observed rate history

No repricing observed since this model entered our archive on 2026-08-28. When a rate moves, the change appears here and on the sitewide ticker.

What the record shows

Mistral Medium 3 has not repriced once since it entered our archive on 2026-08-28. That is the ordinary case rather than the exception: providers ship new models at lower rates far more often than they cut the price of one already in service, which is why a cheaper option usually arrives as a new name rather than as a smaller number on this page.

Against the 44 models we price, it is the 35th cheapest for retrieval-shaped work, which reads a great deal and writes little, and the 31st cheapest for content-shaped work, which does the reverse. The two are far apart, and that is the useful part: Mistral Medium 3 is relatively better value the more your workload writes and the less it reads. A single position in a table sorted one way would have told you the opposite of the truth for the other kind of work.

It charges 5.0 times more for output than for input, against a median of 5.0 times across the set. That is close to typical for the models we track.

Mistral publishes no cache-read rate for it, so a prompt prefix you send on every call costs the same as text the model has never seen. On a cache-heavy workload that is a structural disadvantage against models that publish one, and it is why the agent profile above prices its whole input at the standard rate.

It accepts image input, but Mistral publishes no rule for how images become tokens, so we will not price a scanned page on it. That is a gap in the provider's documentation rather than a limit of the model, and it is the reason this model is excluded from image estimates in our estimator instead of being given a plausible figure.

Its 262k token context window holds roughly 197,000 words of English text in a single request. A workload that does not fit cannot run here at any price, which is a capability limit rather than a cost one.

Priced near this one

Closest to Mistral Medium 3 on a retrieval workload, cheaper and more expensive alike. Cost only; we publish no quality score.

ModelProviderInputOutputvs Mistral Medium 3
Magistral Medium 1.2Mistral$2.00$5.004% less
Claude Sonnet 5Anthropic$2.00$10.004% less
GPT 5.6 TerraOpenAI$2.00$12.0011% more
Gemini 2.5 ProGoogle$1.25$10.0012% less

Retrieval profile: 800k in, 200k out, 70% cached where offered.

Price your own prompt on it Estimate a whole project Plain rate card Mistral's console →

No affiliate or referral arrangements: the console link carries no parameters and nothing on this page is paid placement. Mistral Medium 3 is a trademark of its owner; LLMPrice.io is independent and unaffiliated.