Grok 4.6 API pricing
Published on-demand rates for xAI's xai/grok-4.6, captured 2026-08-27 from public pricing data and committed to our archive. Context window 500k tokens; first observed in our archive 2026-08-17.
| Meter | USD per 1M tokens |
|---|---|
| Input | $2.00 |
| Output | $6.00 |
| Cached input | $0.500 |
Standard on-demand rates. Excludes batch endpoints and negotiated pricing; confirm on xAI's pricing page before committing.
What it costs on a real workload
Per 1M tokens processed, at the rates above. How the four profiles are defined.
| Profile | In / out mix | Cached share | Cost |
|---|---|---|---|
| Retrieval | 800k / 200k | 70% | $1.96 |
| Chat | 500k / 500k | 30% | $3.77 |
| Content | 200k / 800k | 20% | $5.14 |
| Agent | 900k / 100k | 85% | $1.25 |
Observed rate history
No repricing observed since this model entered our archive on 2026-08-17. When a rate moves, the change appears here and on the sitewide ticker.
What the record shows
Grok 4.6 has not repriced once since it entered our archive on 2026-08-17. That is the ordinary case rather than the exception: providers ship new models at lower rates far more often than they cut the price of one already in service, which is why a cheaper option usually arrives as a new name rather than as a smaller number on this page.
Against the 43 models we price, it is the 29th cheapest for retrieval-shaped work, which reads a great deal and writes little, and the 30th cheapest for content-shaped work, which does the reverse. The two are close, so it holds its position whichever way your workload leans.
It charges 3.0 times more for output than for input, against a median of 5.0 times across the set. That is narrower than most, so it punishes long answers less than the typical model does.
Cached input reads at $0.500 against a standard input rate of $2.00, a 75% discount on the part of your prompt that repeats.
It accepts image input, but xAI publishes no rule for how images become tokens, so we will not price a scanned page on it. That is a gap in the provider's documentation rather than a limit of the model, and it is the reason this model is excluded from image estimates in our estimator instead of being given a plausible figure.
Its 500k token context window holds roughly 375,000 words of English text in a single request. A workload that does not fit cannot run here at any price, which is a capability limit rather than a cost one.
Priced near this one
Closest to Grok 4.6 on a retrieval workload, cheaper and more expensive alike. Cost only; we publish no quality score.
| Model | Provider | Input | Output | vs Grok 4.6 |
|---|---|---|---|---|
| Sonar Reasoning | Perplexity | $1.00 | $5.00 | 8% less |
| Gemini 3.5 Flash | $1.50 | $9.00 | 14% more | |
| Gemini 2.5 Pro | $1.25 | $10.00 | 21% more | |
| GPT 5 Chat | OpenAI | $1.25 | $10.00 | 21% more |
Retrieval profile: 800k in, 200k out, 70% cached where offered.
No affiliate or referral arrangements: the console link carries no parameters and nothing on this page is paid placement. Grok 4.6 is a trademark of its owner; LLMPrice.io is independent and unaffiliated.