LLM / providers / grok

xAI Grok API Pricing

xAI's Grok models offer a flagship tier with a large context window and a low output-to-input price ratio, plus a flat cached-input rate, which can suit output-heavy workloads.

Prices verified July 2026 · changes logged in the changelog
Heads up: xAI publishes a Batch API. Grok 4.3 carries a verified −20% batch discount; the newer Grok 4.5 and Grok Build tiers don't publish a batch multiplier yet, so the estimate doesn't assume batch on any tier — your real async cost may be lower where a discount applies. The cached-input rate is flat across tiers rather than scaling with the standard rate.
Model$ input /1M$ output /1M$ cached /1MBatch≈ $/mo *
Grok 4.5FRONTIER $2$6$0.30$342
Grok 4.3FRONTIER $1.25$2.50$0.20−20%$178
Grok Build 0.1MID $1$2$0.20$148

* Example workload — chatbot, 100k requests/mo, 2,000 input / 300 output tokens per request, 70% of input cached. Computed by the same engine as the calculator. Batch: the −20% is xAI's verified Batch API discount; the ≈ $/mo column is computed without it.

Prompt caching

Cache pricing differs per model: Grok 4.3 at $0.20/1M (16% of input); Grok Build 0.1 at $0.20/1M (20% of input); Grok 4.5 at $0.30/1M (15% of input). The calculator models this with your cache share.

Batch / async

The Batch API runs asynchronous jobs at a verified −20% on both input and output across Grok 4.3 — flip the Batch toggle in the calculator to model it.

Context window

Grok 4.3 runs a verified 1M-token context window; Grok 4.5 is 500k tokens; Grok Build 0.1 is 256k tokens. Grok 4.5, Grok 4.3 and Grok Build 0.1 bill higher rates on long-context prompts — this table and the calculator use standard rates.

When Grok is worth it

Use caseVerdict
Output-heavy workloadsGrok's output rate is relatively low versus input
Large context within a single requestThe flagship tier carries a large window
You need a confirmed batch discount up frontGrok 4.3 has a verified −20%; the newer tiers don't publish one yet
Is Grok the right price for your workload?
The calculator puts these three models next to the other 25 we track — at your volume, token mix and cache share.
Open calculator

Frequently asked questions

Yes — Grok 4.3 has a verified −20% batch discount. The newer Grok 4.5 and Grok Build tiers don't publish a multiplier yet, so we don't apply batch to the ≈$/mo figures. Your async cost may be lower than the on-demand figure where a discount applies.
Our data lists a flat cached-input rate across the Grok tiers rather than one that scales with each model's standard rate. The exact figure is in the per-model table above.
All 28 models → OpenAI pricing → Anthropic pricing → Gemini pricing → DeepSeek pricing → Mistral pricing → Grok alternatives → Cheapest LLM API → Price changelog →