Skip to content

Grok 3 vs Grok 4

Data reconciled August 15, 2026

Short answer

Grok 3 and Grok 4 each win on different axes — neither is simply cheaper or better across the board.

Grok 4 holds a 2× larger context window.

Grok 4 returns a 2× longer answer per call.

Side by side

ParameterGrok 3Grok 4
ProviderxAIxAI
Input, $ per 1M tokens$3.00$3.00
Output, $ per 1M tokens$15.00$15.00
Cache read, $ per 1M$0.75
Context window131,072256,000
Max output131,072256,000
Visionnono
Toolsyesyes
Reasoningnono
Statusavailableavailable

What it costs on your own workload

A price gap per million tokens says little until your volumes are plugged in: on a long fixed system prompt the winner is the model with cheap cache reads, not the one with a cheap input rate. Run Grok 3 and Grok 4 against your own numbers.

Open the cost calculator

What is deliberately absent

We do not compare answer quality and we do not reprint third-party benchmarks. We have no quality measurements of our own, and external tables go stale faster than prices while being compiled, almost always, by someone with a stake in the result. What is here is only what we reconcile ourselves every day: prices, limits and supported capabilities.

Other comparisons

Model catalog