xAI

grok-4.20-0309-reasoning

Made by xAI. Released 2026-02-17. Accessibility not recorded. Cheapest standard rate we hold: $2.50 per million output tokens, at < 200k prompt tokens context.

What it costs

TierKindInOutCached inFrom
standardtext · < 200k prompt tokens$1.25$2.50$0.20seller
standardtext · ≥ 200k prompt tokens$2.50$5$0.40seller

per million tokens, US dollars. * marks a rate from a single registry or two that disagree — the best available number rather than a confirmed one. Every other row is confirmed against a seller or a licensed index.

What it scores, by house

BenchmarkHouseScoreEvaluatedScaffold
chess puzzlesEpoch AI0.2±0.042026-07-13
frontiermath tier 4 v2Epoch AI0.2±0.062026-07-13
frontiermath tiers 1 3 v2Epoch AI0.4±0.032026-07-13
gpqa diamondEpoch AI0.9±0.022026-07-13
otis mock aime 2024 2025near its ceilingEpoch AI0.9±0.032026-07-13
simpleqa verifiedEpoch AI0.3±0.012026-08-27
forecastbenchcollected by Epoch AI60.7
proofbenchcollected by Epoch AI0.1

These scores are not comparable between houses and are not added up. A score is dated by when it was evaluated, not when the model was released, and a score without its scaffold is not a measurement — which is why both are printed.

Every rate we hold · Why there is no single best model · Everything from xAI