Skip to content
Bencher

Models · SpaceXAI (formerly xAI) · Out sinceReleased 10 Mar 2026

Grok 4.20

Available on OpenRouterOn OpenRouterAvailable nowGenerally availableClosed (can't be downloaded)Proprietary

Grok 4.20 is made by SpaceXAI (formerly xAI). We don't have enough test results yet to rank it. It's mid-priced to use.

GA in API 2026-03-10 together with Grok 4.20 Multi-agent (x-ai/grok-4.20-multi-agent). Reasoning and non-reasoning variants. No vendor benchmarks found.

Writing & creativity
—
Research & analysis
—
Coding
—
Price
Mid-priced$1.25 / $2.50
Price per 1M (blended)Blended / 1M
$1.56
MemoryContext
1M
Longest answerMax output
—
Test resultsResults
1 (1 independent1 indep.)
Out sinceReleased
10 Mar 2026
Made byVendor
SpaceXAI (formerly xAI)
UnderstandsInputs
text, image

Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.

Thinking levels

Reasoning settings

You can't choose how long this one thinks.

No adjustable reasoning setting listed.

Read more

Links

Using it through OpenRouter

On OpenRouter

OpenRouter is a service that gives access to many AI models in one place. This model is listed there as x-ai/grok-4.20.

Route it as x-ai/grok-4.20 at $1.25 in / $2.50 out per 1M tokens, 2M context. Listed since 31 Mar 2026.

include_reasoninglogprobsmax_tokensreasoningresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_logprobstop_p

Every test result

Every result

All the numbers,with where they came from.

Every result we've found for this model, with who measured it and a link to where we read it.

All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.

All results for this model
NotesNotes
LMArena Search Arena1189Defaultgrok-4.20-beta1Independent testIndependentLMArena ↗24 Aug 2026—rank 18 (rank range 14-19); 95% CI 1183.2-1195.3; 53921 votes