Skip to content
Bencher

Models · Z.ai (Zhipu) · Out sinceReleased 18 Sep 2026

GLM-5.3-FlashX

Available on OpenRouterOn OpenRouterAvailable nowGenerally availableClosed (can't be downloaded)Proprietary

GLM-5.3-FlashX is made by Z.ai (Zhipu). We don't have enough test results yet to rank it. It's cheap to use.

High-speed (up to ~200 tok/s) serving variant of GLM-5.3-Flash. Price from Z.ai pricing page; no separate benchmarks.

Writing & creativity
—
Research & analysis
—
Coding
—
Price
Cheap$0.37 / $1.25
Price per 1M (blended)Blended / 1M
$0.59
MemoryContext
1M
Longest answerMax output
131K
Test resultsResults
0 (0 independent0 indep.)
Out sinceReleased
18 Sep 2026
Made byVendor
Z.ai (Zhipu)
UnderstandsInputs
text, image, video

Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.

Thinking levels

Reasoning settings

LowHighMax

Read more

Links

Using it through OpenRouter

On OpenRouter

OpenRouter is a service that gives access to many AI models in one place. This model is listed there as z-ai/glm-5.3-flashx.

Route it as z-ai/glm-5.3-flashx at $0.37 in / $1.25 out per 1M tokens, 1M context. Listed since 18 Sep 2026.

include_reasoningmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetoolstop_ktop_p

Every test result

Every result

All the numbers,with where they came from.

Every result we've found for this model, with who measured it and a link to where we read it.

All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.

No results yetNo results yet

Test results for this model appear after our next daily check.Benchmark results for this model appear after the next daily run.