GLM-5.3-FlashX
Zhipu / Z.ai · released 2026 · proprietary API
- 1M context
- Multimodal (video/image/text/file in)
The faster half of the first native multimodal pair in the GLM-5 series; API-only, with no published weights unlike GLM-5.3 and GLM-5.3-Flash.
Context window
1M
1,000,000 tokens
Input price
$0.37
per 1M tokens
Output price
$1.25
per 1M tokens
Blended
$0.59
3:1 in:out mix
Weights
Proprietary
API only
Released
2026
Zhipu / Z.ai
Long-context / tier pricing — Cached input $0.075/1M. Max output 128K. ~200 tokens/s. Not yet on the GLM Coding Plan. Listed prices are the base rate for prompts up to ~200K tokens.
Where it ranks
Of 32 models tracked.
Cost to run
#7
of 32
Input price
#7
of 32
Context window
#6(tied)
of 32
Compare with
The nearest models by blended cost to run.
Put it to work
Specs from BitByteCore's hand-verified Frontier Ledger (v2026-09-19), this row verified 19 Sep 2026 against Zhipu / Z.ai's official pricing ↗. Model pricing moves fast, so confirm with Zhipu / Z.ai before you rely on a figure.