Skip to content

GLM-5.3-FlashX

Zhipu / Z.ai · released 2026 · proprietary API

The faster half of the first native multimodal pair in the GLM-5 series; API-only, with no published weights unlike GLM-5.3 and GLM-5.3-Flash.

Context window

1M

1,000,000 tokens

Input price

$0.37

per 1M tokens

Output price

$1.25

per 1M tokens

Blended

$0.59

3:1 in:out mix

Weights

Proprietary

API only

Released

2026

Zhipu / Z.ai

Long-context / tier pricingCached input $0.075/1M. Max output 128K. ~200 tokens/s. Not yet on the GLM Coding Plan. Listed prices are the base rate for prompts up to ~200K tokens.

Where it ranks

Of 32 models tracked.

Cost to run

#7

of 32

Input price

#7

of 32

Context window

#6(tied)

of 32

See the full Frontier Index →

Compare with

The nearest models by blended cost to run.

Put it to work

Specs from BitByteCore's hand-verified Frontier Ledger (v2026-09-19), this row verified 19 Sep 2026 against Zhipu / Z.ai's official pricing. Model pricing moves fast, so confirm with Zhipu / Z.ai before you rely on a figure.