reroute

Z.ai: GLM 5.3 FlashX

z-ai/glm-5.3-flashx

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Modalities
In / Out Price
$0.37 / $1.25 per 1M
Context
1M
Released
Sep 18, 2026
Knowledge Cutoff
—

Uptime

Share of successful requests per carrier, measured over the last 5 minutes, 30 minutes and day. When one carrier dips, the router fails over to the next.

Carriers
1
Best carrier, 30m
100.00%
Healthy now
1 / 1
ProviderStatusLast 5mLast 30mLast day
Z.AIOperational100.00%100.00%100.00%