reroute

Z.ai: GLM 5.3 FlashX

z-ai/glm-5.3-flashx

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Modalities
In / Out Price
$0.37 / $1.25 per 1M
Context
1M
Released
Sep 18, 2026
Knowledge Cutoff
—

Activity

Tokens processed for this model through Reroute, per day, over the last 30 days.

No usage through Reroute yet

Daily token volume shows up here as soon as the first request for this model completes.