reroute

Z.ai: GLM 5.3 Flash (batch)

z-ai/glm-5.3-flash:batch

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Modalities
In / Out Price
$0.06 / $0.20 per 1M
Context
1M
Released
Aug 26, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.06 / M
Output tokens
$0.20 / M
Cache read
$0.012 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
DeepInfra$0.06$0.20$0.012—50%