GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...
Top apps by tokens routed through Reroute in the last 30 days. Apps appear here when they send an X-Title header.
Once apps call this model with an X-Title header, they're ranked here by tokens.