GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 203K | $0.40 | $1.75 | $0.08 | — | — | 100.00% | |
| 198K | $0.4 | $1.93 | $0.0801 | — | — | 100.00% | |
| 205K | $0.54 | $1.98 | $0.099 | — | — | 100.00% | |
| 200K | $0.60 | $2.20 | — | — | — | 100.00% | |
| 203K | $0.60 | $2.20 | $0.11 | — | — | 100.00% | |
| 131K | $0.70 | $2.50 | — | — | — | 100.00% |