GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 1.1M | $1.25 | $7.50 | $0.125 | — | — | 33.33% | |
| 1.1M | $2.50 | $15 | $0.25 | — | — | 100.00% | |
| 1.1M | $2.50 | $15 | $0.25 | — | — | 100.00% | |
| 1.1M | $2.75 | $16.50 | $0.275 | — | — | — | |
| 1.1M | $2.75 | $16.50 | $0.275 | — | — | 100.00% | |
| 1.1M | $2.75 | $16.50 | $0.275 | — | — | — | |
| 1.1M | $5 | $30 | $0.50 | — | — | — |