GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 400K | $0.025 | $0.20 | $0.0025 | — | — | 100.00% | |
| 400K | $0.05 | $0.40 | $0.01 | — | — | 100.00% | |
| 400K | $0.05 | $0.40 | $0.005 | — | — | 100.00% | |
| 400K | $0.055 | $0.44 | $0.011 | — | — | 100.00% |