Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 262K | $0.09 | $1.10 | — | — | — | 100.00% | |
| 131K | $0.0975 | $0.78 | — | — | — | 100.00% | |
| 262K | $0.10 | $1.10 | $0.07 | — | — | 100.00% | |
| 262K | $0.15 | $1.20 | — | — | — | 100.00% |