Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.
Each carrier sets its own price. The router weighs these against uptime when it picks one.
| Provider | Input /M | Output /M | Cache read /M | Cache write /M | Discount |
|---|---|---|---|---|---|
| $0.0481 | $0.193 | — | — | 55% | |
| $0.09 | $0.30 | — | — | — | |
| $0.09 | $0.30 | — | — | — | |
| $0.10 | $0.30 | — | — | — | |
| $0.13 | $0.52 | — | — | — |