The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.
Each carrier sets its own price. The router weighs these against uptime when it picks one.
| Provider | Input /M | Output /M | Cache read /M | Cache write /M | Discount |
|---|---|---|---|---|---|
| $0.39 | $2.34 | — | — | — | |
| $0.45 | $3 | $0.22 | — | — | |
| $0.50 | $3.60 | $0.30 | — | — | |
| $0.55 | $3.50 | $0.11 | — | — | |
| $0.55 | $3.50 | $0.225 | — | — | |
| $0.55 | $3.50 | $0.55 | — | — | |
| $0.60 | $3.60 | $0.12 | — | — | |
| $0.60 | $3.60 | — | — | — | |
| $0.60 | $3.60 | — | — | — | |
| $0.75 | $4.50 | — | — | — |