reroute

OpenAI: GPT-5.6 Luna

openai/gpt-5.6-luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

Modalities
In / Out Price
$0.20 / $1.20 per 1M
Context
1.1M
Released
Jul 9, 2026
Knowledge Cutoff
Feb 2026

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.20 / M
Output tokens
$1.20 / M
Cache read
$0.02 / M
Cache write
$0.25 / M
Web search
$10 / K requests
Long-context pricing

Prompts over 272K tokens: $0.40 input / $1.80 output per M.

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
OpenAI$0.10$0.60$0.01$0.125—
Azure$0.20$1.20$0.02$0.25—
OpenAI$0.20$1.20$0.02$0.25—
Azure$0.22$1.32$0.022$0.275—
Amazon Bedrock$0.22$1.32$0.022$0.275—
Azure$0.22$1.32$0.022$0.275—
OpenAI$0.40$2.40$0.04$0.50—