reroute

OpenAI: GPT-6 Luna

openai/gpt-6-luna

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

Modalities
In / Out Price
$0.10 / $0.50 per 1M
Context
1.1M
Released
Sep 22, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.10 / M
Output tokens
$0.50 / M
Cache read
$0.01 / M
Cache write
$0.125 / M
Web search
$10 / K requests
Long-context pricing

Prompts over 272K tokens: $0.20 input / $0.75 output per M.

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
OpenAI$0.05$0.25$0.005$0.0625—
OpenAI$0.10$0.50$0.01$0.125—
Azure$0.10$0.50$0.01$0.125—
Azure$0.11$0.55$0.011$0.138—
Azure$0.11$0.55$0.011$0.138—
Amazon Bedrock$0.11$0.55$0.011$0.138—
OpenAI$0.20$1$0.02$0.25—