OpenAI
GPT-6 Luna pricing and benchmarks
OpenAI’s most efficient GPT-6 model, for focused, high-volume tasks.
Input
$0.10
per 1M tokens
Cached input
$0.01
per 1M tokens
Output
$0.50
per 1M tokens
How capable is GPT-6 Luna?
Epoch AI has not published capability scores for GPT-6 Luna yet. See benchmarks for rated models.
What GPT-6 Luna costs in practice
| Workload | Monthly cost |
|---|---|
Support chatbot 100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each. | $35.00 |
Document Q&A (RAG) 50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each. | $55.00 |
Coding agent 10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens. | $24.80 |
Pricing details
- Cache writes are billed at $0.125 per 1M tokens.
- Fast mode costs 2x the standard rate; Flex costs half.
- Regional (data residency) endpoints add a 10% uplift, and EU data residency is available only with standard processing.
Estimate your GPT-6 Luna bill
- GPT-6 LunaOpenAI$55.00/mo$0.0011 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.