Polarison

OpenAI

GPT-6 Luna pricing and benchmarks

OpenAI’s most efficient GPT-6 model, for focused, high-volume tasks.

Input
$0.10
per 1M tokens
Cached input
$0.01
per 1M tokens
Output
$0.50
per 1M tokens

How capable is GPT-6 Luna?

Epoch AI has not published capability scores for GPT-6 Luna yet. See benchmarks for rated models.

What GPT-6 Luna costs in practice

WorkloadMonthly cost
Support chatbot
100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each.
$35.00
Document Q&A (RAG)
50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each.
$55.00
Coding agent
10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens.
$24.80

Pricing details

  • Cache writes are billed at $0.125 per 1M tokens.
  • Fast mode costs 2x the standard rate; Flex costs half.
  • Regional (data residency) endpoints add a 10% uplift, and EU data residency is available only with standard processing.

Estimate your GPT-6 Luna bill

Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.