OpenAI · Small model

GPT-5.6 Luna pricing & cost calculator

Prices last updated:

Input

$0.20

per 1M tokens

Output

$1.20

per 1M tokens

Context window

1.05M

1,050,000 tokens

Inputs above 272K tokens are billed at a higher long-context rate ($0.40 / $1.80).

About GPT-5.6 Luna pricing

GPT-5.6 Luna is OpenAI’s small, low-cost model built for high-volume, latency-sensitive tasks. It is billed per token: $0.20 for every million tokens you send and $1.20 for every million tokens it generates, so output costs 6.0× as much as input.

For a typical request of 1,000 input and 500 output tokens, GPT-5.6 Luna costs about $0.0008 — the 2nd-cheapest of the 14 models we track. One million input tokens is roughly 2,000 pages of English prose.

With OpenAI’s batch API (50% off, results within 24 hours) the rate drops to $0.10 input and $0.60 output per million tokens.

OpenAI’s official pricing page is the source of truth; see our guide to how AI pricing works for caching and long-context rules.

GPT-5.6 Luna cost calculator

Per request
$0.0008
Per day
$0.80
Per month (30 days)
$24.00

Count tokens from your own text in the full calculator →

What GPT-5.6 Luna costs for common tasks

Standard pricing, no caching or batch discount. Token sizes are typical estimates — see token cost by use case.
TaskInput tokensOutput tokensPer requestPer 1,000 requests
Short chat reply500300$0.00046$0.46
Support chatbot turn with history3,000400$0.00108$1.08
RAG answer (retrieved context)6,000600$0.00192$1.92
Summarize a 20-page document12,0001,000$0.0036$3.60
Long-document analysis100,0002,000$0.0224$22.40

Compare GPT-5.6 Luna with…

Other OpenAI models first.

GPT-5.6 Luna pricing FAQ

How much does GPT-5.6 Luna cost per token?
GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens through the OpenAI API. That is $0.00 per input token and $0.000001 per output token.
How much does a typical GPT-5.6 Luna request cost?
A request with 1,000 input tokens and 500 output tokens costs about $0.0008 on GPT-5.6 Luna, or $0.80 per thousand requests.
What is the context window of GPT-5.6 Luna?
GPT-5.6 Luna accepts up to 1,050,000 tokens (about 1.05M) of context per request, which covers the prompt, any documents you include and the conversation history.
Does GPT-5.6 Luna have batch pricing?
Yes. OpenAI's batch API processes requests asynchronously (typically within 24 hours) at 50% off, bringing GPT-5.6 Luna to $0.10 input and $0.60 output per million tokens.
How accurate is the token count for GPT-5.6 Luna?
Exact. Tokenova counts GPT-5.6 Luna tokens in your browser with the o200k_base tokenizer that OpenAI's current models use.