Home/OpenAI/GPT-5.6 Luna
OpenAI pricing

GPT-5.6 Luna API Pricing & Cost Calculator

GPT-5.6 Luna is the efficient tier of the GPT-5.6 family. After the July 2026 price cut it runs at $0.20/$1.20 per 1M tokens — 20x cheaper than the flagship — while keeping the full 922K context window. Cached input at $0.02/1M.

Input
$0.20
per 1M tokens
Output
$1.20
per 1M tokens
Context window
922K
tokens
Released
2026-06
Cutoff
✓ Exact tokenizer·$0.20 in·$1.20 out (per 1M)
Quick start with a use case
Total cost per call$0.000800
Input$0.000200
Output$0.000600
Cost comparison
Standard
$0.000800
With Caching
$0.000710
Save 11% ↓
With Batch
$0.000400
Save 50% ↓
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Detailed pricing

GPT-5.6 Luna pricing breakdown

All pricing dimensions including caching and batch discounts.

TypePrice (per 1M tokens)
Input$0.2000
Output$1.2000
Cached input$0.0200
Cache write$0.2500
Batch input$0.1000
Batch output$0.6000
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Last verified 2026-08-24 · OpenAI official pricing · ⚠️ Spotted a wrong price? Report in 30s →

How it compares

GPT-5.6 Luna vs alternatives

Single-call cost (1000 input + 500 output tokens) ranked from cheapest.

ModelPer call
GPT-5.6 Luna
OpenAI · this page
$0.000800
DeepSeek V3.2
DeepSeek
$0.000480
Gemini 3.1 Flash Lite
Google
$0.001000
DeepSeek V4-Flash
DeepSeek
$0.001100
GPT-5 mini
OpenAI
$0.001250
Mistral Large 3
Mistral
$0.001250
Recommended use

When to choose GPT-5.6 Luna

GPT-5.6 Luna shines for general-purpose tasks, image and document understanding, high-throughput, and low-latency tasks. Token counts on this page are exact via the official tokenizer.

Context window of 922K tokens handles long conversations and large documents.
Prompt caching available — significant savings for repeated system prompts.
Batch API support for non-realtime workloads at ~50% discount.
Tool / function calling supported.
FAQ

Frequently asked questions

GPT-5.6 Luna costs $0.20 per 1M input tokens and $1.20 per 1M output tokens. A typical chat call (1000 input + 500 output tokens) costs approximately $0.0008. Use the calculator above to estimate your specific use case.
GPT-5.6 Luna supports a 922K token context window with a max output of 128,000 tokens.
Yes. Cached input is priced at $0.02 per 1M tokens — 90% cheaper than uncached input. This is especially valuable for repeated system prompts, long-context retrieval, and chat threads with shared history.
Yes. GPT-5.6 Luna supports the Batch API at $0.10 input / $0.60 output per 1M tokens — typically 50% off standard pricing. Batch jobs complete within ~24 hours, ideal for non-realtime workloads like overnight data processing or content generation pipelines.
Yes, GPT-5.6 Luna accepts image input. Vision token pricing is generally calculated based on image dimensions and folded into the input token count. Specific image-specific pricing varies — refer to OpenAI's official documentation.
Token counts for GPT-5.6 Luna use the official o200k_base tokenizer (exact). Cost calculations use the prices verified on 2026-08-24. For final billing accuracy, always verify with OpenAI's usage dashboard.
GPT-5.6 Luna does not currently support fine-tuning. Consider another model in the OpenAI family if fine-tuning is required.