Home/OpenAI/GPT-6 Luna
OpenAI pricing

GPT-6 Luna API Pricing & Cost Calculator

GPT-6 Luna is priced at $0.10 input / $0.50 output per 1M tokens, with a 922K context window and 128K max output. Cached input $0.01/1M. That is 56% cheaper per call than GPT-5.6 Luna ($0.20/$1.20). A 1,000-in / 500-out call costs $0.00035.

Listed automatically on 2026-09-23

Prices and limits come straight from the LiteLLM registry and are verified daily. The description is generated from those numbers and has not yet been reviewed by a person, so it may miss context about how this model relates to others in its family.

Input
$0.10
per 1M tokens
Output
$0.50
per 1M tokens
Context window
922K
tokens
Released
Cutoff
✓ Exact tokenizer·$0.10 in·$0.50 out (per 1M)
Quick start with a use case
Total cost per call$0.000350
Input$0.000100
Output$0.000250
Cost comparison
Standard
$0.000350
With Caching
$0.000305
Save 13% ↓
With Batch
$0.000175
Save 50% ↓
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Detailed pricing

GPT-6 Luna pricing breakdown

All pricing dimensions including caching and batch discounts.

TypePrice (per 1M tokens)
Input$0.1000
Output$0.5000
Cached input$0.0100
Cache write$0.1250
Batch input$0.0500
Batch output$0.2500
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Last verified 2026-09-23 · OpenAI official pricing · ⚠️ Spotted a wrong price? Report in 30s →

How it compares

GPT-6 Luna vs alternatives

Single-call cost (1000 input + 500 output tokens) ranked from cheapest.

ModelPer call
GPT-6 Luna
OpenAI · this page
$0.000350
DeepSeek V3.2
DeepSeek
$0.000480
GPT-5.6 Luna
OpenAI
$0.000800
DeepSeek V4-Flash
DeepSeek
$0.000900
Gemini 3.1 Flash Lite
Google
$0.001000
GPT-5 mini
OpenAI
$0.001250
Recommended use

When to choose GPT-6 Luna

GPT-6 Luna shines for general-purpose tasks, complex multi-step reasoning, image and document understanding, long documents and large codebases, high-throughput, and low-latency tasks. Token counts on this page are exact via the official tokenizer.

Context window of 922K tokens handles long conversations and large documents.
Prompt caching available — significant savings for repeated system prompts.
Batch API support for non-realtime workloads at ~50% discount.
Tool / function calling supported.
FAQ

Frequently asked questions

GPT-6 Luna costs $0.10 per 1M input tokens and $0.50 per 1M output tokens. A typical chat call (1000 input + 500 output tokens) costs approximately $0.0003. Use the calculator above to estimate your specific use case.
GPT-6 Luna supports a 922K token context window with a max output of 128,000 tokens.
Yes. Cached input is priced at $0.01 per 1M tokens — 90% cheaper than uncached input. This is especially valuable for repeated system prompts, long-context retrieval, and chat threads with shared history.
Yes. GPT-6 Luna supports the Batch API at $0.05 input / $0.25 output per 1M tokens — typically 50% off standard pricing. Batch jobs complete within ~24 hours, ideal for non-realtime workloads like overnight data processing or content generation pipelines.
Yes, GPT-6 Luna accepts image input. Vision token pricing is generally calculated based on image dimensions and folded into the input token count. Specific image-specific pricing varies — refer to OpenAI's official documentation.
Token counts for GPT-6 Luna use the official o200k_base tokenizer (exact). Cost calculations use the prices verified on 2026-09-23. For final billing accuracy, always verify with OpenAI's usage dashboard.
GPT-6 Luna does not currently support fine-tuning. Consider another model in the OpenAI family if fine-tuning is required.