Home/OpenAI/o4-mini
OpenAI pricing

o4-mini API Pricing & Cost Calculator

o4-mini is OpenAI's reasoning-focused model optimized for STEM, coding, and math. Chain-of-thought strong at a fraction of o-series flagship cost.

OpenAI plans to retire o4-mini on 2026-10-23

Pricing below is still current. Worth factoring the retirement date into anything you are building on it long-term.

Still available from OpenAI: GPT-5.5 Pro, GPT-5.6 Cyber, GPT-6 Astra

Input
$1.10
per 1M tokens
Output
$4.40
per 1M tokens
Context window
200K
tokens
Released
2025-04
Cutoff 2024-10
✓ Exact tokenizer·$1.10 in·$4.40 out (per 1M)
Quick start with a use case
Total cost per call$0.003300
Input$0.001100
Output$0.002200
Cost comparison
Standard
$0.003300
With Caching
$0.002888
Save 12% ↓
With Batch
Not supported
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Detailed pricing

o4-mini pricing breakdown

All pricing dimensions including caching and batch discounts.

TypePrice (per 1M tokens)
Input$1.1000
Output$4.4000
Cached input$0.2750
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Last verified 2026-08-24 · OpenAI official pricing · ⚠️ Spotted a wrong price? Report in 30s →

How it compares

o4-mini vs alternatives

Single-call cost (1000 input + 500 output tokens) ranked from cheapest.

ModelPer call
o4-mini
OpenAI · this page
$0.003300
DeepSeek V3.2
DeepSeek
$0.000480
GPT-5.6 Luna
OpenAI
$0.000800
DeepSeek V4-Flash
DeepSeek
$0.000900
Gemini 3.1 Flash Lite
Google
$0.001000
GPT-5 mini
OpenAI
$0.001250
Recommended use

When to choose o4-mini

o4-mini shines for complex multi-step reasoning, code generation and review, mathematical problem-solving, and scientific research and analysis. Token counts on this page are exact via the official tokenizer.

Context window of 200K tokens handles long conversations and large documents.
Prompt caching available — significant savings for repeated system prompts.
Tool / function calling supported.
FAQ

Frequently asked questions

o4-mini costs $1.10 per 1M input tokens and $4.40 per 1M output tokens. A typical chat call (1000 input + 500 output tokens) costs approximately $0.0033. Use the calculator above to estimate your specific use case.
o4-mini supports a 200K token context window with a max output of 100,000 tokens. Knowledge cutoff: 2024-10.
Yes. Cached input is priced at $0.28 per 1M tokens — 75% cheaper than uncached input. This is especially valuable for repeated system prompts, long-context retrieval, and chat threads with shared history.
Yes, o4-mini accepts image input. Vision token pricing is generally calculated based on image dimensions and folded into the input token count. Specific image-specific pricing varies — refer to OpenAI's official documentation.
Token counts for o4-mini use the official o200k_base tokenizer (exact). Cost calculations use the prices verified on 2026-08-24. For final billing accuracy, always verify with OpenAI's usage dashboard.
o4-mini does not currently support fine-tuning. Consider another model in the OpenAI family if fine-tuning is required.