Home/Anthropic/Claude Opus 4.8
Anthropic pricing

Claude Opus 4.8 API Pricing & Cost Calculator

Claude Opus 4.8 succeeds Opus 4.7 at identical pricing ($5/$25) with stronger coding and agentic performance. 1M context, 128K max output, 90% cache discount.

Anthropic plans to retire Claude Opus 4.8 on 2027-05-28

Pricing below is still current. Worth factoring the retirement date into anything you are building on it long-term.

Still available from Anthropic: Claude Fable 5, Claude Mythos 5, Claude Opus 5

Input
$5.00
per 1M tokens
Output
$25.00
per 1M tokens
Context window
1000K
tokens
Released
2026-06
Cutoff
≈ Estimated tokenizer·$5.00 in·$25.00 out (per 1M)
Quick start with a use case
Total cost per call$0.0175
Input$0.005000
Output$0.0125
Cost comparison
Standard
$0.0175
With Caching
$0.0152
Save 13% ↓
With Batch
$0.008750
Save 50% ↓
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Detailed pricing

Claude Opus 4.8 pricing breakdown

All pricing dimensions including caching and batch discounts.

TypePrice (per 1M tokens)
Input$5.0000
Output$25.0000
Cached input$0.5000
Cache write$6.2500
Batch input$2.5000
Batch output$12.5000
Cutting costs? Open models start at $0.28/1M

DeepSeek, Kimi & 200+ open models on Novita AI — often 5–20x cheaper than proprietary APIs for comparable tasks.

Try Novita AI

Affiliate link — we may earn a commission at no extra cost to you.

Last verified 2026-08-01 · Anthropic official pricing · ⚠️ Spotted a wrong price? Report in 30s →

How it compares

Claude Opus 4.8 vs alternatives

Single-call cost (1000 input + 500 output tokens) ranked from cheapest.

ModelPer call
Claude Opus 4.8
Anthropic · this page
$0.0175
DeepSeek V3.2
DeepSeek
$0.000480
GPT-5.6 Luna
OpenAI
$0.000800
Gemini 3.1 Flash Lite
Google
$0.001000
DeepSeek V4-Flash
DeepSeek
$0.001100
GPT-5 mini
OpenAI
$0.001250
Recommended use

When to choose Claude Opus 4.8

Claude Opus 4.8 shines for complex multi-step reasoning, code generation and review, general-purpose tasks, and long documents and large codebases. Token counts are estimated within ~10-20% margin.

Context window of 1000K tokens handles entire codebases or book-length documents.
Prompt caching available — significant savings for repeated system prompts.
Batch API support for non-realtime workloads at ~50% discount.
Tool / function calling supported.
FAQ

Frequently asked questions

Claude Opus 4.8 costs $5.00 per 1M input tokens and $25.00 per 1M output tokens. A typical chat call (1000 input + 500 output tokens) costs approximately $0.0175. Use the calculator above to estimate your specific use case.
Claude Opus 4.8 supports a 1000K token context window with a max output of 128,000 tokens.
Yes. Cached input is priced at $0.50 per 1M tokens — 90% cheaper than uncached input. This is especially valuable for repeated system prompts, long-context retrieval, and chat threads with shared history.
Yes. Claude Opus 4.8 supports the Batch API at $2.50 input / $12.50 output per 1M tokens — typically 50% off standard pricing. Batch jobs complete within ~24 hours, ideal for non-realtime workloads like overnight data processing or content generation pipelines.
Yes, Claude Opus 4.8 accepts image input. Vision token pricing is generally calculated based on image dimensions and folded into the input token count. Specific image-specific pricing varies — refer to Anthropic's official documentation.
Token counts for Claude Opus 4.8 are estimated from character ratios (~10-20% margin). Cost calculations use prices verified on 2026-08-01. For final billing accuracy, always verify with Anthropic's usage dashboard.
Claude Opus 4.8 does not currently support fine-tuning. Consider another model in the Anthropic family if fine-tuning is required.