Skip to content
API Pricing

Claude API Pricing 2026:
Fable 5.1, Opus 5.5, Sonnet 5.5 & Haiku 4.5 Costs

Current Anthropic Claude API pricing for Fable 5.1, Opus 5.5, Sonnet 5.5 and Haiku 4.5, plus the previous generation still on the API. Compare token costs, prompt caching, batch discounts and context windows. Last verified: 2026-10-01.

10 min read·Updated October 2026
Claude API Cost at a Glance
$1.00
Haiku 4.5 input / 1M
$2.00
Sonnet 5.5 input / 1M
$4.00
Opus 5.5 input / 1M
$10.00
Fable 5.1 input / 1M

Anthropic's current Claude lineup has four tiers: Claude Fable 5.1, its most capable widely released model, for the hardest reasoning and long-horizon agentic work; Claude Opus 5.5 for complex analysis and agents; Claude Sonnet 5.5 as the everyday production default; and Claude Haiku 4.5 for high-volume, cost-sensitive workloads. Fable 5.1, Opus 5.5 and Sonnet 5.5 all have a 1 million token context window; Haiku 4.5 has 200K.

Claude API Standard Pricing 2026

ModelInput / 1M tokensOutput / 1M tokensContext windowBest for
Claude Fable 5.1$10.00$50.001M tokensHardest reasoning and long-horizon agentic work
Claude Opus 5.5$4.00$20.001M tokensComplex analysis, agents, demanding coding
Claude Sonnet 5.5$2.00$10.001M tokensProduction default for coding, agents and enterprise work
Claude Haiku 4.5$1.00$5.00200K tokensHigh-volume, lower-cost workloads
What changed in 2026: Opus 5.5 costs $4/$20 per million tokens — less than Opus 4.6 ($5/$25) — and Sonnet 5.5 costs $2/$10, down from Sonnet 4.6's $3/$15. Claude Mythos 5.1 shares Fable 5.1's pricing but is available only to approved organizations, so it is not listed here.

Previous-Generation Claude Models Still on the API

ModelInput / 1MOutput / 1MContext
Claude Fable 5$10.00$50.001M
Claude Opus 5$5.00$25.001M
Claude Opus 4.8 / 4.7 / 4.6$5.00$25.001M
Claude Sonnet 5$2.00$10.001M
Claude Sonnet 4.6$3.00$15.001M

For new projects, Opus 5.5 and Sonnet 5.5 are cheaper than the models they replace, so migrating usually lowers the bill as well as improving quality.

Prompt Caching Pricing

Prompt caching bills repeated prompt prefixes — system prompts, tool definitions, shared documents — at a fraction of the normal input price. Cache reads cost:

ModelStandard input / 1MCache read / 1MDiscount
Claude Fable 5.1$10.00$0.2597.5%
Claude Opus 5.5$4.00$0.2095%
Claude Sonnet 5.5$2.00$0.2090%
Claude Haiku 4.5$1.00$0.1090%

Writing a prompt into the cache costs more than a normal input token, so caching pays off when the same prefix is reused across many requests.

Batch API Pricing

Anthropic's Message Batches API gives a 50% discount for non-real-time workloads (classification, data processing, evals). Batches complete within 24 hours.

ModelBatch Input / 1MBatch Output / 1M
Claude Fable 5.1$5.00$25.00
Claude Opus 5.5$2.00$10.00
Claude Sonnet 5.5$1.00$5.00
Claude Haiku 4.5$0.50$2.50

Claude vs OpenAI: Quick Price Comparison

TierClaude modelClaude in / out per 1MOpenAI modelOpenAI in / out per 1M
BudgetHaiku 4.5$1.00 / $5.00GPT-6 Luna$0.10 / $0.50
BalancedSonnet 5.5$2.00 / $10.00GPT-6.1 Sol$2.00 / $10.00
PremiumOpus 5.5$4.00 / $20.00GPT-6 Astra$10.00 / $50.00
Top tierFable 5.1$10.00 / $50.00GPT-6 Astra$10.00 / $50.00
ContextSonnet / Opus / Fable1MGPT-6 family1.05M

GPT-6 Luna is far cheaper than Haiku 4.5 for simple, high-volume tasks. In the middle of the market Sonnet 5.5 and GPT-6.1 Sol cost exactly the same, and Opus 5.5 costs less than half of GPT-6 Astra. OpenAI prices above 272K input tokens double for input; check long-context rates before sending very large prompts.

Which Claude Model Should You Choose?

Use caseBest modelReason
High-volume chatbots, classificationHaiku 4.5Lowest Claude cost, handles most routine tasks
Customer support, summarizationSonnet 5.5Best quality-to-cost ratio, 1M context
Complex coding, content creationSonnet 5.5Production default for demanding tasks
Complex agentic workflows, researchOpus 5.5Higher accuracy on long, multi-step work at $4/M input
Hardest problems where quality outweighs costFable 5.1Anthropic's most capable widely released model

Real-World Cost Examples

Customer Support Bot — 50,000 conversations/month

  • System prompt: 2,000 tokens (cached after the first call)
  • Per conversation: 500 tokens fresh input + 300 tokens output
  • Sonnet 5.5 without caching: 125M input × $2 + 15M output × $10 = ~$400/month
  • Sonnet 5.5 with prompt caching: 100M cached × $0.20 + 25M × $2 + $150 output = ~$220/month
  • Savings: about 45% overall — caching removes 90% of the system-prompt cost, and output tokens make up most of what remains

Document Analysis Pipeline — 10,000 docs/month

  • Average: 5,000 tokens input + 800 tokens output (50M input, 8M output per month)
  • Haiku 4.5: $50 input + $40 output = $90/month
  • Sonnet 5.5: $100 input + $80 output = $180/month
  • Opus 5.5: $200 input + $160 output = $360/month

API Access

  • Sign up at console.anthropic.com — no approval process required
  • Pay-as-you-go billing, no monthly minimum
  • Enterprise: custom terms for high-volume spend — contact Anthropic sales

Frequently Asked Questions

Is Claude Sonnet 5.5 the best default production model?

For most teams, yes. It handles complex tasks at $2/M input and $10/M output, and its 1M-token context handles large documents without chunking. Haiku 4.5 is the better choice when volume is high and cost is the main constraint.

Does Claude batch processing reduce cost?

Yes — Anthropic's Message Batches API gives a 50% discount on input and output tokens. Batches complete within 24 hours, which makes them a good fit for classification, evals and other pipelines that are not latency-sensitive.

Is Opus 5.5 more expensive than Opus 4.6?

No. Opus 5.5 is $4/M input and $20/M output, compared with $5/$25 for Opus 4.6, Opus 4.7, Opus 4.8 and Opus 5. Sonnet 5.5 is also cheaper than Sonnet 4.6 ($2/$10 vs $3/$15).

When is Haiku 4.5 worth it over GPT-6 Luna?

GPT-6 Luna ($0.10/M input, $0.50/M output) is ten times cheaper per token than Haiku 4.5 ($1/$5). Haiku can still win on total cost if it needs fewer retries or shorter prompts for your task, so benchmark both on real traffic before choosing.

Which Claude model is best for RAG or coding?

Sonnet 5.5 is the default for both. For RAG, its 1M context window can hold large document sets; for coding, it handles complex multi-file tasks and follows detailed instructions reliably. Opus 5.5 and Fable 5.1 are worth the premium for the longest, most complex reasoning or research tasks.

Calculate Your Claude API Costs

Get an accurate monthly cost estimate for your Claude API usage.

Open API Cost Calculator