Anthropic API pricing (2026): the Claude per-token cost table

Short answer: Per million tokens, the Claude API costs $10 in / $50 out for Fable 5.1 (the top model, released Sept 1, 2026), $4 in / $20 out for Opus 5.5 (released Sept 22, 2026), $2 in / $10 out for Sonnet 5.5 (released Sept 28, 2026), and $1 in / $5 out for Haiku 4.5 — the cheapest current model. A prompt-cache read costs 10% of the input rate (Fable 5.1's cache reads are a flat $0.25/MTok), and the Batch API takes 50% off both input and output. Verified against Anthropic's official pricing page.

Claude API model pricing (per 1M tokens)

ModelInputOutputBest for
Claude Fable 5.1 (top model, Sept 2026)$10$50Flagship frontier model; Mythos 5.1 is the same model under restricted-access safeguards
Claude Opus 5.5 (Sept 22, 2026)$4$20Hardest reasoning, agentic and critical work at below-Fable cost
Claude Sonnet 5.5 (Sept 28, 2026)$2$10Production default; best quality-to-cost balance
Claude Opus 5 / 4.8 / 4.7 / 4.6$5$25Prior Opus releases, still available
Claude Sonnet 5$2$10Prior Sonnet release, still available
Claude Sonnet 4.6$3$15Older Sonnet; now costs more than Sonnet 5.5
Claude Haiku 4.5 (cheapest)$1$5Routing, classification, high-volume simple tasks

Prices are USD per million tokens (MTok) for the first-party Claude API with default global routing. Fable 5.1, Opus 5.5, Sonnet 5.5, Opus 5, Opus 4.8/4.7/4.6 and Sonnet 4.6 include the full 1M-token context window at standard per-token pricing (Fable 5.1, Opus 5.5 and Sonnet 5.5 have a 128K-token max output). Haiku 4.5 has a 200K context window. Older models (Opus 4.1 and earlier, at the legacy $15/$75 Opus rate) are deprecated.

What is the cheapest Claude API model?

Haiku 4.5, at $1 / $5 per million tokens, is the cheapest current Claude model. You can push the effective cost lower in three ways, which stack:

If you need to go cheaper than Haiku for trivial tasks, Google's Gemini Flash tier undercuts it — see OpenAI & Claude API alternatives.

Prompt caching and Batch API discounts

FeatureMultiplier vs base inputEffect
5-minute cache write1.25xPays off after one cache read
1-hour cache write2xPays off after two cache reads
Cache read (hit)0.1x90% off repeated context
Batch API0.5x in & out50% off async (non-time-sensitive) jobs

Web search as a server tool is billed separately at $10 per 1,000 searches plus token costs. Caching and batching discounts stack with each other. Claude Fable 5.1 is the exception on caching: instead of the 0.1x multiplier, its cache reads are priced at a flat $0.25 per million tokens — a 75% cut from Fable 5's $1.00 cache-read rate.

Worked example: cost of a typical Sonnet 5.5 request

A request that sends 20,000 input tokens and generates 2,000 output tokens on Sonnet 5.5:

If 16,000 of the input tokens are served from a prompt cache (0.1x), the input drops to (4,000 × $2 + 16,000 × $0.20) / 1,000,000 = $0.0112, taking the request to about $0.031 — roughly half.

How Claude API pricing compares to OpenAI

At the top tier, Claude Fable 5.1 ($10 in / $50 out) is priced identically to OpenAI's GPT-6 Astra (released Sept 2026). In the middle, Claude Sonnet 5.5 and OpenAI's GPT-6.1 Sol both cost $2 / $10, with Claude Opus 5.5 ($4 / $20) sitting between the tiers. OpenAI keeps the edge at the cheap end — GPT-6 Luna ($0.10 / $0.50) undercuts Haiku 4.5 ($1 / $5). For the full side-by-side, see ChatGPT vs Claude API cost and the OpenAI API pricing table.