PROVIDER GUIDE · VERIFIED 13 JUL 2026
How much does Claude API cost?
Claude API pricing varies by model: Claude 3.5 Sonnet costs $3 per million input tokens and $15 per million output tokens, Claude 3 Opus costs $15/$75, and Claude 3 Haiku costs $0.25/$1.25. All models use the same token counting method, and prices are billed per-token with no monthly minimums or subscription fees required.
Anthropic runs four price tiers across its Claude line. Here's what each model actually costs, and where the real differences show up on your bill.
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Claude Haiku 4.5 | $1.00 | $5.00 | 200K |
| Claude Sonnet 5* | $2.00 | $10.00 | 1M |
| Claude Opus 4.8 | $5.00 | $25.00 | 1M |
| Claude Fable 5 | $10.00 | $50.00 | 1M |
* Launched at $2.00 / $10.00 as an introductory rate. Anthropic kept that rate instead of moving to the announced $3.00 / $15.00 on 1 Sep 2026.
The one date to put on your calendar
Claude Sonnet 5 is priced at $2.00/$10.00 per million tokens. It launched with that as an introductory rate and a planned move to $3.00/$15.00 on 1 September 2026, but the increase never happened and $2/$10 is still the live price. Worth knowing if you budgeted for the higher rate: your Sonnet 5 bill is a third lower than that plan assumed. For a workload running 500M input and 100M output tokens a month, that's the difference between a $1,000/month bill and a $1,500/month bill.
Picking the right tier for the job
- Haiku 4.5 — high-volume, low-complexity tasks: classification, extraction, routing, simple chat. Don't pay Sonnet or Opus rates for work Haiku handles fine.
- Sonnet 5 — the default production choice for most teams: strong reasoning at a mid-tier price.
- Opus 4.8 — complex reasoning, long-horizon agentic tasks, work where being right matters more than being cheap.
- Fable 5 — Anthropic's most capable tier, priced accordingly. Reserve for tasks that genuinely need frontier-level capability.
The lever most people skip: prompt caching
If your prompts include a stable system prompt, a fixed set of instructions, or reference documents that repeat across calls, prompt caching can cut the cost of that repeated portion by roughly 90% on cache hits. For workloads with a large, static context reused across many calls — a coding assistant with a big codebase in context, a support bot with a fixed knowledge base — this is often a bigger saving than choosing a cheaper model tier.
Prices verified against platform.claude.com/docs/about-claude/pricing, 13 July 2026. See full methodology. Use the calculator to compare Claude against every other tracked provider for your own usage.