Anthropic API pricing
15 Anthropic models tracked, updated daily. Full rates below, with guidance on which tier fits which workload.
Cached input rates and full detail on the pricing page. Prices in USD per 1M tokens.
Anthropic's lineup is more concentrated than most providers': a Standard tier (the Sonnet line) and a Premium tier (Opus, plus the Fable flagship and the Fast-suffixed Opus variants). There's no Budget-tier Claude model in the current lineup, which matters for how you should think about routing — see the live table above for current rates.
Sonnet vs. Opus — the real trade-off
Claude Sonnet models are Anthropic's workhorse tier: strong general-purpose capability at a meaningfully lower price than Opus, and the right default for most chat, summarization, and moderate-complexity coding workloads. Opus is priced as a genuine premium tier — reserve it for tasks where Sonnet's capability ceiling is the actual bottleneck, not as a default upgrade "just in case."
A practical test: run a representative sample of your workload on Sonnet first. If quality is consistently sufficient, stay there — the price gap to Opus is large enough that upgrading "for safety" without evidence of a capability gap is usually a wasted spend rather than a genuine hedge.
The "(Fast)" Opus variants
The Fast-suffixed Opus models trade some latency characteristics for the same Premium-tier pricing bracket as standard Opus. If your application is latency-sensitive and you've already decided you need Opus-level capability, the Fast variant is worth checking — but it doesn't change the fundamental Sonnet-vs-Opus cost decision, which is about capability need, not speed.
Anthropic's prompt caching is unusually strong
Anthropic was an early mover on prompt caching and it remains one of the more impactful discounts available on any provider — commonly cutting the cached portion of input to a fraction of the standard rate. Because Claude is frequently used in agentic and tool-calling setups (where the tool schema and system instructions repeat on every step of a loop), caching often has an outsized effect on Anthropic bills specifically compared to providers where caching support is newer or narrower. See Prompt Caching Explained for the exact mechanics and a break-even calculation you can apply directly to a Sonnet or Opus workload.
Coding-specific guidance
Claude models — Sonnet in particular — consistently perform well on agentic coding tasks in third-party evaluations. Before assuming you need Opus for a coding workflow, check the SWE-bench Verified and Aider Polyglot cost-vs-quality charts: Sonnet frequently sits closer to the value frontier than Opus for coding specifically, even when Opus leads on raw score.
When Anthropic vs. a competitor
Claude tends to be priced as a premium-positioned option relative to the lower-cost providers, and the case for choosing it is usually capability- or reliability-driven rather than price-driven — strong instruction-following, long-context handling, and agentic tool use are commonly cited reasons teams pay the premium. If price is the primary constraint, compare directly against OpenAI, Google, or a lower-cost provider on the compare page before assuming Claude is worth the gap for your specific task.
Check current Sonnet-vs-Opus pricing live above, or put Claude head-to-head with any other model on compare.
