Pricing hub

Anthropic Pricing

Anthropic's pricing is straightforward to compare when you are planning complex prompts, especially if you want to evaluate the flagship, balanced, and lightweight tiers.

Live pricing table

Use this table to compare the published input, cached input, and output rates before you run the calculator.

Input 1,000 / Output 250

ModelInput / 1MCached input / 1MOutput / 1MExample requestSource
Claude Opus 4.6

Anthropic's current flagship model for the hardest reasoning, coding, and agentic tasks.

$5.00No public rate$25.0$0.011Anthropic pricing
Claude Sonnet 4.6

Anthropic's current balanced production model for everyday writing, analysis, and automation.

$3.00No public rate$15.0$0.0067Anthropic pricing
Claude Sonnet 4.5

Previous Sonnet tier that remains useful as a stable comparison point.

$3.00No public rate$15.0$0.0067Anthropic pricing
Claude Haiku 4.5

Anthropic’s fastest low-cost model for high-volume, latency-sensitive work.

$1.00No public rate$5.00$0.0023Anthropic pricing

How to use this pricing page

  1. 1. Start with real prompts. Estimate the input and output token counts from the workflow you actually run.
  2. 2. Check the example cost. The table shows a request-level estimate so you can compare the published rate card with your usage pattern.
  3. 3. Validate the savings. Open the LLM cost calculator and test a few prompt sizes before you make a budget call.
  4. 4. Compare the model choice. If you need a second opinion, use one of the linked comparison pages below.

Highlights

  • • Shows the current Claude flagship, balanced, and lightweight tiers.
  • • Useful for procurement and budget conversations that need a clear rate card.
  • • Pairs well with the calculator for request-level forecasting.

Source: Anthropic model overview

FAQ

Which Claude model should I start with?

Use Claude Sonnet 4.6 for balanced production work, Claude Haiku 4.5 for latency-sensitive routines, and Claude Opus 4.6 for the hardest tasks.

How do I use this pricing page?

Start with your real prompt size, estimate output length, then multiply the token mix by the published input and output rates.

Where should I compare Anthropic against other providers?

Use the comparison pages to line it up against OpenAI, Google, xAI, or DeepSeek for the same prompt profile.