Anthropic API Token Costs: Claude Pricing Guide
AI Costs & Pricing

Anthropic API Token Costs Explained: Prices Per Model and Per Workflow

By Paige Gilmore, Founder, NetLift· Published July 28, 2026· Updated August 18, 2026
Prices checked August 18, 2026 against official vendor pages

Anthropic API costs range from $1 to $15 per 1M input tokens and $5 to $75 per 1M output tokens, depending on the model's performance tier. The most affordable option is Claude Haiku 4.5, while Claude Opus 4.1 remains the most expensive.

14-day free trial. No credit card required.

Anthropic API pricing is billed per 1 million tokens, with separate rates for input and output. Claude Haiku 4.5 is the entry-level model at $1 per 1M input tokens, while the mid-range Claude Sonnet 5 costs $2 per 1M input tokens. High-performance models like Claude Opus 5 are priced at $5 per 1M input tokens, though older versions like Opus 4.1 command a higher price of $15.

Costs are also influenced by prompt caching. This feature allows users to pay a "Write" fee to store context and a significantly lower "Read" fee for subsequent requests. For example, on Claude Sonnet 5, the standard input rate of $2 drops to $0.20 per 1M tokens when reading from the cache.

Official per-token prices

Model / rate Price Unit
Anthropic Claude 3.5 Sonnet (Public Extended Access) — Batch Input $3 per 1M input tokens
Anthropic Claude 3.5 Sonnet (Public Extended Access) — On-Demand Input $6 per 1M input tokens
Anthropic Claude 3.5 Sonnet (Public Extended Access) — Batch Output $15 per 1M output tokens
Anthropic Claude 3.5 Sonnet (Public Extended Access) — On-Demand Output $30 per 1M output tokens
Anthropic Claude 3.5 Sonnet v2 (Public Extended Access) — Cache Read $0.6 per 1M input tokens
Anthropic Claude 3.5 Sonnet v2 (Public Extended Access) — Cache Write $7.50 per 1M input tokens
Claude API — Haiku 4.5 (Batch Input) $1 per 1M input tokens
Claude API — Fable 5 (Prompt caching Read) $1 per 1M input tokens
Claude API — Haiku 4.5 - Input $1 per 1M input tokens
Claude API — Sonnet 5 (Batch Input) $2 per 1M input tokens
Claude API — Sonnet 5 - Input $2 per 1M input tokens
Claude API — Sonnet 4.6 (Input) $3 per 1M input tokens
Claude API — Opus 5 (Batch Input) $5 per 1M input tokens
Claude API — Haiku 4.5 (Batch Output) $5 per 1M output tokens
Claude API — Opus 4.8 (Input) $5 per 1M input tokens
Claude API — Opus 5 - Input $5 per 1M input tokens
Claude API — Haiku 4.5 - Output $5 per 1M output tokens
Claude API — Fable 5 (Batch Input) $10 per 1M input tokens
Claude API — Sonnet 5 (Batch Output) $10 per 1M output tokens
Claude API — Fable 5 - Input $10 per 1M input tokens
Claude API — Sonnet 5 - Output $10 per 1M output tokens
Claude API — Fable 5 (Prompt caching Write) $12.50 per 1M input tokens
Claude API — Sonnet 4.6 (Output) $15 per 1M output tokens
Claude API — Opus 4.1 (Input) $15 per 1M input tokens
Claude API — Opus 5 (Batch Output) $25 per 1M output tokens
Claude API — Opus 4.8 (Output) $25 per 1M output tokens
Claude API — Opus 5 - Output $25 per 1M output tokens
Claude API — Fable 5 (Batch Output) $50 per 1M output tokens
Claude API — Fable 5 - Output $50 per 1M output tokens
Claude API — Opus 4.1 (Output) $75 per 1M output tokens

14-day free trial. No credit card required.

How do prompt caching rates impact the budget?

Prompt caching introduces a two-tier pricing structure for input tokens. When you first send a large block of text to be cached, you pay a "Write" rate, such as $1.25 for Haiku 4.5 or $2.50 for Sonnet 5. Once cached, "Read" requests for that same data are discounted to $0.10 and $0.20 per 1M tokens, respectively. This makes long-context applications significantly more predictable for finance teams.

What is the price difference between model versions?

Selecting the specific model version is a primary cost driver. While Claude Opus 5 costs $5 per 1M input tokens, the legacy Opus 4.1 costs three times as much at $15. Similarly, output costs for Opus 5 are $25 per 1M tokens, compared to $75 for Opus 4.1. Organizations should verify which version of Sonnet or Opus they are calling to avoid unnecessary overhead.

Measuring the value of Anthropic's API requires looking past the invoice to see how token spend converts into operational efficiency. A more expensive model like Opus 5 might save more employee time on complex reasoning than a cheaper model, resulting in a higher net return. Using NetLift, finance teams can correlate API costs against a baseline of manual task completion to quantify the true payback of their AI adoption.

14-day free trial. No credit card required.

Sources

All pricing on this page comes from official vendor pages:

Frequently Asked Questions

Ready to see your AI return?

14-day free trial. No credit card required.

About the author

Paige Gilmore · Founder, NetLift

Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption.

Paige Gilmore on LinkedIn

Keep reading