Anthropic API Token Costs: Claude Pricing Guide
ai-costs

Anthropic API Token Costs Explained: Prices Per Model and Per Workflow

By Paige Gilmore, Founder, NetLift· Published July 28, 2026· Updated July 28, 2026

Anthropic API costs range from $1 to $15 per 1M input tokens and $5 to $75 per 1M output tokens, depending on the model's performance tier. The most affordable option is Claude Haiku 4.5, while Claude Opus 4.1 remains the most expensive.

14-day free trial. No credit card required.

Anthropic API pricing is billed per 1 million tokens, with separate rates for input and output. Claude Haiku 4.5 is the entry-level model at $1 per 1M input tokens, while the mid-range Claude Sonnet 5 costs $2 per 1M input tokens. High-performance models like Claude Opus 5 are priced at $5 per 1M input tokens, though older versions like Opus 4.1 command a higher price of $15.

Costs are also influenced by prompt caching. This feature allows users to pay a "Write" fee to store context and a significantly lower "Read" fee for subsequent requests. For example, on Claude Sonnet 5, the standard input rate of $2 drops to $0.20 per 1M tokens when reading from the cache.

Official per-token prices

Model / rate Price Unit
Anthropic Claude 3.5 Sonnet — Batch Input tokens $3 per 1M input tokens
Anthropic Claude 3.5 Sonnet — Public Extended Access Input tokens $6 per 1M input tokens
Anthropic Claude 3.5 Sonnet — Batch Output tokens $15 per 1M output tokens
Anthropic Claude 3.5 Sonnet — Public Extended Access Output tokens $30 per 1M output tokens
Anthropic Claude Opus — On-Demand Input tokens $6 per 1M input tokens
Anthropic Claude Opus — On-Demand Output tokens $30 per 1M output tokens
Anthropic Claude Sonnet 5 — Promotional Input tokens $2 per 1M input tokens
Anthropic Claude Sonnet 5 — Promotional Output tokens $10 per 1M output tokens
Claude API — Opus 4.8 - Input $5 per 1M input tokens
Claude API (Batch Legacy) — Sonnet 4.6 Input $3 per 1M input tokens
Claude API (Batch Legacy) — Opus 4.8 Input $5 per 1M input tokens
Claude API (Batch) — Haiku 4.5 Input $1 per 1M input tokens
Claude API (Batch) — Sonnet 5 Input $2 per 1M input tokens
Claude API (Batch) — Opus 5 Input $5 per 1M input tokens
Claude API (Batch) — Haiku 4.5 Output $5 per 1M output tokens
Claude API (Batch) — Fable 5 Input $10 per 1M input tokens
Claude API (Batch) — Sonnet 5 Output $10 per 1M output tokens
Claude API (Batch) — Opus 5 Output $25 per 1M output tokens
Claude API (Batch) — Fable 5 Output $50 per 1M output tokens
Claude API (Fable 5) — Prompt caching - Read $1 per 1M input tokens
Claude API (Fable 5) — Input $10 per 1M input tokens
Claude API (Fable 5) — Prompt caching - Write $12.50 per 1M input tokens
Claude API (Fable 5) — Output $50 per 1M output tokens
Claude API (Haiku 4.5) — Prompt caching - Read $0.1 per 1M input tokens
Claude API (Haiku 4.5) — Input $1 per 1M input tokens
Claude API (Haiku 4.5) — Prompt caching - Write $1.25 per 1M input tokens
Claude API (Haiku 4.5) — Output $5 per 1M output tokens
Claude API (Opus 4.1) — Input $15 per 1M input tokens
Claude API (Opus 4.1) — Output $75 per 1M output tokens
Claude API (Opus 5) — Prompt caching - Read $0.5 per 1M input tokens
Claude API (Opus 5) — Input $5 per 1M input tokens
Claude API (Opus 5) — Prompt caching - Write $6.25 per 1M input tokens
Claude API (Opus 5) — Output $25 per 1M output tokens
Claude API (Sonnet 4.6) — Input $3 per 1M input tokens
Claude API (Sonnet 4.6) — Output $15 per 1M output tokens
Claude API (Sonnet 5) — Prompt caching - Read $0.2 per 1M input tokens
Claude API (Sonnet 5) — Input $2 per 1M input tokens
Claude API (Sonnet 5) — Prompt caching - Write $2.50 per 1M input tokens
Claude API (Sonnet 5) — Output $10 per 1M output tokens
Claude Platform — Prompt Caching (Read) - Opus 5 $0.5 per 1M input tokens

14-day free trial. No credit card required.

How do prompt caching rates impact the budget?

Prompt caching introduces a two-tier pricing structure for input tokens. When you first send a large block of text to be cached, you pay a "Write" rate, such as $1.25 for Haiku 4.5 or $2.50 for Sonnet 5. Once cached, "Read" requests for that same data are discounted to $0.10 and $0.20 per 1M tokens, respectively. This makes long-context applications significantly more predictable for finance teams.

What is the price difference between model versions?

Selecting the specific model version is a primary cost driver. While Claude Opus 5 costs $5 per 1M input tokens, the legacy Opus 4.1 costs three times as much at $15. Similarly, output costs for Opus 5 are $25 per 1M tokens, compared to $75 for Opus 4.1. Organizations should verify which version of Sonnet or Opus they are calling to avoid unnecessary overhead.

Measuring the value of Anthropic's API requires looking past the invoice to see how token spend converts into operational efficiency. A more expensive model like Opus 5 might save more employee time on complex reasoning than a cheaper model, resulting in a higher net return. Using NetLift, finance teams can correlate API costs against a baseline of manual task completion to quantify the true payback of their AI adoption.

14-day free trial. No credit card required.

Sources

All pricing on this page comes from official vendor pages:

Frequently Asked Questions

Ready to see your AI return?

14-day free trial. No credit card required.

About the author

Paige Gilmore · Founder, NetLift

Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption.

Paige Gilmore on LinkedIn

Keep reading