Claude API Pricing (2026)

This page lists Anthropic's current Claude API prices, explains the two discount mechanisms that change your effective rate, and covers a tokenizer detail that makes Claude bills differ across model generations even at the same list price.

Claude model prices

List prices verified 2026-08-10. USD per 1 million tokens, standard tier.

ModelInput $/1MOutput $/1MContextMax output
Claude Fable 5$10.00$50.001M128K
Claude Opus 5$5.00$25.001M128K
Claude Opus 4.8$5.00$25.001M128K
Claude Sonnet 5$3.00$15.001M128K
Claude Sonnet 4.6$3.00$15.001M128K
Claude Haiku 4.5$1.00$5.00200K64K
Claude Sonnet 5 has introductory pricing of $2.00 input / $10.00 output through 2026-08-31. The standard $3.00 / $15.00 rate applies after that date.

Prompt cache pricing

Prompt caching changes the input rate on repeated content. The mechanics:

The write premium is small next to the read discount. Content that gets written once and read many times, a system prompt, tool definitions, a document under discussion, costs 90% less on every read after the first request. For agent and coding workloads where the same prefix repeats on every step, cached reads become most of the input bill.

Batch API: 50% off

The Batch API applies a 50% discount to both input and output. It suits any job that does not need an immediate response, such as bulk classification, evaluation runs, and offline content generation.

The tokenizer difference between generations

Claude has no public tokenizer. Exact counts come from the free count_tokens API. That matters here because the tokenizer differs by model generation. Opus 4.7 and later, including Opus 5 and Fable 5, use a newer tokenizer that produces roughly 1x to 1.35x the tokens of the 4.6-and-older family on the same text. Sonnet 5 counts about 30% more tokens than Sonnet 4.6 for identical text.

So the same prompt can bill differently across Claude generations for two separate reasons: the price per token, and the number of tokens the prompt becomes. Sonnet 5 and Sonnet 4.6 share a list price, but the same text costs about 30% more on Sonnet 5 because it counts as more tokens. When you compare Claude models, count the prompt with the count_tokens API for each model rather than assuming one count.

Related pages