Last checked: September 30, 2026

Claude API Pricing

Every price below links to the page it came from.

Anthropic's API is billed per token. Two models are on sale today: Claude Sonnet 5.5 at $2 per million input tokens, and Claude Opus 5.5 at $4. Claude Haiku 5.5 has been announced but isn't priced yet.

Claude API pricing: the rate card

Model Input Output Cache read Cache write Context Max output
Claude Sonnet 5.5 $2.00 $10.00 $0.20 $2.50 1M tokens 128K tokens
Claude Opus 5.5 $4.00 $20.00 $0.20 $5.00 1M tokens 128K tokens
Claude Haiku 5.5 — — — — — —

All figures are US dollars per million tokens on Anthropic's direct API. Prices verified September 30, 2026.

Claude Haiku 5.5 is not priced yet. Anthropic's launch pages say it "will join the Claude 5.5 family in the coming weeks." We don't publish a number we can't source. This row fills in the day it ships.

What your Claude API bill will actually cost

Per-token prices are hard to feel. Fill in three numbers and this page will tell you what a request costs and what a month costs.

Claude Sonnet 5.5 Lowest cost
Per request —
Per month —
Claude Opus 5.5 Lowest cost
Per request —
Per month —

At 100,000 requests a month, the support-bot preset runs about $900/month on Sonnet 5.5 and about $1,800/month on Opus 5.5. The cheaper model is highlighted.

These are list-price estimates. They exclude caching, batch discounts and any negotiated rate.

What Claude API pricing buys you

Sonnet 5.5 is the everyday model. Anthropic positions it for well-scoped tasks: fixing bugs, producing documents, slides and spreadsheets. It scores 70.6% on Terminal-Bench 4.0, against 10.3% for Sonnet 5, and generates output 30%+ faster. If your workload is high-volume and well-defined, this is the one you want.

Opus 5.5 is the judgment model. It's for complex, open-ended work where a wrong call is expensive: codebase-wide migrations, audits, long agent runs. Anthropic says it costs 40% less to run than Opus 5 on typical workloads, mostly because it uses fewer tokens per task rather than because the rate card is lower.

Which one to pick is mostly a question of shape, not quality. If your requests are short and well defined, like extracting fields, classifying tickets or drafting replies, Sonnet 5.5 is the sensible default, and the 2× gap in input and output rates is real money at volume. Opus 5.5 earns its rate when a wrong answer costs more than the tokens: codebase-wide migrations, audits, and long agent runs where one bad decision cascades.

There is one place the two models price identically, and it matters more than people expect. A cache read costs $0.20 on both, so a workload built around one large repeated prefix gets the same discount on either model.

Context window
1M tokens
Both models
Max output
128K tokens
Both models
Knowledge cutoff
June 2026
Both models
Deprecation
Sep 2027+
Sonnet: Sep 28 / Opus: Sep 22

Both models share the same envelope: 1M-token context window. Roughly 750,000 English words, about ten novels, or a large codebase. 128K max output per request. Knowledge cutoff: June 2026.

Thinking is on by default. Sonnet 5.5 uses adaptive thinking with configurable effort (low, medium, high, xhigh, max); the default is high on the Claude Platform and medium in the Claude apps. Opus 5.5's thinking is always on and cannot be disabled; its default effort is medium. Both stay available through at least September 2027. Anthropic commits to no earlier than September 28, 2027 for Sonnet 5.5 and September 22, 2027 for Opus 5.5, with a six-month legacy window after that.

Opus 5.5 Fast mode: runs up to 2.5× speed at $8 per million input tokens and $40 per million output tokens, available in Claude Code and the Claude Platform.

How prompt caching changes your Claude API bill

Caching is where most teams find their savings, and where the arithmetic surprises people.

+25%
Write premium on both models
1/10
Sonnet read cost vs input price
1/20
Opus read cost vs input price

Writing to the cache costs more than normal input. On both models the premium is 25%: $2.50 vs $2.00 on Sonnet 5.5, $5.00 vs $4.00 on Opus 5.5.

Reading from the cache costs a fraction of normal input. On Sonnet 5.5 a cache read is $0.20, one tenth of the $2.00 input price. On Opus 5.5 it is $0.20 against a $4.00 input price, one twentieth.

That asymmetry is the whole point: you pay a 25% premium once to write a prefix, and one tenth to one twentieth of the input price every time you reuse it. Any prefix you send twice has already paid for itself.

Anthropic notes that cache reads make up the majority of agentic and coding work costs, which is why the Opus 5.5 cache-read price was cut 60% from Opus 5.

Caching is not free money, though. It pays when a prefix repeats. If every request carries a fresh document, you pay the 25% write premium and never read it back, which costs more than sending that text as ordinary input. The break-even is easy to hold in your head: a prefix has to be reused at least twice before caching beats paying the input price twice, and a cache entry has a lifetime, so a prefix that shows up once an hour may expire between uses. The workload that wins is the boring one, where the same system prompt, the same instructions and the same reference material go out on every request.

How we check Claude API prices

Every number on this page points to the page we read it from, and every row carries the date we last checked it.

Prices are read directly from Anthropic's model announcement pages. We open the page, read the pricing table, and record the date. Specifications (context window, max output, knowledge cutoff, end-of-life commitment, thinking defaults) are read from the official AWS Bedrock model cards for each model.

No aggregator data. If a figure isn't on a page we opened ourselves, it doesn't go in the table. That's why the Haiku 5.5 row is empty. Nothing is scraped or auto-imported. A stale row is worse than a blank one.

Prices are per million tokens and change when Anthropic ships or re-prices a model, not on a schedule. We re-check when a model is announced, and run a full audit of this table every month. The "last checked" date above is the real date of the last audit.

How this page makes money

Right now: it doesn't. There are no ads and no affiliate links on this page.

If that changes, the rules are simple. Any commercial link will be labelled, and it will only ever point to something that is genuinely cheaper than paying list price for the same tokens. We won't accept payment to move a number on this page.

This page can't chase scale; the sites that do are bigger and update faster. The only thing it can trade on is being right and showing its work.

Frequently asked questions

No. The API bills per token, with separate input and output prices. The free tier people are thinking of is claude.ai, the chat product, a different thing with a different price (nothing).
$2.00 for input, $10.00 for output, $0.20 for cache reads and $2.50 for a five-minute cache write.
A cache write does extra work (it stores the prefix so it can be reused), so it costs a premium over normal input: 25% on both models. A cache read just fetches that stored prefix, so it's priced far below input: one tenth on Sonnet 5.5, one twentieth on Opus 5.5.
We haven't verified reseller pricing and we won't guess. The figures on this page are Anthropic's direct API list prices. Note that on Bedrock the same models are billed through AWS Marketplace rather than on the Bedrock pricing page itself.
Three numbers and one formula. Take your requests per month, multiply by the input tokens per request and by the input price, then do the same for output tokens and add the two.
A worked example: 50,000 requests a month at 3,000 input and 400 output tokens each comes to 150 million input tokens and 20 million output tokens. On Sonnet 5.5 that is $300 of input plus $200 of output, so roughly $500 a month. The same traffic on Opus 5.5 lands near $1,000.
If 2,000 of those input tokens are the same prefix on every call, caching changes the input line: 100 million tokens come back from the cache at $0.20 per million instead of $2.00, which is about $120 instead of $300, plus the one-off cost of writing that prefix. The calculator above does all of this for you.
Anthropic says it will join the Claude 5.5 family "in the coming weeks." It has not published input or output prices. This page will carry the numbers the day they're public.
Whenever a model is announced or re-priced, plus a full audit every month. Each row shows the date it was last checked, and we won't post a date we haven't earned.