// the unit changed
Claude uses a newer tokenizer from Opus 4.7 on, and it cuts the same text into more pieces. The token counting page puts it plainly:
Sonnet 5 is on the new one; Sonnet 4.6 is on the old one ("Claude Sonnet 4.6 and earlier models use the previous tokenizer"). So a Sonnet 4.6 to Sonnet 5 move crosses the boundary. The Opus 5.5 migration guide gives the range: the tokenizer "may use roughly 1x to 1.35x as many tokens when processing text compared to models before Claude Opus 4.7 (up to ~35% more, varying by content)". About 30% is typical, not a constant.
Think of it as a pizza. The price per slice fell, but the new cutter slices the same pizza into more pieces.
A third off the list, about an eighth off your text
Anthropic's own Sonnet 5 page says the cut does not carry straight through:
The $2/$10 price is permanent: it "is now the standard price. The previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur." Our arithmetic for the same text, at list price:
The multiplier applies to text you send and text you get back alike, so the ratio is the same on both sides: roughly an eighth off, not a third. Your own ratio decides the exact number.
Still cheaper, per solved task
None of this means the new model costs you more. Anthropic's cost guide makes the opposite point, and names the unit to compare on:
For this exact move it measured: "Sonnet 5's saving comes from its lower per-token price, which more than offsets the extra tokens it uses per task compared with Sonnet 4.6: 15% less per solved task for 5 more points." That is one benchmark, at shipped defaults and list rates, but it is Anthropic's own and it points the right way.
And the general pattern, in its words: "in Anthropic's measurements each newer model solved at least as many tasks as the one before it, usually for less per solved task". Usually is doing work there. The same page shows Opus 4.8 to Opus 5 solving "12 more points of tasks at 21% more per solved task", and a Fable 5 to 5.1 upgrade costing 41% more per task on DeepResearch Bench II at high effort. Measure your own workload.
The check costs nothing
Three lines, from the token counting docs:
input_tokens values."Token counting is free to use but subject to requests per minute rate limits" based on your usage tier. Swap in the model IDs you actually call.
Where the counter stops. It counts the prompt, not the reply. On Sonnet 5 adaptive thinking is on by default, and max_tokens "is a hard limit on total output (thinking plus response text)", so the output side of the bill is not in that number. The count "is an estimate" and "might differ by a small amount". And it refuses a few inputs the Messages API accepts: server tools such as web search, web fetch and code execution (the advisor tool is the exception), the MCP connector, and image or document blocks with a url or file source (send those as base64 to count them).
When the cut is the cut
The gap only opens when a move crosses the Opus 4.7 boundary. Opus 5 to Opus 5.5 does not: both are on the newer tokenizer, which "later Opus models, including Claude Opus 5.5, also use". Opus 5.5 "costs $4 USD per million input tokens and $20 USD per million output tokens, below Claude Opus 5's $5 and $25", so that 20% per-token cut is a 20% cut on the same text.
One knock-on worth knowing: a million tokens now holds fewer words. "1M tokens is roughly 555k words or 2.5M Unicode characters on the current tokenizer (introduced with Claude Opus 4.7); models before it fit about 750k words in 1M tokens." The pricing page's rule of thumb of 4 characters or 0.75 words per token describes the older tokenizer.
Related: output token costs covers why the reply side of the bill is the bigger one; agent run cost covers how a loop multiplies every request; context window limits covers what fits in one call.
Sources: Claude Platform docs: Token counting; Pricing; What's new in Claude Sonnet 5; Claude Sonnet 5 overview; Optimizing for cost and intelligence; Migrating to Claude Opus 5.5; What's new in Claude Opus 5.5. Verified 2026-09-26. The 13% and 10% figures are our arithmetic at list prices; Anthropic's is the 15% per solved task.
One concept a week. Free.
The deeper, copy-paste version of each ToolCall short — in your inbox.
// total: 0.00 · spam: void · unsubscribe: one click
