40% cheaper than Opus 5.
Every token
got cheaper
Cache writes $5 (was $6.25). Fast mode: $8 in / $40 out for up to 2.5× speed.
What would
you pay?
List prices, millions of tokens a month. Excludes fast mode and batch. Opus 5.5 also uses fewer tokens per task, so real savings run higher.
fewer tokens too
Full track = what Opus 5 used on the same job. Early-tester reports. Output also streams 30%+ faster.
Effort is
your new dial
Where it shines
Prompt less.
Tune effort.
Drop “think step by step”
It decides how much to think. Replies start sooner.
Give agents a checklist
A text-only end of turn is a report, not “done”.
Name what to avoid
“No cream background, no pill buttons” beats “don't look generic”.
Tag pasted text
Wrap it in tags so hidden instructions aren't followed.
Two quiet traps
Changing effort resets the cache
Fix: change effort per message (beta). The cache stays.
Asking for its reasoning can be refused
Fix: set display: "summarized" and read the thinking blocks.
Prompt kit
Written from the patterns in Anthropic's “Prompting Claude Opus 5.5” guide. Official wording at platform.claude.com.
Four breaking
changes
Same envelope.
Better engine.
claude-opus-5-5API, AWS, GCP, Azurehand it first?
Sources: anthropic.com/claude-opus-5-5 and the Opus 5.5 docs on platform.claude.com. Tester numbers are self-reported.