Claude Opus 5 vs 5.5: Price, Performance and a Real Test

Jonhisking · October 9, 2026

Claude Opus 5 vs 5.5: Price, Performance and a Real Test

Claude Opus 5.5 is Anthropic's new top model, released on September 22, 2026. Compared with Opus 5, its token prices are 20% lower and cache reads are 60% cheaper. Anthropic says it costs "40% less to run on typical work"; our own math for a coding agent workload came out about 37% cheaper. When we gave both models the same game-building request, Opus 5.5 used 64% fewer tokens. Below: price, performance and measured results side by side, plus a Max 20x user's experience with both models and whether it is worth switching.

Opus 5 vs 5.5 pricing

Opus 5 vs 5.5 pricing: Item, Opus 5, Opus 5.5, Difference

Price per million tokens, from Anthropic's official pricing page.

Item Opus 5 Opus 5.5 Difference
Input $5 $4 20% cheaper
Output $25 $20 20% cheaper
Cache write (5 min) $6.25 $5 20% cheaper
Cache read $0.50 $0.20 60% cheaper
Fast mode – $8 / $40 New in Opus 5.5

Hands-on: my Max 20x weekly limit started to last

I (Jaehyun) have used both Opus 5 and Opus 5.5 on the Claude Max 20x plan while building games. This is what it felt like in daily use, not a measured figure.

The official prices above and the calculations below point the same way. Doing the same work for less cost (that is, less of your limit) means the same plan lasts longer.

We measured it with the same request

Same request, measured: the brick-breaker games built by Opus 5 and Opus 5.5 side by side

Feelings are not enough, so on October 9, 2026 we gave both models the same request once each in Claude Code, changing only the model name: build a brick-breaker game that runs in the browser (two levels, score and lives) as a single HTML file. Usage was measured with the free tool ccusage.

Item Opus 5 Opus 5.5 Difference
Output tokens 36,640 15,180 59% fewer
Cache write tokens 74,533 63,917 14% fewer
Cache read tokens 1,774,025 595,337 66% fewer
Total tokens 1,885,246 674,452 64% fewer
Cost (at API prices) $2.55 $0.93 64% cheaper
Game code produced about 684 lines 409 lines 40% shorter

Token chart for the same request: Opus 5 used 1.89M tokens, Opus 5.5 used 674K

ccusage measurement record: Opus 5 and Opus 5.5 tokens and cost on October 9, 2026

Cost on real workloads

Cost on real workloads: Task, Opus 5, Opus 5.5, Difference

We calculated the cost of the same work on both models, keeping token counts equal and changing only the prices.

Task Opus 5 Opus 5.5 Difference
1 chat request (2,000 in / 500 out) $0.0225 $0.018 20% cheaper
1 document question (50,000 in, 45,000 cached / 1,000 out) $0.0725 $0.049 32% cheaper
1 coding agent feature (1.6M in, mostly cached / 17,500 out) $1.79 $1.14 37% cheaper
Coding agent for a month (110 features) about $197 about $125 37% cheaper

Example: Opus 5.5 document question = 5,000 new input × $4 + 45,000 cached × $0.20 + 1,000 output × $20 (all per million tokens) = $0.049.

The more you cache, the more you save. Plain chat is only 20% cheaper, but a coding agent that keeps re-reading the conversation gets close to 37% cheaper. This table assumes equal token counts. Anthropic says its "40% less" comes from lower prices plus fewer tokens per task at default settings, and our test above also used 64% fewer tokens.

Performance (as reported by Anthropic)

Performance (as reported by Anthropic): benchmark bar chart, Opus 5 vs Opus 5.5

Benchmarks published by Anthropic. Opus 5.5 scores use the highest setting (max effort), except Terminal-Bench 4.0, which uses one step lower (xhigh). All are Anthropic's own measurements.

Benchmark Opus 5 Opus 5.5
Terminal-Bench 4.0 (terminal tasks) 52.3% 66.4%
CursorBench 4.0 (coding) 46.6% 57.8%
FrontierCode v1.1 (coding) 48.0% 54.4%
AutomationBench (work automation) 26.9% 40.0%
OSWorld 2.1 (computer use, subset) 74.0% 81.8%
Humanity's Last Exam (with tools) 63.6% 67.7%

Independent results: Artificial Analysis

There are outside scores too. On the Intelligence Index from independent evaluator Artificial Analysis, Opus 5.5 at max effort took first place with 58 (per press reports).

Model Intelligence Index
Claude Opus 5.5 (max effort) 58
GPT-6 Astra 53
Claude Fable 5.1 53
Claude Opus 5 51

Effort settings change the savings

For Pro and Max plan users

Pros and cons

Opus 5.5 pros and cons summary card

Pros

Cons

What to change when moving your API code

Per Anthropic's migration guide, some Opus 5 code returns a 400 error on Opus 5.5.

Change What to do on Opus 5.5
Thinking can't be turned off: disabled and budget_tokens error Drop the thinking setting and control cost with effort
No forced tool use: tool_choice any and tool error Use auto with strict tools, or structured outputs
Old computer-use tool (computer_20251124) errors on the API and Google Cloud Switch to computer_toolset_20260801
Replaying thinking blocks after editing earlier turns errors (accounts created on or after August 31, 2026) Keep conversation history append-only
Notes between tool calls move into thinking blocks Set thinking.display if you show them to users
Wider refusals (stop_reason: "refusal") Handle refusals and set up a fallback

Priority Tier is also not supported on Opus 5.5. If you use it through Claude Code or a subscription plan, there is no code for you to change.

Should you switch from Opus 5?

FAQ

How much cheaper is Opus 5.5 than Opus 5? Token prices are 20% lower and cache reads 60% lower. On real workloads, chat is 20% cheaper and cache-heavy coding agent work about 37% cheaper.

Do token counts change? The same input text gives the same token count (same tokenizer). The tokens spent doing the work differ: Opus 5.5 used 64% fewer in our test, and it can use more at the highest setting.

Opus 5.5 or Sonnet 5.5? Sonnet 5.5 is enough for most work at half the price per token. Use Opus 5.5 for hard coding, long agent runs and writing where quality matters.

Can I reuse my Opus 5 API code as is? No. Settings such as turning thinking off or forcing a tool with tool_choice return errors on Opus 5.5. See "What to change when moving your API code" above.

Calculate it for your own work

In the coding agent cost calculator, enter task size and tasks per day to compare a month on Opus 5.5, Sonnet 5.5 and GPT-6 Sol. Check the cost of a single prompt in the token counter. To decide between a plan and the API, see Claude Code cost per month.

Prices as of October 9, 2026. Prices change; check Anthropic's pricing page before you rely on them.

Sources

More articles

Enter task size and tasks per day to compare a month on Opus 5.5, Sonnet 5.5 and GPT-6 Sol.

Open the coding agent cost calculator →