Research · released 2026-09-22

Claude Opus 5.5: cheaper, stronger, with a new API

Pricing, benchmarks, what changes in the API and what to update in each system to use the new model.

Claude Opus 5.5 banner: research and upgrade plan
Summary

Cheaper than Opus 5 and ahead of it on almost everything

With four API changes that break older code. API ID claude-opus-5-5, 1M context, 128K output, default effort medium.

Price vs Opus 5
−20%

US$ 4/20 versus 5/25 per million tokens

Cache read
−60%

US$ 0.20 versus 0.50

Cost per task
−40%

approximate; Anthropic's estimate at default settings

SWE-bench Pro
89.9

Opus 5: 79.2 (system card, max effort)

Terminal-Bench 4.0
66.4%

+14 points over Opus 5

AA Intelligence Index
58

#1 on Artificial Analysis on 2026-09-22

Pricing

The cheapest Opus to date

With the largest cache discount in the Opus line.

ModelInputOutputCache readBatch (in/out)Fast (in/out)
Fable 5.110500.255 / 25—
Opus 5.54200.202 / 108 / 40
Opus 55250.502.50 / 12.5010 / 50
Opus 4.85250.502.50 / 12.5010 / 50
Sonnet 52100.201 / 5—
Haiku 4.5150.100.50 / 2.50—

US$ per million tokens. Source: Anthropic's official pricing page [PR]. Fast mode is a research preview, available only on the Claude API.

What this means in practice
  • With cache reads at 0.05x of input, long agent sessions that re-read the same prefix (bots, Claude Code) get much cheaper. On the other hand, a cache miss weighs more in relative terms.
  • On subscriptions (Claude Code, claude.ai) per-token pricing is not visible, but Anthropic raised the 5-hour limits on paid plans.
Benchmarks

Beats Opus 5 and Fable 5.1 on almost every system card test

And loses to GPT-6 Astra on three.

Read with care
  • The official numbers are self-reported and measured at max effort (Terminal-Bench at xhigh), with production safeguards on. Several were run by partners (Cognition, Cursor, Zapier, Artificial Analysis).
  • Anthropic re-evaluates older models in every system card. Opus 5 scores 1861 on GDPval-AA v2 in its own card and 1708 on v2.1 in the 5.5 card. Do not mix tables.
  • Where Opus 5.5 loses: Terminal-Bench-Science (58.7 versus 64.6 for GPT-6 Astra), AutomationBench (40.0 versus 41.4), FrontierSWE v2 (62.3 versus 65.5) and OfficeQA (to Fable 5.1).
  • Not reported for 5.5: SWE-bench Verified, ARC-AGI, GPQA, AIME/USAMO, BrowseComp and τ-bench. The model does not appear on LMArena yet.
  • Anthropic itself says that "benchmark margins have become a less reliable guide" and that the real gap to Fable 5.1 "is smaller than the scores suggest".
Effort × cost

At high, Opus 5.5 already beats Fable 5.1 at max

For less than a quarter of the cost per task, according to Artificial Analysis.

In practice: the same app at six levels

Third-party test (2026-09-24): the same /goal in Claude Code at every level, turning ~105 GB of event recordings into an explorable 3D conference. One run per level, the author’s visual judgment, and cost estimated at API prices.

EffortTimeCost (est.)TokensChecksResult
low16m43sUS$ 3.91191k22Works, but wrong branding and static videos
medium1h13mUS$ 12.44419k23The biggest quality jump
high1h07mUS$ 16.31~559k22More detail and storytelling
xhigh1h30mUS$ 25.92~723k34Author’s pick
max2h28mUS$ 50.381.18M512× the cost of xhigh, with more bugs
ultracode1h35mUS$ 18.69?42Spawned no workflows; behaved like xhigh

The GPT-6 Astra experiment with the same design reached the same pattern: the top level was not the preferred one, and there the winner was medium. Charts, the eight patterns and the data limits are in Effort in practice.

Nei’s take

Running Opus 5.5 on medium is the best option for everyday work. Go up to high when you need more reasoning and keep xhigh for extreme cases. The same goes for GPT-6 Astra. max stays out. This is a recorded opinion, not a rule.

API changes

Four breaking changes and one silent change

Compared with Opus 5. Anyone coming from 4.x also inherits the earlier restrictions.

#What changesErrorHow to migrate
1Thinking cannot be turned off. {type:"disabled"} and budget_tokens are rejected at every effort level.400Omit thinking (or use adaptive) and control it with output_config.effort. For low latency, low.
2No forced tool use. tool_choice any/tool are rejected, including in Batches and count_tokens.400auto + strict: true + an instruction in the prompt, and check that the call happened. To extract JSON, structured outputs.
3Preserved thinking. Thinking blocks are bound to the model and the conversation. Accounts created on or after 2026-08-31 get a 400 if the history is edited.400 / droppedAppend-only history. A system message mid-conversation instead of editing the system. Or the thinking-binding-controls beta with drop_block.
4Computer use only through computer_toolset_20260801. computer_20251124 is rejected (API and Google Cloud).400Change the tool declaration and the agent loop (the action becomes the block's name, several per turn).
5Text between tool calls now arrives as progress thinking blocks, empty under the default display.nonethinking.display: "updates" (beta) or "summarized", and render those blocks.

Before and after (Python)

# Before: accepted on Opus 5, 400 on Opus 5.5
client.messages.create(model="claude-opus-5", max_tokens=16000,
    thinking={"type": "disabled"}, temperature=0.4, messages=[...])

# After: thinking always on, effort is the control
client.messages.create(model="claude-opus-5-5", max_tokens=16000,
    output_config={"effort": "low"}, messages=[...])

Also inherited (if you come from 4.x)

Safety

The first Opus with cyber, bio and reasoning-extraction classifiers (reasoning_extraction). A refusal comes back as HTTP 200 with stop_reason: "refusal". The recommendation is to enable server-side fallback (fallbacks: "default"), knowing that the fallback model runs without the 5.5 thinking blocks. The system card cites no ASL level and lists caveats: the model follows more malicious instructions pasted by the user and accepts more authorization claims that cannot be verified.

Systems

Where Claude runs today in INEMA's systems

Read-only server inventory on 2026-09-22: model, how it is called, what to do and priority.

Two findings that come before any model switch
  • There are two claude installs. ~/.local/bin/claude is 2.1.280 and /usr/bin/claude is 2.1.63, from an old global npm. The 17 yt-scheduler*.service units and two crons put /usr/bin first in PATH, so they run the old version, which defaults to claude-opus-4-6. Changing the model in config does not reach them.
  • openpcbotv2 uses Agent SDK 0.2.50, which ships an embedded CLI 2.1.50. Update the SDK first, then change the ID.
SystemTodayHow it callsWhat to doPriority
yt-pub-lives (17 schedulers)claude-opus-4-6 (2.1.63 default)claude -p, old CLIFix the units' PATH (or update /usr/bin/claude) and pass --model opus explicitly.HIGH
openpcbotv2claude-opus-5, effort lowAgent SDK 0.2.50Update the SDK, switch to claude-opus-5-5 in the agent.yaml files, keep effort explicit and add xhigh to the type.HIGH
openpcbotv3alias opus → already 5.5CLI 2.1.280Nothing on the model. Check that effort is explicit in the yaml (the 5.5 default is medium).OK
openpcbotv3 (OpenRouter tiers)haiku-4.5 / sonnet-5OpenRouter, temperature 0.4If the premium tier moves to Opus 5.5, remove temperature. Fix the opus-5 price table (it says 15/75; the right value is 5/25).MEDIUM
cerebro-vip (chat API)opus-4.8, sonnet-5, fable-5 (default glm-5.2)OpenRouter, temperature 0.2Add Opus 5.5 to the list and stop sending temperature to Claude 4.7+ models.MEDIUM
cerebro-vip and telegramtopicosindex cronsalias sonnetCLI 2.1.63Same PATH fix as the schedulers.MEDIUM
inemaccvbot (stopped)claude-opus-5, effort lowCLISwitch to claude-opus-5-5 when it is reactivated.MEDIUM
musicavideoclaude-fable-5CLIProduct decision: test Opus 5.5, which beats Fable 5.1 on most system card tests and costs 60% less.MEDIUM
inemaccbot (10 profiles)alias sonnet, effort lowCLINothing now. Sonnet 5.5 arrives within weeks and the alias upgrades on its own.LOW
Portal (site chat)haiku-4.5OpenRouterNothing now. Wait for Haiku 5.5.LOW
iccmonitclaude-haiku-4-5-20251001Python SDKKeep it. If it ever moves to Opus 5.5, read the response by type (it uses content[0].text today, which breaks with thinking).LOW
claudebot, inemabot, dsh-sandboxsonnet-4-6, opus-4-6, sonnet-4.5variousStopped legacy. Update only if it runs again.LOW

Not using Claude: inemanews, inemaeventos and the portal news feed (Groq), webmcp-readiness, inemacbot, agentehermes, dsh-orchestrator and cerebro-pro.

Upgrade plan · step by step

From what unblocks the most to fine-tuning

Five steps with real commands. Order matters: without step 1, model switches never reach the services running the old CLI.

1

Unblock the CLI

Find out which claude each service sees and make all of them use the current version.

# see both installs
which -a claude; /usr/bin/claude --version; ~/.local/bin/claude --version
# option A: update the old global npm
npm i -g @anthropic-ai/claude-code@latest
# option B: in the yt-scheduler*.service units, put ~/.local/bin first in PATH
Environment=PATH=/home/nmaldaner/.local/bin:/usr/local/bin:/usr/bin:/bin
systemctl --user daemon-reload && systemctl --user restart 'yt-scheduler*'
2

Confirm where the alias points

The opus alias follows the newest model. Prefer the alias over pinning an ID; pin only where reproducibility matters.

# which model actually answered
claude --model opus -p "ok" --output-format json | jq -c '.modelUsage|keys'
# on 2026-09-22: ["claude-opus-5-5"]
3

Change the ID where it is pinned

In openpcbotv2, update the Agent SDK first and then change the model in the agent.yaml files, keeping effort explicit.

# 1) current SDK (0.2.50 ships an embedded CLI 2.1.50)
npm i @anthropic-ai/claude-agent-sdk@latest
# 2) agents/*/agent.yaml
model: claude-opus-5-5
effort: low
4

Clean up calls through OpenRouter

Before pointing to Claude 4.7 or newer, remove temperature, top_p and top_k from the request body. The Opus 5.5 slug on OpenRouter still needs checking.

# no temperature/top_p/top_k for Claude 4.7+
{ "model": "anthropic/claude-opus-...", "messages": [...] }
5

Official checklist for API code

For any code that calls the API directly (Anthropic's migration guide).

TypeItem
BLOCKSID claude-opus-5-5 (Bedrock: anthropic.claude-opus-5-5).
BLOCKSRemove thinking disabled and budget_tokens. Size max_tokens for thinking + reply. Read blocks by type.
BLOCKSReplace tool_choice any/tool with auto + strict, or structured outputs.
BLOCKSComputer use through computer_toolset_20260801.
BLOCKSAppend-only history (preserved thinking). Handle stop_reason: "refusal" and enable fallback.
TUNESet effort explicitly and measure low/medium before going higher.
TUNEUIs that showed text between tool calls: display: "updates".
TUNERe-measure cost and latency. Review prompts written for Opus 5 (verbosity, verification).
Nothing has been changed yet
  • This page is the research and the plan. Applying it to production systems is a separate step.
Timeline

Nine releases in ten months

API dates, according to Anthropic's release notes.

2025-11-24
Opus 4.5The effort parameter enters beta.
2026-02-05
Opus 4.6Adaptive thinking; budget_tokens deprecated; prefill removed.
2026-04-16
Opus 4.7New tokenizer; sampling parameters rejected.
2026-05-28
Opus 4.81M context by default.
2026-06-09
Fable 5And Mythos 5, restricted to Project Glasswing.
2026-06-30
Sonnet 5US$ 2/10 per million tokens.
2026-07-24
Opus 5US$ 5/25 per million tokens.
2026-09-01
Fable 5.1And Mythos 5.1; cache reads at US$ 0.25.
2026-09-22
Opus 5.5US$ 4/20; the subject of this page.
Coming soon
Sonnet 5.5 · Haiku 5.5"In the coming weeks", according to the announcement.
Method and sources

Where the numbers come from

Research done on 2026-09-22, launch day: Anthropic's official documentation, third-party coverage and a read-only local inventory. Each chart uses a single source table. Self-reported numbers are marked as such.

Official (Anthropic)

  1. [A] Announcement — anthropic.com/claude-opus-5-5
  2. [MP] Model overview — platform.claude.com/…/opus-5-5/overview
  3. [W] What's new — platform.claude.com/…/whats-new-opus-5-5
  4. [MG] Migration guide — platform.claude.com/…/opus-5-5/migration-guide
  5. [MO] Models overview — platform.claude.com/…/models/overview
  6. [PR] Pricing — platform.claude.com/…/pricing
  7. [RN] Release notes — platform.claude.com/…/release-notes
  8. [SC] System Card Opus 5.5 — PDF
  9. [SC5] System Card Opus 5 — PDF

Third parties

  1. Artificial Analysis — release · leaderboard
  2. GitHub Changelog — Opus 5.5 / Copilot
  3. CodeRabbit — "More catches, different misses"
  4. OrcaRouter — Opus 5.5 vs Opus 5
  5. The Decoder — the-decoder.com
  6. MacRumors · 9to5Mac · XDA · Thurrott