Claude Opus 5.5 banner showing the 22 September 2026 release, $4 input and $20 output API pricing, $0.20 cache reads, 40% lower typical workload cost, 1M-token context, 128K max output and five effort levels
AI Tools

Claude Opus 5.5 Explained: What's New, What It Costs, and Which Plans Include It

35 min read

Anthropic released Claude Opus 5.5 on 22 September 2026, three weeks after Fable 5.1 and two months after Opus 5. It is the first model in a new Claude 5.5 family, and the pitch is unusually plain: Fable 5.1-level results on most work, at Opus prices that just went down by 20% on every token and by 60% on cache reads. Anthropic's own summary is that Opus 5.5 "costs 40% less to run than Opus 5" once you count the tokens it no longer spends.

This guide walks through what actually changed, the published benchmark numbers, the new API price list, how the safeguards work, what breaks if you are moving code from Opus 5, and which Claude plans give you the model. Every figure below comes from Anthropic's launch announcement, its platform documentation or its Help Center, and the source list at the end tells you where each one lives.

Opus 5.5 at a glance

ItemClaude Opus 5.5
Release date22 September 2026
FamilyFirst model of the Claude 5.5 family; Sonnet 5.5 and Haiku 5.5 "will follow in the coming weeks" with many of the same improvements
API model IDclaude-opus-5-5 (Amazon Bedrock: anthropic.claude-opus-5-5)
Context window / max output1,000,000 tokens / 128,000 tokens
Reliable knowledge cutoffJune 2026
Thinking and effortAdaptive thinking, always on; effort levels Low · Medium (default) · High · X-High · Max
API list price$4 per million input tokens · $20 per million output tokens · $0.20 per million cache-read tokens
SpeedOutput generated more than 30% faster than Opus 5; optional Fast mode up to 2.5× speed at $8 / $40
Claude plansPro, Max, Team and seat-based Enterprise plans named in the launch (all with higher five-hour limits); the Free plan is not mentioned
Where it runsClaude apps, Claude Code, Claude Platform (API), Amazon Bedrock, Google Cloud Vertex AI, Microsoft Foundry
SafeguardsA similar class of safeguards to Fable 5.1: flagged cyber requests go to Opus 4.8, flagged biology and frontier-LLM-development requests to Opus 5

First, where Opus 5.5 sits in the Claude line-up

Anthropic's 2026 range now has four public tiers plus one restricted one. At the top is the Mythos class: Mythos 5.1 for approved organisations, and Fable 5.1, the same weights with classifier safeguards that let anyone on a paid plan use it. Below that sits Opus, then Sonnet, then Haiku. What is new this month is that the Opus tier has caught up with Fable on most everyday work, while keeping a price that is 60% lower per token.

There is also a second reason the gap matters. Anthropic says Opus 5.5 is "comparable to Claude Mythos 5.1 in biology and cybersecurity", which is why it ships with Fable-style safeguards rather than the lighter ones Opus 5 had. More on that further down.

ModelWho can use itTypical roleAPI price (in / out, per 1M tokens)
Claude Mythos 5.1Approved organisations in Anthropic's trusted-access programmesDefensive security research, professional life-science R&D$10 / $50
Claude Fable 5.1Paid Claude plans and the APIHardest reasoning and research tasks; the top of the range$10 / $50
Claude Opus 5.5Paid Claude plans and the APIFable-level coding, agents and knowledge work at Opus prices$4 / $20
Claude Opus 5Paid Claude plans and the APIPrevious Opus flagship (24 July 2026); still served; fallback model for some flagged requests$5 / $25
Claude Sonnet 5Broadly available, including the APIFast everyday assistant and high-volume automation$2 / $10
Claude Haiku 4.5Broadly available, including the APICheapest tier for simple, high-frequency tasks$1 / $5

Prices are Anthropic's published list prices in USD. Sonnet 5.5 and Haiku 5.5 are announced but not yet released at the time of writing.

What's new in Opus 5.5

Fable-level results, Opus-level bill

The headline is that a model at Opus prices now lands where Fable 5.1 did on most benchmarks, and above it on several. Anthropic pairs that with three efficiency claims. First, output tokens are generated more than 30% faster than Opus 5 at default settings. Second, the model finishes tasks in fewer steps and with fewer tokens, so the effective cost of a job falls further than the 20% list-price cut suggests, which is where the "40% less to run" figure comes from. Third, a Fast mode, available in Claude Code and on the Claude Platform, pushes speed up to 2.5× for double the token price.

Anthropic also compares the model with OpenAI's on a cost-per-task basis, all at Opus 5.5's default Medium effort. On FrontierCode it beats GPT-6 Astra at about a fifth of the cost per task. On Terminal-Bench 4.0 it beats Opus 5 at Max effort for about a fifth of the cost and matches GPT-6 Astra at about 40% of the cost. On CursorBench it beats GPT-5.6 Sol by 11 points at about a third of the cost per task. These are Anthropic's calculations, not independent ones.

The effort dial: Medium is now the default

Like Fable 5.1, Opus 5.5 exposes five effort levels, and the default has changed. Opus 5 defaulted to High; Opus 5.5 defaults to Medium, because Anthropic found the model already matches or beats older models there. The interesting evidence is that several testers report Opus 5.5 at Low effort matching Opus 5 at High: Deloitte says Opus 5.5 at low effort caught 72% of known bugs in its code reviews, where Opus 5 at high effort caught 56%; Rogo says Opus 5.5 at its lowest effort beat Opus 5 at high effort on its BigFinance Bench with about 60% fewer output tokens; Walleye Capital says the model "largely solved" its evaluation task at the lowest setting.

Two things to know before you touch the dial. Thinking can no longer be switched off on Opus 5.5; effort is the only control, and asking the API to disable thinking returns an error. And at a given effort level the model thinks more per turn than Opus 5 did, so developers should leave a bigger output budget for thinking than before.

Diagram of Claude Opus 5.5's five effort levels: Low, Medium (default), High, X-High and Max, with published tester results showing low effort matching Opus 5 at high effort

The benchmark picture

Anthropic published one comparison table across Opus 5.5, Fable 5.1, Opus 5, GPT-6 Astra and GPT-5.6 Sol. The biggest jumps are on agentic coding and business-workflow benchmarks, which reward finishing multi-step jobs rather than answering single questions. On Terminal-Bench 4.0, Opus 5.5 scores 66.4% against 55.8% for Fable 5.1 and 52.3% for Opus 5. On AutomationBench, Zapier's business-workflow test, it reaches 40.0% against 31.4% and 26.9%. Reasoning gains are smaller: on Humanity's Last Exam with tools, the model edges Fable 5.1 by about two points.

Two benchmarks go the other way. GPT-6 Astra is ahead on Terminal-Bench-Science (64.6% to 58.7%) and marginally ahead on AutomationBench (41.4% to 40.0%). Anthropic's footnotes matter here: AutomationBench was run and reported by Zapier without fallback models, so every safeguard intervention counted as a failure, and on Terminal-Bench-Science the standard error is roughly ±3.5–5 points per model.

BenchmarkWhat it measuresOpus 5.5Fable 5.1Opus 5GPT-6 AstraGPT-5.6 Sol
Terminal-Bench 4.0Agentic coding66.4%55.8%52.3%57.9%37.3%
FrontierCode v1.1Hard software engineering54.4%50.3%48.0%53.3%47.5%
CursorBench 4.0Agentic coding in an IDE57.8%51.8%46.6%–41.7%
AutomationBenchBusiness workflows40.0%31.4%26.9%41.4%28.8%
Terminal-Bench-Science 0.1Agentic scientific research58.7%52.6%29.0%64.6%22.4%
OSWorld 2.0 (partial credit)Computer use81.8%80.7%74.0%––
Humanity's Last Exam (with tools)Multidisciplinary reasoning67.7%65.6%63.6%57.2%–
Chartography (with tools)Reading charts and figures89.0%88.4%83.4%––
GDPval-AA v2.1Knowledge work (rating, not %)18461735170815421588

Source: Anthropic, "Introducing Claude Opus 5.5", 22 September 2026. Opus 5.5 results use adaptive thinking at Max effort unless noted; Terminal-Bench 4.0 shows Opus 5.5 at X-High and GPT-6 Astra at High, each model's best score. Opus 5.5 was tested with production safeguards on; where they intervened, cyber tasks were completed by Opus 4.8 and biology tasks by Opus 5, which "likely reduces" its scores. Terminal-Bench 4.0 carries a ±2.6-point standard error for Opus 5.5.

Six bar charts comparing Claude Opus 5.5 with Fable 5.1, Opus 5, GPT-6 Astra and GPT-5.6 Sol on Terminal-Bench 4.0, FrontierCode, CursorBench, AutomationBench, Terminal-Bench-Science and Humanity's Last Exam

Fewer steps, fewer tokens

Most of the customer evidence in the launch is about efficiency rather than raw capability, and the numbers are consistent across very different companies. GitHub says its agents solved more terminal tasks in less than half the steps Opus 5 needed. Box reports that in its evaluations Opus 5.5 used a third of the tokens Opus 5 did, with answers 40% less verbose and no loss of accuracy. Kiro measured about 40% fewer tool calls and half the tokens. Optiver saw the same quality as Opus 5 "in about half the turns, time and output tokens", a 40 to 50% cost cut on agentic coding. Factory says Opus 5.5 at Medium effort matched Opus 5 at High with 20 to 25% fewer output tokens, the first model it would default to at Medium. Lovable says builds finish in a third to a half fewer steps with fewer, more complete edits.

Against Fable 5.1 the comparison Anthropic highlights is a rewrite of the HAProxy codebase that both models were given: both passed nearly all of HAProxy's regression tests, but Opus 5.5 finished in 9.5 hours instead of 12 and cost 51% less. Two unnamed early testers add the largest-scale examples: a 680,000-line code migration completed in under a day, and an audit-and-fix of a 200,000-line codebase in under three hours, where Opus 5 took over 20 hours and used 2.5× the tokens.

Built for runs that last all night

The long-horizon theme from Fable 5.1 continues. Clio left Opus 5.5 running unattended for more than 18 hours on an engineering task spanning six repositories and reports it hit milestones faster with minimal rework. Stripe used it to direct a multi-day rebase of 40 stacked pull requests, laying out every conflict plainly; all 40 passed CI. Quantium says a complex coding task that previously took 38 prompts over four days took 11 prompts in three hours and came out production-ready. Chicago Trading had it handle a bug autonomously overnight and says its writing is "easy to follow and more coherent than Opus 5's". Column used it to audit a cloud bill and noted stronger self-verification loops.

Anthropic's own alignment data backs the unattended use case from the other side: in its behavioural audit, Opus 5.5 attempted to circumvent boundaries it had been given around 85% less often than Opus 5 or Mythos 5.1, and it is described as much less likely to take hard-to-reverse actions or act outside the boundaries it was set. For an agent you leave running with real credentials, that is the number to watch.

Knowledge work, research and writing

On GDPval-AA v2.1, the third-party rating of professional tasks, Opus 5.5 scores 1846 against 1735 for Fable 5.1 and 1708 for Opus 5, and the domain customers echo it. Hebbia, whose users work on financial documents, measured 86.6% coverage on end-to-end finance workflows graded against expert rubrics, up from 60.3% with Opus 5, alongside its best-ever citation recall and a cost per research task it describes as "in check". Thomson Reuters reports better results in its expert evaluations and internal benchmarks for CoCounsel, with gains in speed and token efficiency. LexisNexis says the model identified highly relevant citations consistently, with particular strength on statutes. Viktor solved twice as many of its hard tasks at nearly half the cost, in fewer steps and tool calls, and Hex says the model "keeps digging past the first plausible answer".

Anthropic also spent a section on writing, which was a frequent complaint about Opus 5. Opus 5.5 is meant to put the most important information first, use less jargon and fewer idiosyncratic phrases, and follow the writing rules you give it. Ramp says a design spec came out usable with minimal edits because the model followed its writing rules. Early testers described the output as clearer and easier to follow, which matters most in long sessions where you read hundreds of model messages.

Sharper eyes

One quieter change: Anthropic's docs say Opus 5.5 reads values off dense charts, diagrams and screenshots more precisely without tools, and on the with-tools Chartography benchmark it scores 89.0% against 83.4% for Opus 5. In practice that means fewer wrong numbers pulled from a pasted dashboard or a scanned table.

Pricing: every number went down

Unlike Fable 5.1, where only cache reads got cheaper, Opus 5.5 cuts the whole price list. Input drops from $5 to $4 per million tokens and output from $25 to $20, a 20% cut on both. Cache reads fall from $0.50 to $0.20, a 60% cut, and cache writes from $6.25 to $5 for the five-minute cache. Batch API pricing is half of list, so $2 in and $10 out.

Prompt caching is where the saving compounds. Any agent that works for more than a few steps re-reads the same instructions, files and history every turn, and those repeated tokens are billed at the cache-read rate rather than the input rate. With Opus 5.5 that rate is 5% of the input price. Combine the cheaper list, the cheaper cache and a model that finishes in fewer steps, and you get Anthropic's estimate of a bill roughly 40% lower than Opus 5 on typical workloads. Versus Fable 5.1, the list price is 60% lower per token before any efficiency difference.

Per 1M tokens (USD)Opus 5Opus 5.5Fable 5.1Sonnet 5
Input$5.00$4.00$10.00$2.00
Output$25.00$20.00$50.00$10.00
Cache read$0.50$0.20$0.25$0.20
Cache write (5-minute)$6.25$5.00$12.50$2.50
Cache write (1-hour)$10.00$8.00$20.00$4.00
Batch API (in / out)$2.50 / $12.50$2.00 / $10.00$5.00 / $25.00$1.00 / $5.00
Fast mode (in / out)$10.00 / $50.00$8.00 / $40.00––

Anthropic list prices. Cache writes are 1.25× (5-minute) and 2× (1-hour) the input price; Fable 5.1 cache reads are a special 0.025× rate and Opus 5.5's a special 0.05× rate. US-only inference adds a 1.1× multiplier on all token types. There is no long-context surcharge: the full 1M-token window is billed at the standard rate. Fast mode is a research preview on the Claude API only and is not available through the Batch API.

A concrete example: a request with 20,000 input tokens and 5,000 output tokens costs $0.18 on Opus 5.5, $0.225 on Opus 5 and $0.45 on Fable 5.1, before caching or batching. Subscription plans (Pro, Max, Team) are billed monthly, not per token, so these figures matter for API users and for understanding why usage limits stretch further.

Infographic comparing Claude Opus 5 and Opus 5.5 API prices: input $5 to $4, output $25 to $20, cache reads $0.50 to $0.20, with Anthropic's estimate of 40% lower running cost

Safeguards: Fable-style routing comes to Opus

This is the part Opus 5 users will notice first. Because Anthropic judges Opus 5.5 comparable to Mythos 5.1 in biology and cybersecurity, it ships with "a similar class of safeguards to Fable 5.1" on cybersecurity, biology and distillation, rather than the lighter Opus 5 set. A classifier watches each request, including attached files, memory, connector content and web results, not just your last message. If it flags a higher-risk cybersecurity request, the conversation is handed to Opus 4.8; if it flags a dual-use biology or frontier-LLM-development request, it goes to Opus 5, a new route for the Opus tier, since Opus 5 itself never fell back on biology. You see a notice that the model switched and the reply is labelled with the model that answered. The Help Center's model-switching article for Opus 5 describes the app controls, which carry over: you can turn the automatic switch off under Settings → Capabilities ("Switch models when a message is flagged"); with it off, a flagged request pauses the conversation instead of switching models, and editing your previous message before retrying often clears it.

On the API, the picture is slightly different: a declined request returns a normal HTTP 200 with stop_reason: "refusal" and a stop_details object naming the policy area, and a beta fallbacks: "default" option retries on Anthropic's recommended model server-side. Two safeguard categories are new for the Opus tier: a biology safety classifier, and a reasoning_extraction category for requests that try to make the model reproduce its internal reasoning.

If your request is…Who answersWhat you see
Everyday coding, writing, analysis, research (the vast majority)Opus 5.5A normal reply
Higher-risk cybersecurity: exploit generation, binary vulnerability scanning, penetration testingOpus 4.8A notice that the model switched; reply labelled Opus 4.8
Dual-use biology or frontier LLM developmentOpus 5A notice that the model switched; reply labelled Opus 5
Either of the above, with automatic switching turned offNobody yetThe conversation pauses; edit and retry, or turn switching back on

Anthropic has not published false-positive reduction percentages for Opus 5.5 the way it did for Fable 5.1, so treat the intervention rate as "similar to Fable 5.1" until it says otherwise. What it has published is a route around the routing: the Cyber Verification Program is being expanded "in the coming weeks" so that verified security practitioners can use Opus 5.5 for offensive-security work, and the Life Sciences Verification Program is open for applications from organisations doing advanced biology.

On the alignment side, Anthropic reports that Opus 5.5 scored better than any recent Claude model on nearly every measure of misaligned behaviour in its automated behavioural audit, nearly 2,000 scenarios covering things like biased reasoning, attempts to escape a sandbox and harmful actions taken after concluding a situation was simulated. Opus 5.5 also ties Fable 5.1 for the lowest prompt-injection success rate of any model tested on Gray Swan's benchmark and matches or beats Opus 5 in every setting Anthropic tried. The company adds two honest caveats: catching every failure before deployment "remains an unsolved problem", and the model often suspects it is being evaluated, which complicates measuring real-world behaviour.

Flow diagram showing how Claude Opus 5.5 routes flagged requests: everyday work answered by Opus 5.5, higher-risk cybersecurity requests handed to Opus 4.8, dual-use biology requests handed to Opus 5

Preserved thinking, watermarks and "pacing the frontier"

Three policy items carry over from Fable 5.1. Preserved thinking, Anthropic's anti-distillation measure, now applies to Opus 5.5: API accounts created on or after 31 August 2026 cannot edit the model's prior context to extract its reasoning, and thinking blocks are bound to the conversation prefix that produced them. Anthropic's September 2026 threat-intelligence report describes the kind of attack this targets, with thousands of fake accounts used to pull a model's capabilities out at industrial scale. Watermarking of generated text, which Fable 5.1 shipped with to meet the EU AI Act's transparency rules, is on for Opus 5.5 as well. And zero data retention remains available, as on previous Opus models.

The framing Anthropic chose is worth noting. Opus 5.5 is "our first model since we called for pacing the frontier", the company's position, argued by CEO Dario Amodei, that progress should be paced so that safety practice stays ahead of capability. In practice that meant external testing by METR and Frontier Design before release, safeguards matched to the model's measured risk, and a 5.5 rather than a 6.

For developers: what breaks when you move from Opus 5

Swapping the model ID is not enough. Anthropic's migration guide lists four breaking changes and a few behaviour shifts, all of which surface as 400 errors or subtly different output if you ignore them.

ChangeWhat happensWhat to do
Thinking cannot be disabledthinking.type: "disabled" returns a 400Remove the thinking field; control depth with output_config.effort
Forced tool use removedtool_choice of any or tool returns a 400Use auto with strict tool use, or structured outputs
Thinking blocks bound to model and prefixOpus 5.5 does not read thinking from Fable or Mythos; a changed system prompt, tools or earlier messages returns a 400 by defaultUse the thinking-binding-controls-2026-08-01 beta header with prefix_mismatch_behavior: "drop_block" to drop affected blocks instead
New computer-use toolcomputer_20251124 not supported on the Claude API and Google CloudMove to computer_toolset_20260801 (Bedrock still accepts the old tool)
Default effort is MediumWas High on Opus 5Re-run your effort sweeps; set explicitly if you relied on High
Progress text lives in thinking blocksText between tool calls is omitted by default (thinking.display: "omitted")Use thinking.display: "updates" (beta header thinking-display-updates-2026-08-18) if you stream status updates

Fast mode is enabled with speed: "fast" under the fast-mode-2026-02-01 beta header and is Claude API only; Bedrock, Vertex AI and Foundry do not offer it. Prompt caching needs at least 512 tokens to cache.

Which Claude plan includes Opus 5.5?

The launch post is specific about three things and silent on the rest. It says Anthropic is increasing five-hour usage limits on Pro, Max, Team and seat-based Enterprise plans, which is the same as saying the model is available on those plans, and it introduces a rate-limit reset for subscribers that "you can now save and use whenever you choose", in other words a stored reset you trigger when a session limit bites at a bad moment. It does not give a percentage for the limit increase, does not mention the Free plan, and does not say whether Opus 5.5 replaces Opus 5 as a default model. For reference, Opus 5 launched on 24 July as the default model on Max and the strongest model on Pro; if Anthropic follows the same pattern, Opus 5.5 takes those slots, but check the model picker rather than assuming.

PlanOpus 5.5What the launch says
FreeNot statedThe launch post does not mention the Free plan
ProYesFive-hour usage limits increased; saved rate-limit reset
Max 5x and Max 20xYesFive-hour usage limits increased; saved rate-limit reset
TeamYesFive-hour usage limits increased; saved rate-limit reset
Enterprise (seat-based)YesFive-hour usage limits increased; zero data retention available
Claude CodeYesIncluded with Fast mode; update to the current release
Claude API / Bedrock / Vertex AI / FoundryYesclaude-opus-5-5 at the prices above

Which plan makes sense? The economics changed in a specific way. Fable 5.1 on a Max plan is capped at 50% of your weekly limit and metered with credits on Pro, so heavy Fable use pushes people up the ladder. Opus 5.5 is an ordinary Opus model with ordinary limits, and those limits just grew, while the model spends fewer tokens per job. If Fable 5.1 was the only reason you were looking at Max, Opus 5.5 on Pro is worth trying first for everyday coding and writing; keep Fable for the problems where Opus at High or X-High still misses something. If you run Claude Code sessions for hours or several agents at once, the Max tiers are still built for that, and Max 20x remains the option with the most session capacity.

Where Hermanos License fits in

Our AI Tools category holds the assistant subscriptions we sell, with digital delivery to your email and account dashboard and 15-day store support on every order. Availability changes, so each product page shows live stock, price and delivery details. At the time of writing the category includes:

ProductWhat you getWhere it helps with Opus 5.5
Claude Max 20x SubscriptionAnthropic's highest consumer tier, with Opus 5.5 and Fable 5.1 (Fable up to 50% of weekly limits), full Claude Code and Cowork accessLong Claude Code sessions, several agents at once, and switching between Opus 5.5 and Fable 5.1 without credits
Google AI Pro (Gemini) – 18-Month SubscriptionGemini's paid tier for 18 months, delivered as an activation linkA second assistant for research and cross-checking Claude's answers; Google Workspace integration
Google AI Pro – Gemini Advanced + 5 TB StorageGemini Advanced plus 5 TB of Google storageStorage for the documents, datasets and repos you feed into long-running agents

We add new AI subscriptions to this category as suppliers list them, so the category page is the place to check rather than this table.

Hermanos License Services banner: ready to put Claude Opus 5.5 to work, with links to Claude Max 20x, Google AI Pro and design software in the AI Tools category

Opus 5.5 does its best work inside software you already own, and the pairings from the Fable 5.1 guide still hold:

If you use Opus 5.5 for…It pairs well withFind it in
Coding and Claude Code sessionsWindows 11 Pro for a proper local dev machine: Hyper-V, Windows Sandbox, BitLocker, Remote Desktop hostingWindows 11
Decks, reports and financial analysisMicrosoft Office, so the .pptx and .xlsx files the model produces open exactly as intendedOffice Licenses
Process diagrams and project plansMicrosoft Visio and Project: Opus drafts, you refineVisio & Project
Design work and UI prototypesCorelDRAW or Figma to finish the model's mock-upsDesign Tools
Reading PDFs, contracts and scansAdobe Acrobat Pro for editing and signing what the model summarisesAdobe Acrobat Pro 2020
Agents that browse the web for youEndpoint protection and a VPN, because agents open pages you never seeAntivirus & VPN
SEO audits and content plansScreaming Frog SEO Spider for the crawl data the model analysesScreaming Frog SEO Spider

All Hermanos License licences and subscriptions are delivered digitally. Claude subscriptions remain subject to Anthropic's Consumer Terms, including session and weekly limits.

Frequently asked questions

Is Opus 5.5 better than Fable 5.1?

On Anthropic's published table it scores higher than Fable 5.1 on every benchmark listed, at 40% of the token price. Anthropic's own wording is more careful: it "performs at the level of Claude Fable 5.1 on most work". Fable 5.1 remains the top of the range and the model to reach for when Opus at High or X-High effort still misses a condition.

Why did Anthropic release 5.5 instead of Claude 6?

The company frames Opus 5.5 as its first release since it called for "pacing the frontier": improving efficiency, alignment and safeguards on a model whose risks it says it largely understands, rather than making a larger capability jump. Sonnet 5.5 and Haiku 5.5 follow in the coming weeks "with many of the same improvements to performance, efficiency, and safety".

Can I use Opus 5.5 on the Free plan?

The launch post names Pro, Max, Team and seat-based Enterprise plans and does not mention Free. Check the model picker in your account for the definitive answer.

Why did my Opus 5.5 conversation switch to Opus 4.8 or Opus 5?

A safety classifier flagged the request as higher-risk cybersecurity (handled by Opus 4.8) or dual-use biology or frontier-LLM development (handled by Opus 5). The reply is labelled with the model that answered. Opus 5 users did not see biology routing before; it is new with Opus 5.5 because the model's biology capability is judged comparable to Mythos 5.1.

Is Opus 5.5 really 40% cheaper? The list price only fell 20%.

Both are true. The per-token list price fell 20% (60% for cache reads). The "40% less to run" figure is Anthropic's estimate for typical workloads and includes the model finishing tasks in fewer steps and tokens; customer reports of a third to half the tokens are consistent with it, but your own workload will land somewhere specific.

Can I turn thinking off to save money?

No. Adaptive thinking is always on for Opus 5.5 and the API returns an error if you try to disable it. Use the effort parameter instead; Low effort is where the cost savings live, and several testers report it matching Opus 5 at High.

What is the "rate limit reset"?

A reset of your session usage limit that subscribers can now save and trigger at a time of their choosing, instead of waiting for the five-hour window to roll over. Anthropic introduced it alongside the increased five-hour limits on Pro, Max and Team.

Sources

Anthropic, "Introducing Claude Opus 5.5" (22 September 2026) · Claude Platform Docs, "What's new in Claude Opus 5.5" and "Models overview" (model IDs, context window, output limits, effort defaults, breaking changes) · Claude Platform Docs, "Pricing" (list, cache, batch, fast-mode and data-residency prices) · Claude Platform Docs, "Claude Opus 5" model page (Opus 5 prices and release date) · Claude Help Center, "Why Claude switched models in your conversation" (routing, labelling, settings toggle) · Claude Help Center, release notes (22 September 2026) · Anthropic, "Introducing Claude Opus 5" (24 July 2026). Benchmark figures and customer quotes are Anthropic's published numbers and have not been independently verified.

You might also be interested