Technical research with cited sources. Original measurements are identified in the article.

This article was researched and written with AI assistance. Editorial responsibility: Julian Dominic Altmann. How this site is made

Published: September 22, 2026 Updated: September 22, 2026

About the author

Checked September 22, 2026. Anthropic released Claude Opus 5.5 on September 22, 2026. Direct API pricing is $4 per million input tokens and $20 per million output tokens. Prompt-cache reads are $0.20 per million tokens and cache writes are $5 per million. Anthropic also reports more than 30% higher output speed than Opus 5 and roughly 40% lower cost on typical workloads. Anthropic, Sep. 22, 2026

The more useful story is the combination of capability and economics. Anthropic’s own table shows material gains on several coding and agent benchmarks while token prices fall. But Opus 5.5 launched today, so most published scores have not yet been independently reproduced in a public, identical harness. Reuters, TechCrunch, and VentureBeat independently confirm the release and pricing; the detailed benchmark numbers remain primarily first-party measurements. Reuters VentureBeat

Claude Opus 5.5: pricing and key specifications

ItemClaude Opus 5.5Status
Release dateSep. 22, 2026official
API model IDclaude-opus-5-5official
Input$4 / 1M tokens20% below Opus 5
Output$20 / 1M tokens20% below Opus 5
Cache read$0.20 / 1M tokens60% below Opus 5
Cache write$5 / 1M tokens20% below Opus 5
Fast mode$8 input / $40 outputup to 2.5× speed, per Anthropic
Terminal-Bench 4.066.4%Anthropic result, max effort
FrontierCode v1.1 Main54.4%Anthropic result
CursorBench 4.057.8%Anthropic result
GDPval-AA v2.11846 EloAnthropic result
AutomationBench40.0%early-access setup
HLE with tools67.7%Anthropic result

Pricing and score source: Anthropic, checked September 22, 2026. Benchmark rows are vendor-reported, not measurements performed by this publication.

What is Claude Opus 5.5?

Opus 5.5 is the first model in Anthropic’s Claude 5.5 family. Anthropic positions it for demanding coding, agentic workflows, and knowledge work, saying it reaches Fable 5.1-level performance on most work while costing less. The company says Sonnet 5.5 and Haiku 5.5 are planned for the following weeks. Anthropic

For developers, the launch page gives the model identifier claude-opus-5-5. Anthropic also says the model is available across its platforms and through Amazon Web Services, Google Cloud, and Microsoft Azure. Provider catalogs can lag a launch by hours or days, so an older cloud documentation snapshot should not be treated as proof that the model is unavailable.

API pricing: a 20% cut on standard input and output

Compared with Opus 5, Opus 5.5 reduces standard input pricing from $5 to $4 per million tokens and output pricing from $25 to $20. Both are 20% cuts. Cache-read pricing drops more aggressively, from $0.50 to $0.20 per million tokens, a 60% reduction. Anthropic

Billing componentOpus 5Opus 5.5Change
Input / 1M tokens$5.00$4.00−20%
Output / 1M tokens$25.00$20.00−20%
Cache write / 1M tokens$6.25$5.00−20%
Cache read / 1M tokens$0.50$0.20−60%

Original cost example: same token volumes

Consider a cache-heavy agent workload that consumes 10M regular input tokens, 2M output tokens, 100M cache-read tokens, and 10M cache-write tokens during a billing period.

Opus 5.5: 10×$4 + 2×$20 + 100×$0.20 + 10×$5 = $150.

Opus 5: 10×$5 + 2×$25 + 100×$0.50 + 10×$6.25 = $212.50.

With identical token volumes, Opus 5.5 is 29.4% cheaper in this example. That is deliberately different from Anthropic’s “40% less on typical workloads” claim. Anthropic attributes part of the larger typical-workload saving to Opus 5.5 completing work with fewer tokens; we have not independently reproduced that efficiency claim.

Claude Opus 5.5 API price changes versus Opus 5.
Opus 5.5's largest direct cut is cache-read pricing; checked Sep. 22, 2026.

Fast mode: 2× token pricing for up to 2.5× speed

Anthropic lists Opus 5.5 Fast mode at $8/M input and $40/M output. That doubles the standard token price. The advertised upside is up to 2.5× speed. The economics therefore depend on whether latency has business value: interactive coding agents may justify the premium, while asynchronous batch work often benefits more from the standard tier. Anthropic

Benchmarks: strong vendor-reported gains, with a launch-day caveat

Anthropic reports 66.4% on Terminal-Bench 4.0, compared with 52.3% for Opus 5 in its table. It also reports 54.4% on FrontierCode v1.1 Main, 57.8% on CursorBench 4.0, and an 1846 Elo result on GDPval-AA v2.1. Anthropic

Those numbers are useful but they are not a universal model ranking. Terminal and agent benchmarks can change materially with harness, tool access, effort level, sampling, retry policy, and scoring rules. Snorkel’s Terminal-Bench material and independent trackers highlight why version and harness consistency matters. Snorkel AI LLMCompare

The timing matters too. On launch day, we did not find a broad independent same-harness reproduction covering all of Anthropic’s Opus 5.5 scores. VentureBeat similarly notes gaps in clean public same-harness comparisons with current OpenAI models. That does not invalidate Anthropic’s results; it defines the evidence class: vendor benchmark data awaiting broader independent replication. VentureBeat

Selected Anthropic-reported Opus 5.5 scores for Terminal-Bench, FrontierCode, CursorBench, AutomationBench, and HLE.
Vendor benchmark results are not independent reproductions.

Why coding agents are the most interesting test case

Anthropic’s strongest claims concern long-running tool use, repository work, and coding agents. The company describes an early test on a roughly 200,000-line codebase that finished in under three hours with Opus 5.5 versus more than 20 hours with Opus 5. In an internal HAProxy C-to-Rust project, Anthropic reports 9.5 hours with Opus 5.5 versus 12 hours with Fable 5.1 and 51% fewer tokens. Anthropic

These are informative product examples, not neutral lab tests. Teams evaluating a migration should run their own paired pilot: identical issues, repositories, tools, permissions, harness settings, and several repetitions. Record completion rate, wall-clock time, tokens, human interventions, and regression rate. That produces a metric public leaderboards cannot give you: cost per successful real task.

Safety changes: stronger safeguards and explicit fallbacks

Opus 5.5 adopts Fable 5.1-class safeguards for certain cyber and biology risks. Anthropic says some sensitive cybersecurity requests can be transparently routed to Opus 4.8, while normal software-engineering tasks such as finding and fixing bugs remain allowed. In an approximately 2,000-scenario behavioral audit, Anthropic reports 85% fewer containment-circumvention attempts than Opus 5 or Mythos 5.1. The company also flags evaluation awareness as an ongoing limitation when interpreting behavioral tests. Anthropic The Verge

Other launch details include an available zero-data-retention option, EU AI Act-related watermarking, and a change that means thinking can no longer be fully switched off. Organizations handling sensitive data should still verify the exact account, region, cloud-provider, and contractual configuration rather than treating a product-level statement as a blanket compliance guarantee.

Five key safety and operating changes in Claude Opus 5.5.
Summary of published safety and operating changes.

Context window and launch-day documentation lag

Anthropic’s general platform documentation says Claude 4.6 and later use the full 1M-token context window at standard pricing, which logically includes Opus 5.5. But the launch-day documentation itself illustrates why fast-moving model catalogs need source hierarchy: when checked on September 22, the generic pricing table had not yet been fully updated with an Opus 5.5 row even though the model-specific launch page already published the new prices. Anthropic pricing docs

For same-day releases, prioritize the model-specific primary source, then use generic documentation and cloud-provider pages as cross-checks once they catch up.

Should Opus 5 users migrate?

For existing Opus 5 workloads, the direction of the price change is unambiguous: standard input and output are 20% cheaper, cache reads are 60% cheaper, and Anthropic reports higher coding and agent performance. That makes Opus 5.5 a particularly relevant candidate for context-heavy agents that repeatedly reuse large prompts or repositories.

A production migration should still be gated by a paired test. Measure:

  1. Success rate on your real tasks.
  2. Total tokens per successful task.
  3. P50/P95 latency and time-to-completion.
  4. Human corrections and retries.
  5. Behavior changes caused by mandatory thinking.
  6. Safety fallbacks relevant to your domain.
  7. Actual provider/region pricing and data-retention terms.
Five-step process for a fair Opus 5 to Opus 5.5 migration test.
Compare with identical tasks, tools and settings.

Opus 5.5 vs. Fable 5.1: use cost per successful task

Anthropic says Opus 5.5 performs at Fable 5.1 level on most work while carrying much lower standard token prices. A deliberately simple uncached calculation of 1M input plus 1M output tokens costs $24 on Opus 5.5. Under Anthropic’s published Fable 5.1 pricing, the same raw token volume costs $60. That makes Opus 5.5 60% cheaper in this artificial equal-token example.

Real workloads are not equal-token experiments. Models may need different token counts, retries, tool calls, or human corrections. The more decision-relevant KPI is dollars per successfully completed task, ideally paired with quality and latency thresholds.

Claude Opus 5.5: what the launch changes

Claude Opus 5.5 is a substantive Opus update rather than a naming refresh. Anthropic cuts direct token prices, especially cache-read pricing, reports higher output speed, and publishes large gains on several coding and agent benchmarks. The official API identifier is claude-opus-5-5, with availability through Anthropic and the major cloud partners. Anthropic Reuters

The launch-day limitation is equally important: most detailed performance figures are still vendor-reported. Teams choosing a production model should benchmark Opus 5.5 against Opus 5, Fable 5.1, and relevant competitors in the same harness, then rank candidates by quality-adjusted cost per successful task rather than by a single leaderboard score.

Transparency

Sources and review basis

7

These primary and reference sources form the basis of the technical assessment. Vendor claims and external benchmarks are identified as such in the article.

  1. anthropic.com claude-opus-5-5
  2. reuters.com business / anthropic-unveils-claude-opus-55-2026-09-22
  3. venturebeat.com technology / anthropic-releases-claude-opus-5-5-beating-fable-5-1-on-key-agentic-benchmarks-at-60-cheaper-api-price
  4. snorkel.ai leaderboard / terminal-bench-4-0
  5. llmcompare-drab.vercel.app benchmarks / terminal-bench-4
  6. theverge.com 998868 / anthropic-claude-opus-5-5-cybersecurity
  7. platform.claude.com about-claude / pricing