Claude Opus 5.5 Pricing: 40% Lower Typical Cost
Claude Opus 5.5 costs $4 input, $0.20 cache reads, and $20 output per million tokens. See what changed and how it compares with Opus 5 and GPT-6.
By AI Pricing Guru Editorial Team
AI Pricing Guru articles are maintained by the editorial workflow behind the site: daily pricing snapshots, provider source checks, and review passes for model launches, subscription limits, and billing changes.
TL;DR
- Claude Opus 5.5 is live at $4 input, $0.20 cache reads, $5 five-minute cache writes, and $20 output per 1M tokens.
- Input and output are 20% cheaper than Opus 5; cache reads are 60% cheaper. Anthropic estimates 40% lower cost on a typical default-settings workload.
- Anthropic recommends Opus 5.5 for most workloads, with Fable 5.1 reserved for demanding work that still benefits from the frontier tier.
- Benchmark wins are configuration-specific. Replay your own tasks at fixed effort and compare accepted-result cost before migrating.
Current premium model token rates
USD per 1M tokens. Input and output rates are charted separately.
Estimate an Opus 5.5 migration
Assumes 75% input tokens and 25% output tokens using current per-million rates.
GPT-6 Sol
openai
$40.00
- Input share
- $15.00
- Output share
- $25.00
Claude Opus 5.5
anthropic
$80.00
- Input share
- $30.00
- Output share
- $50.00
Claude Opus 5
anthropic
$100.00
- Input share
- $37.50
- Output share
- $62.50
GPT-6 Astra
openai
$200.00
- Input share
- $75.00
- Output share
- $125.00
Claude Opus 5.5 versus Opus 5, Fable 5.1, and GPT-6
| Model | Provider | Input / 1M | Cached / 1M | Output / 1M |
|---|---|---|---|---|
| Claude Opus 5.5 | anthropic | $4.00 | $0.2 | $20.00 |
| Claude Opus 5 | anthropic | $5.00 | $0.5 | $25.00 |
| Claude Fable 5.1 | anthropic | $10.00 | $0.25 | $50.00 |
| Claude Sonnet 5 | anthropic | $2.00 | $0.2 | $10.00 |
| GPT-6 Sol | openai | $2.00 | $0.2 | $10.00 |
| GPT-6 Astra | openai | $10.00 | $1.00 | $50.00 |
Built from pricing.json at publish time.
Anthropic launched Claude Opus 5.5 on September 22, 2026. The model is live as claude-opus-5-5 with a 1-million-token context window and up to 128,000 synchronous output tokens.
The practical headline is not a benchmark trophy. It is the bill: standard input and output rates are 20% below Opus 5, while cache reads are 60% cheaper. Anthropic says a typical workload at default settings costs 40% less because Opus 5.5 also uses fewer tokens per task.
Claude Opus 5.5 pricing
Standard pricing per 1 million tokens is $4 input, $0.20 cache reads, $5 five-minute cache writes, $8 one-hour cache writes, and $20 output. Batch input/output is $2/$10. Fast mode, which Anthropic says can run up to 2.5 times faster, costs $8 input and $40 output.
Compared with Opus 5, a workload using 1M uncached input tokens and 1M output tokens falls from $30 to $24 before tool fees—a 20% reduction. A cache-heavy workload benefits more because reads fall from $0.50 to $0.20 per 1M.
Anthropic’s 40% typical-workload saving is a provider estimate, not a universal discount. Your result depends on output length, cache hits, thinking effort, retries, and whether the stronger model finishes with fewer attempts.
Use the live Claude pricing page for the full rate table and the token calculator for your prompt and output mix.
What changed besides price
Anthropic positions Opus 5.5 as the default choice for most coding and knowledge work. Fable 5.1 remains the escalation route when demanding reasoning or long-horizon work still performs better in a team’s own evaluation.
The launch reports output more than 30% faster than Opus 5 and highlights clearer, less jargon-heavy writing. Opus 5.5 uses adaptive thinking with medium as the API default effort. The model is available through Anthropic’s API, Amazon Bedrock, Google Cloud, Microsoft Foundry, and Claude Platform on AWS.
Subscription users also get higher five-hour limits on Pro, Max, Team, and seat-based Enterprise plans, plus a rate-limit reset they can save. Anthropic did not announce new seat prices, so this is an allowance change—not a subscription price cut.
Do the benchmark claims hold?
Anthropic reports 66.4% on Terminal-Bench 4.0, 54.4% on FrontierCode, 57.8% on CursorBench, and 1846 Elo on GDPval-AA. Its launch page says default-effort Opus 5.5 beats GPT-6 Astra’s max-effort result on selected FrontierCode and GDPval comparisons at about one-fifth of the provider-calculated cost per task.
That qualification matters. The published top-line benchmark table generally uses max or xhigh settings, while the cost-efficiency examples compare particular effort levels, harnesses, and tasks. “One-fifth the cost” is not a blanket statement that every Opus 5.5 API workload costs 80% less than Astra.
The safer interpretation is that Opus 5.5 deserves an immediate replay against both Opus 5 and the new GPT-6 Sol and Luna tiers.
For the buying decision, see Claude Opus 5.5 vs GPT-6 Sol: cost per task. Opus 5.5 costs twice as much as Sol for equal fresh-input and output volumes, so it must save steps, retries, or review time to win the bill.
Independent Labs result
Both exact OpenRouter routes scored 49/49 with zero errors on AI Pricing Guru’s 49 deterministic tasks. Opus 5.5 cost $0.05422 ($0.0011065 per correct answer), while GPT-6 Sol cost $0.015922 ($0.00032494 per correct answer). Opus was 3.4× more expensive on this narrow suite.
Against Opus 5, the new model improved from 48/49 at $0.09366 to 49/49 at $0.05422—43% lower cost per correct answer. These short machine-graded tasks establish route availability and a repeatable baseline; they do not reproduce repository-scale coding or Anthropic’s agent benchmarks. See the Labs leaderboard and coverage note.
OpenClaw and Hermes compatibility
The OpenClaw build checked at launch did not yet list Opus 5.5. Hermes’ experimental DirectSDK plugin does list the subscription-backed route, but reports claude -p metering around 1.7 times its interactive Claude Code TUI test. Direct integrations must also handle always-on thinking, unsupported forced tool choice, bound thinking blocks, and a computer-use change. Test the exact harness before switching.
Who should test it first
Coding-agent teams are the obvious first testers because repository sessions can reuse large cached prefixes. Run the same tasks at fixed effort and record accepted results, retries, latency, cache use, output, reviewer minutes, and total billed cost. The agent cost calculator converts those inputs into cost per successful task.
For a managed open-model control, run the same replay through Novita’s OpenAI-compatible API and compare cost per accepted result.
Affiliate disclosure: AI Pricing Guru may earn a commission from the sponsored Novita link at no extra cost to you. It does not affect this analysis.
Buyer verdict
Opus 5.5 is a real price move. Test it first for premium Claude work, keep Fable 5.1 as a measured escalation, and use GPT-6 Sol or Sonnet 5 as lower-rate controls.
Sources: Anthropic’s official Opus 5.5 announcement, model overview, and API pricing table. Pricing, availability, limits, and benchmark qualifications verified September 22, 2026 at 20:55 UTC.