Claude Opus 5.5 cost model
Updated September 22, 2026 · first published September 22, 2026
Anthropic released Claude Opus 5.5 (claude-opus-5-5) on September 22, 2026. The headline price is $4 per million input tokens and $20 per million output tokens, down from $5 and $25 on Opus 5. That is a 20% cut. For most production workloads the actual cut is larger, because the rate that moved most is the one the headline leaves out.
The rates that matter
| Token type | Opus 5.5 | Opus 5 | Fable 5.1 |
|---|---|---|---|
| Input (uncached) | $4.00 / MTok | $5.00 / MTok | $10.00 / MTok |
| Output | $20.00 / MTok | $25.00 / MTok | $50.00 / MTok |
| Cache write (5-minute) | $5.00 / MTok | $6.25 / MTok | $12.50 / MTok |
| Cache read | $0.20 / MTok | $0.50 / MTok | $0.25 / MTok |
| Fast mode (input / output) | $8 / $40 | — | — |
Every row is 20% lower except the cache read, which is 60% lower. At $0.20 it is now cheaper than Fable 5.1's $0.25. That reverses the one row where Fable used to win. In our Fable 5.1 cost model, a heavily cached workload could come out cheaper on Fable than on Opus 5. Against Opus 5.5 that no longer happens. Fable costs more on every token type.
Why a typical task costs about 40% less
Anthropic's claim is a 40% drop in cost for a typical job, not 20%. The extra comes from token efficiency. Anthropic says Opus 5.5 reaches the same result with about 20% fewer input and output tokens and writes about 40% less verbose output. A lower rate times fewer tokens compounds: 0.8 × 0.8 = 0.64, which is roughly 36–40% off depending on the mix.
Here is one agent step: 200,000 input tokens, 90% served from cache, and 20,000 output tokens.
| Line | Opus 5 | Opus 5.5 (same tokens) | Fable 5.1 |
|---|---|---|---|
| 20K uncached input | $0.100 | $0.080 | $0.200 |
| 180K cached input | $0.090 | $0.036 | $0.045 |
| 20K output | $0.500 | $0.400 | $1.000 |
| Total | $0.690 | $0.516 | $1.245 |
On identical tokens, Opus 5.5 is 25% cheaper than Opus 5, not 20%, because of the cache discount. If the 20% token reduction holds for your workload, the step drops to about $0.41, which is 40% below Opus 5 and a third of the Fable cost. Output is still about 80% of the bill in this example. That is where the verbosity reduction pays off.
Cost per task against GPT-6 Sol, GPT-6 Astra and Fable 5.1
List prices per million tokens, and the same agent step as above (200K input, 90% cached, 20K output) priced on each model.
| Model | Input $/MTok | Output $/MTok | Cache read $/MTok | Same agent step |
|---|---|---|---|---|
| GPT-6 Sol | $2.00 | $10.00 | $0.20 | $0.276 |
| Claude Opus 5.5 | $4.00 | $20.00 | $0.20 | $0.516 |
| Claude Opus 5 | $5.00 | $25.00 | $0.50 | $0.690 |
| Claude Fable 5.1 | $10.00 | $50.00 | $0.25 | $1.245 |
| GPT-6 Astra | $10.00 | $50.00 | $1.00 | $1.380 |
Measured cost per task on the Artificial Analysis Intelligence Index v4.3.2, read 22 September 2026, sorted by score. GPT-6 Sol is the cheapest way to any score up to 47.5. Above that, Opus 5.5 is the cheapest at every level: Opus 5.5 high scores 53.6 for $1.82, while GPT-6 Astra max scores 52.7 for $3.26 and Fable 5.1 max scores 53.4 for $7.63.
| Model | Intelligence Index | Cost per task | Output tokens per task |
|---|---|---|---|
| GPT-6 Sol · medium | 39.8 | $0.25 | 6,500 |
| GPT-6 Sol · xhigh | 44.1 | $0.53 | 16,000 |
| GPT-6 Astra · low | 45.8 | $0.82 | 4,400 |
| GPT-6 Sol · max | 47.5 | $1.06 | 31,200 |
| Claude Opus 5 · high | 48.1 | $3.61 | 46,200 |
| Claude Opus 5 · max | 50.8 | $5.86 | 72,500 |
| Claude Opus 5.5 · medium | 51.2 | $1.34 | 25,700 |
| Claude Fable 5.1 · high | 51.2 | $3.91 | 38,100 |
| GPT-6 Astra · max | 52.7 | $3.26 | 27,200 |
| Claude Fable 5.1 · max | 53.4 | $7.63 | 78,100 |
| Claude Opus 5.5 · high | 53.6 | $1.82 | 35,600 |
| Claude Opus 5.5 · max | 57.6 | $5.98 | 119,200 |
Every GPT-6 Astra setting above low, every Opus 5 setting and every Fable 5.1 setting costs more than a Sol or Opus 5.5 setting that scores the same or higher. Cost per task depends on the benchmark's task mix. Replay your own traffic before switching models.
Is it good enough to replace Fable?
Anthropic positions Opus 5.5 as matching Fable 5.1 on most tasks. The published numbers support that for coding and agentic work:
- Terminal-Bench 4.0: 66.4%, against 52.3% for Opus 5 and 55.8% for Fable 5.1.
- FrontierCode v1.1: 54.4%, against 48.0% for Opus 5 and 50.3% for Fable 5.1.
- CursorBench 4.0: 57.8%, against 46.6% for Opus 5.
- GDPval-AA v2.1: 1846 Elo, against 1708 for Opus 5.
Artificial Analysis ranked it first on its Intelligence Index on launch day. Anthropic also says that at default (medium) effort it beats GPT-6 Astra at max effort on knowledge work for about a fifth of the cost per task. That is a vendor comparison, so treat it as a hypothesis to test. The cost math above is the part you can verify yourself.
Fast mode costs 2×
Fast mode costs $8 in and $40 out, exactly double the standard rate, for up to 2.5× the speed. The standard model is also about 30% faster than Opus 5. Use fast mode for latency-bound interactive paths where a user is waiting. Keep it off batch jobs, background agents and evals. When fast mode is enabled by default, the unit cost doubles and the dashboard shows nothing unusual until the invoice arrives.
What to do this week
- Pin the model ID. Point production at
claude-opus-5-5explicitly and record the ID on every request, so cost per request can be split by model after the migration. - Re-baseline on your own traffic. Replay a sample of real requests on both models and compare tokens per completed task, not per call. The 20% token-efficiency claim is the part most likely to vary by workload.
- Re-check effort settings. If a workload moved to Fable or to max effort for quality, test Opus 5.5 at medium first. That is the setting Anthropic benchmarks against.
- Re-run the Fable routing decision. Any route that sent cache-heavy traffic to Fable to get the cheaper cache read should be re-evaluated. That cost advantage is gone.
- Watch cache hit ratio. At $0.20 against $4.00, a cache miss now costs 20× a hit. A prompt-template change that breaks the prefix costs more than it did on Opus 5.
Anthropic says Sonnet 5.5 and Haiku 5.5 will follow in the coming weeks. Keep the replay harness from step 2 so each release can be re-baselined in an afternoon.
Related
- Claude Fable 5.1 cost model
- Prompt caching economics
- Opus 5.5 effort levels are the real price
- Opus 5.5 fast mode economics
Want this applied to your own LLM spend? FinOps LLM runs a free audit of your AI costs and shows where the savings are. Book free audit →