1. Home
  2. Blog
  3. AI Tools

Claude Opus 5.5: What's New, Pricing, and What It's Best For

Lindy Drope
Lindy Drope
Founding GTM at Lindy
Lindy leads GTM at Lindy and is the team’s most prolific automation builder. She publishes weekly educational videos and articles on building AI assistants – And yes, she’s a real person!
Lindy Drope
Written by
Lindy Drope
Flo Crivello
Flo Crivello
Founder and CEO of Lindy
Flo Crivello is the founder and CEO of Lindy. Before that, he founded Teamflow and was a product manager at Uber. He writes about technology, startups, and the future of work on his blog.
Flo Crivello
Reviewed by
Flo Crivello
Last Updated:
September 30, 2026
Expert Verified

New flagship models usually show up with two things: a stack of benchmark charts and a bigger bill. Claude Opus 5.5 arrived on September 22 with the charts, sure, but the bill went the other way.

It costs less per token than the Opus 5 it replaces, and Anthropic says it finishes the same work with fewer tokens, too. (If you're already hunting for the catch, same. There are a couple, and they're hiding in the fine print.)

So I went through everything Anthropic published about it: the launch post, the model docs, the pricing pages, and the Claude Code configuration notes.

The footnotes alone could fill a weekend, and they're where the surprises live, like a new default setting that can change your results before you've touched a line of code.

The short version is right below. After that, I'll get into what changed from Opus 5, how it scores, what it costs, where it runs, and which effort setting to start on before you move real work over.

TL;DR:

  • Released: September 22, 2026, as the first model in Anthropic's new Claude 5.5 family.
  • What it is: Anthropic's Opus model for long-running agentic coding and knowledge work, pitched as matching Claude Fable 5.1 on most work.
  • Price: $4 per million input tokens and $20 per million output tokens, with cache reads at $0.20 (Opus 5 was $5, $25, and $0.50).
  • Real-world cost: about 40% less than Opus 5 on typical workloads, per Anthropic, because it's cheaper per token and uses fewer tokens per task.
  • Where to use it: Claude's Pro, Max, Team, and Enterprise plans, Claude Code, the Claude API, Amazon Bedrock, Claude Platform on AWS, Google Cloud, and Microsoft Foundry.
  • The setting to know: effort now defaults to medium, and thinking can't be switched off.

What is Claude Opus 5.5?

Claude Opus 5.5 is Anthropic's newest Opus model, built for long-running agentic coding and knowledge work. Anthropic announced it on September 22, 2026, as the first release in its new Claude 5.5 family.

The headline claim is that it performs at the level of Claude Fable 5.1 on most work and costs 40% less to run than Opus 5. That's Anthropic's framing, so treat it as a strong hypothesis to test on your own tasks.

You get a 1M-token context window and up to 128K output tokens, a reliable knowledge cutoff of June 2026, and the API model ID claude-opus-5-5. Adaptive thinking is always on, and you control how hard it thinks with an effort setting.

If the Claude lineup were a restaurant kitchen, Fable 5.1 would be the head chef you call in for the hardest dishes, Sonnet 5.5 the fast line cook, and Opus 5.5 the sous chef who can run most of the service alone.

Anthropic's own model docs now tell developers to start with Opus 5.5 for most workloads. Sonnet 5.5 followed on September 28, and Haiku 5.5 is due in the coming weeks.

If you're still deciding whether Claude fits your work, it's worth comparing a few Claude alternatives first.

What changed from Opus 5

Opus 5.5 is cheaper, faster, and more frugal with tokens, and it also changes two defaults that can trip up anyone migrating code. Here's how the two compare on the points that touch your bill and your workflow:

🔍 What changed ⏮️ Opus 5 ⏭️ Opus 5.5
Input / output price $5 / $25 per MTok $4 / $20 per MTok
Cache reads $0.50 per MTok $0.20 per MTok
Output speed Baseline 30%+ faster
Default effort High Medium
Turning thinking off Allowed at high or below Not allowed

Cost per task dropped more than the price did. Per-token prices fell 20%, but Anthropic says Opus 5.5 also uses fewer tokens to finish the same job, which is how it gets to roughly 40% cheaper on typical workloads.

One early tester used it to audit and fix a 200,000-line codebase in under three hours, a job where Opus 5 took over 20 hours and used 2.5x as many tokens.

The defaults moved. A request that doesn't set an effort level now runs at medium, one notch below Opus 5's high, and thinking can't be disabled. Code that already ran on Opus 5 with thinking on needs no change on that front.

It writes more clearly. Early testers found its writing easier to follow, and the launch post says it puts the most important information up front and follows the writing rules you give it. (If you've ever asked a model for three bullet points and received a short novella, this one's for you.)

Subscribers get more room. Anthropic is raising five-hour usage limits on Pro, Max, Team, and seat-based Enterprise plans, and it's giving subscribers a rate limit reset they can save and use whenever they choose.

It behaves better under pressure. Anthropic reports that Opus 5.5 posted the best scores of any model it has tested on its automated behavioral audit, and that it resists prompt injection better than Opus 5.

Claude Opus 5.5 benchmarks

Anthropic's launch table puts Opus 5.5 ahead of Fable 5.1 and Opus 5 on every test it reports, and ahead of GPT-6 Astra on four of the six where OpenAI has a score. These are Anthropic's numbers, run at max effort (xhigh on Terminal-Bench 4.0):

📊 Benchmark ⏭️ Opus 5.5 🧠 Fable 5.1 ⏮️ Opus 5 🤖 GPT-6 Astra
Terminal-Bench 4.0 (agentic coding) 66.4% 55.8% 52.3% 57.9%
FrontierCode v1.1 (agentic coding) 54.4% 50.3% 48.0% 53.3%
GDPval-AA v2.1 (knowledge work, Elo) 1846 1735 1708 1542
AutomationBench (business workflows) 40.0% 31.4% 26.9% 41.4%
Humanity's Last Exam (reasoning, with tools) 67.7% 65.6% 63.6% 57.2%
Terminal-Bench-Science (research) 58.7% 52.6% 29.0% 64.6%
OSWorld 2.1 (computer use) 81.8% 80.7% 74.0% n/a

Astra keeps two wins. GPT-6 Astra leads on AutomationBench, Zapier's test of business workflows across connected apps, and on Terminal-Bench-Science. Zapier counted every safeguard intervention as a failure, which Anthropic says left Opus 5.5 with a lower score than it would get in practice.

Efficiency is the clearer win. At its default medium effort, Opus 5.5 matches Astra on Terminal-Bench 4.0 for about 40% of the cost per task, and on FrontierCode it beats Astra's top score for about a fifth of the cost.

To see how the two do on the same real tasks, and where Astra comes out ahead, read our Opus 5.5 vs Astra comparison.

Claude Opus 5.5 pricing

Opus 5.5 is billed per million tokens (MTok) on the Claude API and cloud platforms. Here are the current list prices next to Opus 5:

💵 Price per MTok ⏭️ Opus 5.5 ⏮️ Opus 5
Input tokens $4 $5
Output tokens $20 $25
Cache writes (5-minute) $5 $6.25
Cache writes (1-hour) $8 $10
Cache reads $0.20 $0.50

‍The cache-read price is the one to watch. Cache reads make up the majority of agentic and coding work costs, according to the launch post, and that line fell 60%. If your agents reread the same codebase or document set all day, that's where most of your savings will come from.

A few other pricing options are worth knowing about:

  • Fast mode: up to 2.5x faster output for $8 per MTok input and $40 per MTok output, available on the Claude API and in Claude Code.
  • Batch processing: half price, at $2 input and $10 output per MTok, for work that doesn't need an instant answer.
  • US-only inference: 1.1x the standard input and output prices for workloads that have to run in the US.

If you use Opus 5.5 inside the Claude apps, your plan price covers it, and the plan's usage limits decide how much Opus you get each week.

Where you can use Opus 5.5

In the Claude apps

Opus 5.5 is available to Claude Pro, Max, Team, and Enterprise users. The Free plan doesn't include Opus models, according to Claude's pricing page, so you'll need a paid plan to try it in chat.

In Claude Code

Opus 5.5 works in Claude Code from version 2.1.280 or later, so run claude update if you're behind. On the Anthropic API, Amazon Bedrock, Google Cloud, and Claude Platform on AWS, the opus model alias now points to Opus 5.5.

Microsoft Foundry is the one catch, because the opus alias there still resolves to Opus 4.6, so you'll want to select claude-opus-5-5 by its full name. Fast mode works in Claude Code too, if you're the impatient type (no judgment).

If you run Claude Code alongside other AI coding agents, check each tool's model picker, since support rolls out tool by tool. Kiro, for example, now offers Opus 5.5 with experimental support.

On the API and cloud platforms

Developers can call claude-opus-5-5 on the Claude API today. It's also on Amazon Bedrock, Claude Platform on AWS, Google Cloud, and Microsoft Foundry, with Bedrock using the ID anthropic.claude-opus-5-5.

Fast mode is the exception. It runs only on Anthropic's first-party Claude API, so you won't find it on Bedrock, Claude Platform on AWS, Google Cloud, or Foundry.

‍

{{templates}}

‍

How effort levels work on Opus 5.5

Effort is the dial that decides how much Opus 5.5 thinks before it answers, and because thinking is always on, it's also your main lever on cost and speed. There are five levels:

⚙️ Effort level 🎯 Use it for
Low Simple, scoped tasks and subagents
Medium (default) Everyday agentic work that balances speed and cost
High Complex reasoning and difficult coding problems
xhigh Long coding and agent runs of 30+ minutes
Max The hardest problems, with no cap on token spend

The benchmark scores and the default setting don't match. Those headline numbers come from max effort, while the default is medium, so a workflow that never touches effort gets the cheaper, faster setting.

That's less scary than it sounds. Deloitte told Anthropic that Opus 5.5 at its lowest effort caught 72% of known bugs in code reviews, compared with 56% for Opus 5 at high effort, and Rogo reported a similar lowest-effort win on its finance benchmark.

My suggestion is to start on medium, then rerun your most common tasks at low and high and compare the output and the bill. Anthropic's docs recommend the same thing: run a fresh effort sweep on your own evals whenever you move up from an older model.

What Claude Opus 5.5 is best at

Anthropic pitches Opus 5.5 as a daily driver for serious coding and knowledge work, and the early-tester examples in the launch post back that up with specifics. Keep in mind that these are vendor-published results.

Agentic coding and big migrations

This is the use case Anthropic leans on hardest. It says Opus 5.5 is particularly good at long, sprawling jobs like codebase-wide migrations and audits, and one early tester finished a 680,000-line code migration in less than a day.

In an internal test, Anthropic had Opus 5.5 and Fable 5.1 translate HAProxy from C into Rust. Both rewrites passed nearly all of HAProxy's regression tests, but Opus 5.5 finished in 9.5 hours compared with 12 and cost 51% less.

Long-running agents

Opus 5.5 is built to keep working for hours with little supervision. A developer at Clio handed it a task spanning six repositories, let it run overnight, and said it stayed on task for over 18 hours.

Column's engineering team says it delegates to subagents far more effectively, and Anthropic reports it's much less likely than recent models to take hard-to-reverse actions. That combination matters if you run AI agents unattended.

Research and knowledge work

Anthropic asked Opus 5.5, Fable 5.1, and Opus 5 to write reports on a company's quarterly results from a copy of the web, and a grader failed any invented figure or quote. Sixteen of Opus 5.5's 18 reports passed, while neither of the other models passed once.

It also leads GDPval-AA v2.1, an Artificial Analysis test of real-world work across 44 occupations, by more than 100 Elo points over Fable 5.1.

Charts, screenshots, and computer use

According to Anthropic's docs, Opus 5.5 reads dense charts, diagrams, and screenshots much more precisely than earlier models, and Anthropic calls it its best Opus model for vision and computer use. That helps with document extraction and multi-step tasks that hop between apps.

Trade-offs to know before you switch

Few model launches are all upside, and Opus 5.5 has a few catches worth knowing before you swap model IDs:

  • Thinking is always on: requests that try to disable it get a 400 error, and forced tool use (setting tool_choice to any or a specific tool) is gone too, so API code that relied on either needs a small rewrite.
  • Computer use needs the new toolset: on the Claude API and Google Cloud, Opus 5.5 rejects the older computer use tool and only accepts the newer toolset, so computer use agents built on Opus 5 need updating. Amazon Bedrock keeps working as before.
  • The new default can change your results: requests without an effort setting now run at medium, so a pipeline tuned on Opus 5 at high may behave differently. Pin the effort level you want.
  • Security and biology work hits safeguards: per the launch post, most cybersecurity tasks get re-routed to Opus 4.8, and biology research runs under stricter safeguards unless your organization is approved for its Life Sciences Verification Program.
  • The benchmarks are Anthropic's: most numbers come from Anthropic or its early testers, and Anthropic itself calls benchmark margins "a less reliable guide to real-world differences" at this level.
  • Usage limits still apply: higher five-hour limits help, but subscription plans still cap how much Opus you can use, and the Free plan doesn't include it.

‍

{{cta}}

‍

How Opus 5.5 compares to Fable 5.1 and Sonnet 5.5

Fable 5.1

Fable 5.1 still sits above Opus 5.5 in Anthropic's lineup, at $10 input and $50 output per MTok, which is 2.5x the price. Anthropic's advice is to start with Opus 5.5 and move up to Fable 5.1 for demanding reasoning or when Opus 5.5 at higher effort still falls short.

Opus 5.5 outscores Fable 5.1 on every benchmark in Anthropic's launch table, though Anthropic says the difference in its own use is smaller than the scores suggest.

For most teams, the practical order is to try Opus 5.5 first. Our Opus 5.5 vs Fable 5.1 comparison covers the cases where Fable 5.1 is still worth paying for.

Sonnet 5.5

Sonnet 5.5 arrived on September 28, 2026, six days after Opus 5.5, at half the price: $2 input and $10 output per MTok, with the same $0.20 cache reads. Anthropic pitches it for well-scoped everyday tasks, bug fixes, and polished documents, slides, and spreadsheets.

It even beats Opus 5.5 on Terminal-Bench 4.0 (70.6% vs 66.4%). Anthropic still says Opus 5.5 "remains clearly stronger at complex, open-ended work requiring sustained judgment," so a sensible split is Sonnet 5.5 for quick, scoped work and Opus 5.5 for long agent runs.

Find your cheapest good-enough setting in one afternoon

Claude Opus 5.5 is a rare upgrade where the newer model is also the cheaper one, so for most teams on Opus 5 the real work is picking the right effort setting, which you can do in a single afternoon.

Pick a handful of tasks your team runs every week, like a code review, a bug fix, a research brief, and a spreadsheet model, then run each one on Opus 5.5 at low, medium, and high. Compare the quality and the token bill side by side, and pin the lowest setting that still passes.

If a lot of your team's AI work runs through an AI teammate like Lindy, which lets you pick the model for each task, check which models your workspace lists before you plan the switch around it.

Once the test is done, you'll know which setting to use, and if Anthropic's 40% figure holds on your workload, the savings show up on every invoice after that.

FAQ

What can Claude Opus 5.5 do?

Claude Opus 5.5 handles long-running agentic coding, research, and knowledge work, including codebase migrations, code review, financial models, reports, and multi-step computer use. It has a 1M-token context window and can stay on a single task for hours.

Is Claude Opus 5.5 better than GPT-6 Astra?

Claude Opus 5.5 beats GPT-6 Astra on most benchmarks in Anthropic's launch table, including Terminal-Bench 4.0 (66.4% vs 57.9%), while Astra leads on AutomationBench and Terminal-Bench-Science. These are Anthropic's numbers, so test both if you're choosing between Claude and ChatGPT.

Is Opus 5.5 available in Claude Code?

Opus 5.5 is available in Claude Code from version 2.1.280. On the Anthropic API, Amazon Bedrock, Google Cloud, and Claude Platform on AWS, the opus alias selects it automatically, and Microsoft Foundry users select claude-opus-5-5 by name.

Should I upgrade from Opus 5 to Opus 5.5?

Most Opus 5 users should upgrade to Opus 5.5, because it's cheaper per token, generates output over 30% faster, and uses fewer tokens per task.

Before you switch API code, remove any setting that disables thinking or forces tool use, move computer use agents to the new toolset, and set your effort level explicitly.

How does Opus 5.5 compare to Sonnet 5.5?

Opus 5.5 is the stronger model for complex, open-ended work, while Sonnet 5.5 costs half as much at $2 input and $10 output per MTok and responds faster. Both have a 1M-token context window, so Sonnet 5.5 suits well-scoped everyday tasks and Opus 5.5 suits long, judgment-heavy jobs.

Save 2 Hours Every Day
Lindy is your ultimate AI assistant that manages inbox, meetings, and follow-ups—so you stay ahead of the chaos.
Try Lindy for Free
About the editorial team
Lindy Drope
Lindy Drope
Founding GTM at Lindy

Lindy leads GTM at Lindy and is the team’s most prolific automation builder. She publishes weekly educational videos and articles on building AI assistants – And yes, she’s a real person!

Flo Crivello
Flo Crivello
Founder and CEO of Lindy

Flo Crivello is the founder and CEO of Lindy. Before that, he founded Teamflow and was a product manager at Uber. He writes about technology, startups, and the future of work on his blog.

Ready when you are.

Free to try. In your Slack in two minutes.

Try for free
7-day free trial • Cancel anytime