Claude Opus 4.8
Hybrid reasoning model built for serious coding and AI agents, featuring a 1M context window
Announcements
- NEW
Claude Opus 5.5
Sep 22, 2026
We’re introducing Claude Opus 5.5. It performs at the level of Claude Fable 5.1 on most work and costs 40% less to run than Opus 5.
Read more
Claude Opus 5
Jul 24, 2026
A step-change improvement for the Opus tier: stronger coding, more capable agents, and sharper professional work.
Read more
Claude Opus 4.8
May 28, 2026
Stronger across coding, agentic tasks, and professional work, Opus 4.8 has the consistency and autonomy to keep working on long-running tasks.
Read more
Claude Opus 4.7
Apr 16, 2026
Claude Opus 4.7 brings stronger performance across coding, vision, and complex multi-step tasks. It’s more thorough and consistent on difficult work, with better results across professional knowledge work.
Read more
Claude Opus 4.6
Feb 5, 2026
Claude Opus 4.6 is our most capable model to date. Building on the intelligence of Opus 4.5, it brings new levels of reliability and precision to coding, agents, and enterprise workflows.
Read more
Availability and pricing
Claude Opus 5.5 is our strongest Opus model yet, powering long-running, highly capable agents while delivering improvements in coding and professional work.
For business users and consumers who want to collaborate with a powerful model on complex tasks, Opus 5.5 is available on Claude for Pro, Max, Team, and Enterprise users.
For developers interested in building AI solutions that demand strong intelligence, Opus 5.5 is available on the Claude Platform natively, and in Amazon Web Services, Google Cloud, and Microsoft Foundry. Pricing for Opus 5.5 will cost an estimated 40% less to run than Opus 5 for typical workloads billed by token. Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, 20% below Opus 5. Cache reads, which are a large share of the cost of long-running agentic work, also now cost 60% less than Opus 5, at $0.20 per million tokens. To learn more, check out our pricing page. To get started, use claude-opus-5-5 via the Claude API.
Fast mode for Opus 5.5 is also available now in Claude Code and on the Claude Platform with up to 2.5x faster speed. It costs $8 per million input tokens and $40 per million output tokens.
For workloads that need to run in the US, US-only inference is available at 1.1x pricing for input and output tokens. Learn more.
Use cases
Claude Opus 5.5 is our most capable Opus model yet for coding, agents, and knowledge work. It costs less per token than Opus 5 and uses fewer tokens per task, so work on the Claude Platform, in Claude Code, and in the Claude apps costs about 40% less for work billed by token than on Opus 5.
Opus 5.5 also communicates more clearly. It leads with what matters, avoids jargon, and follows your writing rules, which makes it a better partner over long sessions. Key use cases include:
Advanced coding
Opus 5.5 is our strongest Opus model for agentic coding. It handles long-running work in large codebases, including building features, debugging, refactoring, and code review. It finds the root cause before changing anything, checks its work as it goes, and explains its changes in plain language, so engineers can review and trust them quickly.
Agents
Opus 5.5 is the Opus tier’s strongest agentic model, reliably orchestrating complex multi-tool tasks. It plans deliberately, coordinates subagents, uses memory to learn across sessions, and drives long-running work forward with minimal oversight. Along the way, it reports back clearly on what it did, what it found, and what it needs next.
Enterprise workflows
Opus 5.5 is built to be the enterprise daily driver, powering agents that run projects end-to-end. It follows instructions precisely, stays in scope, and produces professional-grade spreadsheets, slides, and docs that are ready to use.
Financial analysis
Opus 5.5 brings deeper reasoning and precision to financial workflows. It reads dense filings, models, and charts accurately, carries context across an entire deal or reporting cycle, and handles the nuance of compliance-sensitive work, with clear summaries of what it found and how it got there.
Vision & computer use
Opus 5.5 is our best Opus model for vision and computer use. It reads dense documents, charts, screenshots, and diagrams at high fidelity, making it reliable for document extraction, visual analysis, and interpreting complex real-world imagery. Opus 5.5 brings deep reasoning to computer use, handling multi-step tasks that span multiple applications and require planning and judgment.
Benchmarks
Claude Opus 5.5 delivers the intelligence and reliability to be your daily driver for serious coding and knowledge work.

Trust and safety
Extensive testing and evaluation ensures the release of Opus 5.5 meets Anthropic’s standards for safety, security, and reliability. The accompanying system card covers safety results in depth.
Safeguards
Opus 5.5 is the first Opus model to launch with a similar class of safeguards to Fable 5.1 in cybersecurity, biology, and anti-distillation. As our models grow more powerful, stricter safeguards are one way we prevent new capabilities from becoming tools for misuse.
Hear from our customers
Developers want agents that can take on real software work and finish it. In our testing across GitHub Copilot CLI and VS Code, Claude Opus 5.5 used among the fewest tokens and steps we measured. In VS Code, it solved more terminal tasks than Opus 5 in less than half the steps. More than making individual tasks more efficient, it’s making developers’ bigger projects more achievable.
I handed Claude Opus 5.5 a large engineering task across six of our repositories and let it run overnight, unattended. It stayed on task for over 18 hours defining how our services talk to each other and working out how each one should apply that. Compared with Opus 5, it hit milestones faster and required minimal reworking. Its code comments were short and useful instead of long and prose-heavy. I’m struggling to find anything negative to say.
For Lovable builders, Opus 5.5 means faster builds with the same quality, whether you’re starting from scratch or working on a live app. It gathers context once, makes fewer and more complete edits, and doesn’t get stuck retrying, finishing in a third to half fewer steps and using significantly fewer tokens along the way.
On our genomics analysis work, Claude Opus 5 behaves more like a careful scientist than any model we’ve run. It reaches for the right statistical tests to rule out confounders, cross-checks its own results by independent methods, and stays on track through long multi-step analyses.
With Claude Opus 5.5, we’ve seen a clear improvement in token efficiency across our internal evaluations, as we’ve been able to complete the same tasks both cheaper and faster.
We test models on real engineering and trading-desk work. On our agentic coding tasks, Claude Opus 5.5 matched Opus 5’s quality in about half the turns, time and output tokens, cutting the cost of that workload by 40 to 50%. It posted the highest score we’ve recorded on one desk’s trading-support suite, passing tasks earlier Claude models had failed, and topped all eight models on our analysis task.
Claude Opus 5.5 delegates to subagents far more effectively and checks its own work in creative ways. Self-verification loops feel easier to set up. It found savings opportunities in our cloud bill that previous models had missed, and in code review it caught a bug by checking external docs for a third-party integration we’d modeled wrong several commits earlier.
Every call an agent makes is time and cost a developer feels. On a public benchmark of real command-line tasks, Claude Opus 5.5 solved more than Opus 5 while making about 40% fewer calls and using half the tokens. For developers building with Kiro, that means faster, more affordable agent sessions for routine tasks and complex challenges alike. Opus 5.5 will soon be available in Kiro.
Even at its lowest effort setting, Claude Opus 5.5 caught 72% of known bugs in our code reviews to Opus 5’s 56% at high effort, with fewer false alarms and a fraction of the output. On US consulting analysis, low thinking effort matched its higher thinking settings on half the output and passed our quality checks. When more lower thinking efforts are deployed in production, that’s client-ready work delivered efficiently.
Financial firms need outputs that are consistently correct. At its lowest effort setting, Claude Opus 5.5 beat Opus 5 at high effort on our BigFinance Bench with about 60% fewer output tokens. Its answers are shorter and better structured, and its slides come out denser, more in line with industry standards.
Evaluating new models is central to the multi-model approach behind the LexisNexis Legal Intelligence Engine. In our initial evaluations, Claude Opus 5.5 identified highly relevant citations consistently, demonstrated strength with statutes, and structured its answers around the central legal frameworks and key issues. These are the kinds of capabilities we look for to help our customers accomplish more with Lexis+ with Protégé.
In quant research, one wrong assumption can undermine a result. At its lowest effort setting, Claude Opus 5.5 largely solved our evaluation task. At higher settings, it went even further: it detected that the minute indexing in our own instructions was off by one and corrected for it, noting that this would cost it points with the grader. It was right, and no model we’ve tested had caught and acted on that before.
As models get better at data work, we’re seeing more convincing-sounding conclusions the data doesn’t support. Claude Opus 5.5 keeps digging past the first plausible answer. One task in our DataBench benchmark asks whether packages were late or tracking was just slow. Opus 5 checked delivery confirmations and called tracking healthy. Opus 5.5 found the packages were late and tracking was broken too. We’re bringing it into the Hex agent for this work.
CoCounsel combines multiple models with our content and expertise for complex legal work. With Claude Opus 5.5, we’re seeing better results in our expert evaluations and on our internal benchmarks, alongside gains in speed and token efficiency. We’re excited for customers to experience that difference in the back-and-forth with CoCounsel as a sounding board, weighing evidence and refining their thinking in ways benchmarks don’t fully capture.
On end-to-end finance workflows graded against expert rubrics, Claude Opus 5.5 covered 86.6% of what we look for versus 60.3% for Opus 5. On retrieval evals, it achieved our best-ever citation recall with better token efficiency than Opus 5, which keeps our cost per research task in check.
Viktor is an AI employee that lives in Slack and Microsoft Teams, so every step he takes shows up in our costs. At the same effort, Claude Opus 5.5 needs fewer steps and tool calls per task than Opus 5 and costs nearly half as much, while getting twice as many of our hardest tasks right.
Verbose, hard-to-follow output has been my biggest frustration with frontier models, and Claude Opus 5.5 fixes it. It writes like a good colleague, and follows our writing rules. A design spec came out usable with very minimal edits, and when it rewrote one of our prompts I preferred its version to my own. When it optimized our test suite, I could follow its reasoning easily and shipped the change with confidence.
I run long Claude Code sessions every day. On a multi-day rebase of 40 stacked pull requests, one Claude Opus 5.5 session directed a dozen more sessions and laid out every conflict plainly. On the calls that it held, it framed them clearly that after hours away I could answer in minutes. All 40 passed CI the next afternoon. It’s a substantial upgrade over Opus 5.
Our customers use Box AI on enormous amounts of content, so speed and cost are a top priority. In our evaluations, Claude Opus 5.5 used a third of the tokens Opus 5 did, and its answers were 40% less verbose without losing accuracy. We expect that to matter a lot for teams running agents across their content in areas like financial services and the public sector.
Overnight, Claude Opus 5.5 autonomously handled a bug in our Lakehouse services layer that I hadn’t had time to diagnose. It investigated, designed the fix, and implemented it on its own. By morning the change was done and passed our test suite. Its writing is easy to follow and more coherent than Opus 5’s. Our pull requests and user-facing docs have needed almost no editing.
Claude Opus 5.5 is the first model we’d default to at medium effort. In our testing it matched Opus 5 on high effort, while using 20 to 25% fewer output tokens. On long, messy investigations it always came back with a clear, actionable answer. This means our customers get more done for less.
Frequently asked questions
We offer Claude models across the spectrum of speed, price, and performance. We recommend Opus 5.5 as your daily driver for coding and knowledge work—particularly production-ready code, complex document creation, and computer use.
Pricing depends on how you want to use Opus 5.5. To learn more, check out our pricing page.