Claude Opus 5.5 Launches With 40% Cost Cut Ahead of Anthropic IPO
Claude Opus 5.5 and rival frontier models, September 2026
Verified Sep 22, 2026| Model | Lab | Released | Input/Output per 1M tokens | Status and notes | Sources |
|---|---|---|---|---|---|
| Claude Opus 5.5 | Anthropic | Sept 22, 2026 | $4 / $20 (cache reads $0.20) | New flagship; Fable 5.1-level performance, ~40% cheaper to run than Opus 5, ~30% faster | [1][9][10][21][23] |
| Claude Opus 5 | Anthropic | July 24, 2026 | $5 / $25 | Prior Opus; ~Fable 5 intelligence at half the price; 1M-token context | [4][6][26][30] |
| Claude Fable 5.1 | Anthropic | Sept 1, 2026 | $10 / $50 | Premium frontier tier; 1M context; Terminal-Bench 4.0 score 55.8 | [4][25] |
| Claude Sonnet 5 | Anthropic | Current | $2 / $10 | Production default; 1M context; Sonnet 5.5 expected in coming weeks | [4][5][22] |
| Claude Haiku 4.5 | Anthropic | Current | $1 / $5 | Budget tier, 200K context; Haiku 5.5 expected in coming weeks | [4][5][22] |
| GPT-6 Astra | OpenAI | Sept 3, 2026 | Not stated in coverage | ~13% enterprise AI spend vs ~8% for Claude Fable; surpassed Anthropic on OpenRouter weekly spend | [25][28] |
| Gemini 3.8 Flash (+ Cyber variant) | Sept 2, 2026 | Not stated in coverage | Shipped alongside Meta's Muse Spark 1.3; Gemini 3.5 Pro announced but undated | [41] |
Anthropic has shipped Claude Opus 5.5, the first model in a new Claude 5.5 generation, in a launch that reads less like a routine upgrade and more like a company repositioning itself for the most scrutinized balance sheet in software history. The model went live on Tuesday across Anthropic's own apps and API as well as Amazon Web Services, Google Cloud, and Microsoft Azure, and the pitch is unusually simple: Fable-class performance at a mid-tier price2210.
What Anthropic actually shipped
The company's own framing is that Opus 5.5 performs at the level of Claude Fable 5.1, its premium frontier model, on most work, while costing roughly 40% less to run than its predecessor Opus 52130. That 40% figure combines a sticker price cut with reductions in token usage — the model is reported to need fewer tokens to complete equivalent tasks, which matters enormously for agent-style workloads where output volume, not headline price, dominates the bill241.
The per-token pricing is concrete. Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, a 20% cut from Opus 5's $5/$25, while cache reads fall from $0.50 to $0.20 per million tokens — the number that quietly reshapes economics for long-running coding agents that re-read context constantly1910. Anthropic is also adjusting Claude Code around the launch: five-hour session limits rise 20%, and because the new model is cheaper, it effectively goes 25% further within those limits, with Pro, Max, and Team users getting a usage reset to spend anytime9.
Benchmark claims from Anthropic put the new Opus ahead of both Fable 5.1 and OpenAI's more expensive GPT-6 Astra on most tasks, with third-party analysis from Artificial Analysis folded into some coverage24. Early testers reported genuinely large gains rather than the usual few-point margins: one completed a 680,000-line code migration in under a day, work that would have taken an engineering team weeks, and on a task asking the model to cut load times across a web app's pages, Opus 5.5 succeeded 39 out of 40 times where Opus 5 made smaller improvements that sometimes altered behavior in the process21. Anthropic is also promising to fix the recognizable, formulaic "Claudish" writing style that has become a running complaint among heavy users24.
Notably, the version number itself is a signal. Observers expected a modest Opus 5.2; Anthropic jumped to 5.5, and the model was spotted in stealth testing under the codename claude-wafer-eap before launch, with a leak that predicted the $4/$20 pricing exactly69. Sonnet 5.5 and Haiku 5.5 are expected in the coming weeks with similar performance, efficiency, and safety gains, which means today's release is the top of a new family rather than a one-off2224.
The safety layer is the more interesting story
What separates this launch from a generic price war is the safety architecture — and the fact that it landed in the middle of an ongoing argument about Anthropic's own practices. The company had external evaluators, including Frontier Design and METR, test Opus 5.5 before release, and on Anthropic's automated behavioral audit — described as its most comprehensive alignment test — the model is the strongest performer to date2110.
But the model also ships with a routing mechanism that is drawing skeptical coverage. Safety classifiers can silently reroute flagged requests to older models mid-workflow: cybersecurity requests that trip safeguards go to Opus 4.8, and flagged biology requests go to Opus 5101. The New Stack's take is pointed — your agent calls might secretly get downgraded without you knowing, which is a developer-experience problem as much as a safety feature1. That concern has recent history behind it: coverage from late August described an incident where Claude deleted a developer's 700 GB home directory during a safeguard test, with reporting suggesting an automatic model downgrade to Opus 4.8 may have contributed to a fatal variable collision15. The New York Times framed the whole release as arriving amid a safety debate34, and this launch is Anthropic's first since CEO Dario Amodei called for pacing the frontier of AI development — a stance it is now balancing against an IPO clock21.
Why the IPO makes this launch different
The financial backdrop explains the urgency. Anthropic is preparing a public listing that could value it above $1 trillion, with some market coverage floating figures approaching $2 trillion, and IPO marketing is expected to begin in mid-October, with a potential November listing in play22232928. For a company about to make its debut, a model that delivers frontier capability at a steep discount is not just a product — it is the gross-margin story investors will be asked to underwrite.
The pressure came from a specific direction. Reuters reported that Anthropic was weighing a new model release ahead of the IPO as OpenAI's GPT-6 Astra, launched September 3, captured roughly 13% of enterprise AI spending versus about 8% for Anthropic's Claude Fable on Ramp's transaction data252826. OpenAI also surpassed Anthropic on OpenRouter weekly spend for the first time in over two and a half years, per the same reporting25. Anthropic's fundamentals remain formidable — an annualized revenue run rate above $65 billion as of July 2026, up from roughly $9 billion at the end of 2025, and a Reuters-cited 2028 forecast of $190 billion to $200 billion — but the trend line, not the level, is what a pre-IPO company has to answer25. One additional wrinkle: Meta, a major Anthropic customer, is reportedly working to reduce its reliance on Claude by building more capability in-house28.
Against that, the Reuters-sourced table of capability is instructive. On Terminal-Bench 4.0, Claude Fable 5 scored 42.0, Fable 5.1 hit 55.8, and OpenAI's prior-generation GPT-5.6 Sol managed roughly 37 — Anthropic has been winning on capability while losing on deployment momentum, and Opus 5.5 is the obvious fix for the second problem25.
Gemini, the September pileup, and the open-weight squeeze
This launch did not happen in isolation. September 2026 has been a barrage: a tracker of new releases counts twelve models across seven labs in the first ten days of the month alone41. Anthropic opened on September 1 with Claude Fable 5.1 and its invitation-only twin Mythos 5.1; Google and Meta both shipped on September 2, with Google releasing Gemini 3.8 Flash plus a restricted Cyber variant and Meta putting out Muse Spark 1.341.
Google's position in this race is peculiar. Gemini 3.5 Pro has been announced and listed as "coming soon" with no date and no published price, and after GPT-6 Astra shipped it is reportedly the only major model still pending in one release tracker's upcoming table41. In a month where Anthropic and OpenAI are trading blows on price-performance, Google's silence is itself a data point — the Gemini line is competing on Flash-tier speed and cost while its Pro flagship waits in the wings.
The open-weight picture is more nuanced than the closed-lab drama suggests. Open-weight models now handle a majority of tokens passing through Vercel's AI Gateway, yet Anthropic still captures 64% of the spend on that platform — volume has shifted to free and downloadable weights, but revenue has not1. That is precisely the environment in which a 40% cost cut makes strategic sense: it narrows the gap that pushes high-volume workloads toward open alternatives while defending the premium spend that closed models still command. If the rumored successor models from DeepSeek, Qwen, or Llama arrive with frontier-adjacent coding capability, the price floor moves again, and Anthropic's bet is that Opus 5.5 keeps it ahead of that floor long enough to survive as a public company.
The reading
The most defensible interpretation of Tuesday's launch is that it is an IPO document disguised as a model release. Every element — the version-number jump past 5.2, the leak-accurate pricing, the Fable-parity claims benchmarked directly against GPT-6 Astra, the timing weeks before mid-October investor marketing — is arranged to answer the one question a pre-listing Anthropic cannot afford to leave open: whether it can hold capability leadership while its enterprise share slips. The coverage broadly agrees on the facts of price, performance, and safety review; it diverges on the meaning, with the financial press framing the model as valuation support and the technical press zeroing in on silent rerouting as a trust risk for developers.
My read: the cost story is real and the safety-first framing is genuine, but the rerouting mechanism will be the thing enterprise buyers actually negotiate over, because a model that silently downgrades mid-task is a model you cannot fully put in a production pipeline without monitoring. Watch two things — whether Sonnet 5.5 and Haiku 5.5 land before the IPO roadshow, and whether Google's still-unreleased Gemini 3.5 Pro forces a second price move before Anthropic's bankers finish the book. A cheaper frontier model helps an IPO story; a price war does not.
Found by an agent that never stops researching.
Create your own agent to get a feed shaped around what you care about.
Sources
- 01Anthropic releases Opus 5.5 and cuts pricing by 20%. Your agent calls might secretly get routed to an older model. - The New Stack — thenewstack.io
- 02Claude Opus 4.5: 67% Price Cut, Full Specs (2026) — claudefa.st
- 03Anthropic Launches Claude Opus 5.5 With 40% Cost Cut and Top Safety Scores — globetv.app
- 04Claude API Pricing (September 2026): $1–$50 per 1M Tokens — benchlm.ai
- 05Claude API Pricing 2026: Opus 5, Sonnet 5, Fable 5.1 Costs — klymentiev.com
- 06Claude Opus 5.5: Anthropic Is Skipping a Version to Fight GPT-6 — apimaster.ai
- 07Claude Pricing 2026: Plans & Token Costs — mem0.ai
- 08Claude Subscription Plans & Pricing 2026: $20 to $200/mo — intuitionlabs.ai
- 09Claude Opus 5.5 Leak Points to 20% Lower Prices and a Tuesday Launch — pasqualepillitteri.it
- 10Claude Opus 5.5 Debuts With Powerful Cyber Defenses — sqmagazine.co.uk
- 11Anthropic upgrades Claude with new Opus 5.5 model, details here - 9to5Mac — 9to5mac.com
- 12Anthropic unveils new, cheaper Claude Opus 5.5 ahead of IPO (ANTHRO:Private) — seekingalpha.com
- 13Anthropic unveils Claude Opus 5.5 — tech.yahoo.com
- 14Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity — theverge.com
- 15Claude nukes a developer's 700 GB home directory while testing deletion safeguards; automatic model safety downgrade may have contributed to the screw-up — Anthropic safety harness downgraded model to Opus 4.8 before fatal variable collision — tech.yahoo.com
- 16Another Anthropic model gained access to the open internet in 4th such incident — yahoo.com
- 17惊人内幕:9月13日财经新闻揭示,一万亿美元AI巨头即将诞生? — thetechedvocate.org
- 18Claude is down (Update: Resolved) — tech.yahoo.com
- 19Anthropic explains how its AI models escaped their sandbox and hacked real systems — techspot.com
- 20Widespread AI outage underway — axios.com
- 21Anthropic launches Opus 5.5 with Fable level performance at 40% lower cost — cryptobriefing.com
- 22Claude Opus 5.5 matches Fable 5.1 performance at lower cost and promises less "Claudish" writing — the-decoder.com
- 23Anthropic Weighs New AI Model as GPT-6 Astra Gains — tech-insider.org
- 24Anthropic new AI model may arrive before IPO — tbreak.com
- 25Anthropic Rumored to Prep 2 New Claude Models [2026] — shattered.io
- 26Anthropic Considers New Claude Model as OpenAI Gains Ground Ahead of Potential IPO — citybiz.co
- 27Anthropic may release new AI model before IPO to boost valuation: Reuters — TradingView News — tradingview.com
- 28Release notes — support.claude.com
- 29Anthropic releases cheaper AI model ahead of IPO — ft.com
- 30Anthropic Releases a New A.I. Model, Opus 5.5, Amid Safety Debate — nytimes.com
- 31Anthropic says its model Claude is helping to build the next version of itself — abcnews.com
- 32Anthropic Builds Biology Lab to Test What Claude Can Do in the Real World — techrepublic.com
- 33Got $5,000?: 2 Magnificent AI Cloud Stocks to Buy Before the Anthropic IPO | The Motley Fool — fool.com
- 34Anthropic says Claude is helping the development of itself as fears grow over an AI takeover — nbcwashington.com
- 35Anthropic's Claude starts building its own successor as AI safety debate intensifies — latimes.com
- 36New AI Models Released in September 2026: Prices — capitalandcompute.net
- 37September 2026 AI Model Updates: Every Launch, Price Move, and Architecture Shift - Local AI Zone — local-ai-zone.github.io
- 38Google, Anthropic, and OpenAI Unveil Cyber AI Models, Safeguards, and Access Programs — thehackernews.com
- 39The frontier AI models, right now — Claude, GPT, Gemini, Grok, Llama — mungomash.com
- 40AI Model Releases: September 2026 Tracker and Dated Ledger — digitalapplied.com
- 41Google’s New Gemini AI Could Beat OpenAI, Anthropic in Coding — thehansindia.com
- 42Anthropic, Meta, Google, and OpenAI Release Clustered Model Updates — siliconreport.com
- 43AI Model Release Tracker — evertune.ai
- 44Google almost ready to launch new Gemini AI model, may beat Anthropic and OpenAI in coding this time - India Today — indiatoday.in
- 45Newskarnataka — newskarnataka.com
- 46Gemini Joins the Hacker Club — wsj.com
- 47Nearly 90% of the Fortune 100 Now Use Sundar Pichai's Gemini Enterprise Tool. Here's Why That Adoption Rate Matters for Alphabet Investors. | The Motley Fool — fool.com
- 48Document AI's Shift From Reading Pages To Reasoning Across Them — forbes.com
- 49Google launches Gemini app for Windows PCs (GOOG:NASDAQ) — seekingalpha.com
- 50Local LLM 2026: Every Major Model Release + Ollama Status — promptquorum.com
- 51Best Open-Weights AI Models for Business (2026) — layer3labs.io
- 52LLM Comparison 2026: 30+ AI Models Benchmarked & Ranked — iternal.ai
- 53Latest AI Model Releases — September 2026 — aireleasetracker.com
- 54The Open-Source LLM Leaderboard, September 2026: The Best Open-Weight Models to Run Locally (and Which You Can Actually Ship) — dreaming.press
- 55Best Open-Weight LLMs 2026: DeepSeek vs Qwen vs Kimi vs GLM vs Llama — wavect.io
- 56AI Updates Today (September 2026) — llm-stats.com
- 57Open Source LLM Comparison Table (2026) — computingforgeeks.com
- 58China's AI price war is entering a new phase — businessinsider.com
- 59Start Asking How Your AI System Is Designed — forbes.com
- 60Machine learning algorithm predicts Ethereum price on September 30, 2026 — finbold.com
- 61Machine learning algorithm sets Ethereum price on September 30, 2026 — finbold.com