Anthropic 發布 Claude Opus 5.5:效能達 Fable 5.1 等級,運行成本較 Opus 5 低 40%

2026年9月22日 18:59
站內 AI 整理稿

Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family.The team states it performs at the level of Claude Fable 5.1 on most work.It also costs 40% less to run than Opus 5 on typical workloads at default settings.

On Anthropic’s own benchmarks, it leads in agentic coding, computer use, and knowledge work.Is it deployable?Yes, as a managed API model.Anthropic has not released weights, so self-hosting is not an option.

Developers can call claude-opus-5-5 on the Claude Platform, Amazon Web Services, Google Cloud, and Microsoft Azure.Zero data retention is available, as with previous Opus models.Benchmarks: Strong Lead, Not a Clean Sweep Opus 5.

5 scores use adaptive thinking at max effort, with production safeguards enabled.BenchmarkOpus 5.5Fable 5.1Opus 5GPT-6 AstraTerminal-Bench 4.066.4%55.8%52.3%57.9%FrontierCode v1.154.4%50.3%48.0%53.3%CursorBench 4.057.8%51.8%46.6%n/rGDPval-AA v2.1 (Elo)1846173517081542OSWorld 2.081.8%80.7%74.

0%n/rTerminal-Bench-Science 0.158.7%52.6%29.0%64.6%AutomationBench40.0%31.4%26.9%41.4% Terminal-Bench 4.0 is reported at xhigh effort for Opus 5.5.GPT-6 Astra still leads on Terminal-Bench-Science and AutomationBench.

Zapier ran AutomationBench without fallback models, so safeguard interventions counted as failures.Anthropic also cautions that benchmark margins are becoming a less reliable guide.In its own use, the gap to Fable 5.1 is narrower than the scores suggest.The cost-adjusted results are more telling.

At default (medium) effort, Opus 5.5 scores 54.6% on FrontierCode.That beats GPT-6 Astra’s top score of 53.3% at about a fifth of the cost per task.On CursorBench, medium effort scores 52.5%.That is 11 points above GPT-5.6 Sol’s best, at about a third of the cost.Pricing and Speed Opus 5.

5 needs less compute to serve than Opus 5, and pricing reflects that.Per 1M tokensOpus 5.5Opus 5Input$4$5Output$20$25Cache reads$0.20$0.50Cache writes$5$6.25 Cache reads make up most agentic and coding costs, and they drop 60%.Opus 5.5 also uses fewer tokens per task.

Together, that nets out to the 40% cost reduction.Output generation is more than 30% faster than Opus 5.Fast mode in Claude Code and the Claude Platform offers up to 2.5x speed at $8 input and $40 output per million tokens.

Anthropic is also raising five-hour usage limits on Pro, Max, Team, and seat-based Enterprise plans.Subscribers get a rate limit reset they can save and use later.What Early Testers Reported One tester completed a 680,000-line code migration in less than a day.

Another audited and fixed a 200,000-line codebase in under 3 hours.Opus 5 took over 20 hours and 2.5x the tokens.In an internal C to Rust port of HAProxy, Opus 5.5 finished in 9.5 hours.Fable 5.1 took 12 hours, and Opus 5.5 cost 51% less.Deloitte says Opus 5.

5 at lowest effort caught 72% of known review bugs.Opus 5 at high effort caught 56%.In a hard-to-source earnings report test, 16 of 18 Opus 5.5 reports cleared Anthropic’s quality bar.Fable 5.1 and Opus 5 never did.Writing style also changed.Opus 5.

5 puts key information first, uses less jargon, and follows the writing rules you give it.Safety, Safeguards, and API Changes Opus 5.5 is Anthropic’s first release since CEO Dario Amodei called for pacing the frontier.External evaluators including METR and Frontier Design tested it before release.

It posts the best score to date on Anthropic’s automated behavioral audit, which covers nearly 2,000 scenarios.In a new containment test, it tried to circumvent boundaries about 85% less often than Opus 5.Anthropic also notes the model often suspects it is being evaluated.

Its biology and cyber capabilities are comparable to Claude Mythos 5.1.So Opus 5.5 ships with safeguards similar to Fable 5.1: Cybersecurity: Routine bug finding and fixing works.Most other cybersecurity tasks are re-routed to Opus 4.8.The Cyber Verification Program will expand to Opus 5.5.

Biology: Vetted organizations can apply to the Life Sciences Verification Program.Distillation: Preserved thinking stops API users from editing prior context to extract reasoning.It applies to API accounts created on or after August 31, 2026.Two more changes affect integrations.

Thinking can no longer be disabled.Outputs also carry watermarking for EU AI Act compliance.Full details are in the Opus 5.5 System Card.Interactive Explainer (function(){window.addEventListener("message",function(e){var d=e.data;if(!d||d.mtpFrame!=="mtp-o55"||!d.h)return;var f=document.

getElementById("mtp-o55-frame");if(f)f.style.height=d.h+"px";});})(); Key Takeaways Opus 5.5 matches Fable 5.1 on most work and beats both Opus 5 and Fable 5.1 on nearly every reported benchmark.API pricing drops to $4/$20 per 1M tokens, and cache reads fall 60% to $0.20.

Anthropic puts typical workload savings at 40%, with output over 30% faster than Opus 5.Cyber and biology requests hit Fable 5.1-class safeguards, and thinking cannot be switched off.Closed weights: claude-opus-5-5 runs via Claude Platform, AWS, Google Cloud, and Azure.

Check out the Technical details here.All credit goes to the researcher of this project.Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?Connect with us The post Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Than Opus 5 appeared first on MarkTechPost.

Related

相關文章

​韶音首款 AI 耳機 OpenFit 2 AI 亮相雲棲大會,重磅首發“AI 實驗室”四大新功能

作為開放式耳機領域的領頭羊,Shokz 韶音首次重磅參展,不僅攜旗下首款 AI 耳機 OpenFit2AI 核心亮相,更在現場聯合光帆科技首發了“AI 實驗室”全新功能,全面展現了 Qwen 大模型在音頻可穿戴設備中的深度產業落地。作為韶音佈局智能音頻賽道的開山之作,OpenFit2AI 深度融合了通義千問(Qwen)大模型的強大能力,集長時錄音、AI 會議紀要生成與多語種實時翻譯於一體。

剛剛
鈦媒體生成式AI

GPT-6 Sol和Luna上線,打折比梁文鋒還狠

字母AI2026.09.23 11:08 · 來自北京全文5453字00:00 / 14:27Opus 5.5發佈90分鐘,OpenAI來打價格戰了。文 | 字母AI今夜註定無眠。Opus 5.5剛剛發佈90分鐘,OpenAI就把GPT-6 Sol和Luna一起端了上來。前腳Anthropic剛把Opus的價格砍了一刀,後腳OpenAI直接把刀砍得更深。GPT-6 Sol每百萬Token輸入只要2美元、輸出10美元,價格正好只有Opus 5.5的一半,Astra的五分之一。Luna更誇張,只要0.1美元和0.

剛剛

Claude Code v2.1. 280 落地:Opus 5. 5 正式接棒默認王座,百萬上下文配 4 美元骨折價把賬單壓進地心

最新推送的Claude Code v2.1.280版本里,Claude Opus5.5(claude-opus-5-5)正式被任命為默認Opus模型,這也讓前陣子開發者圈裡流傳的"跳級突襲"傳聞徹底塵埃落定。這枚新旗艦一口氣吞下100萬Token的上下文窗口,定價牌掛得異常兇悍:每百萬Token輸入4美元、輸出20美元,緩存讀取更是壓到每百萬Token0.

剛剛
鈦媒體生成式AI

Meta Muse 爆紅,為什麼騰訊股價大漲?

深流研究所2026.09.23 10:39 · 來自北京全文3733字00:00 / 10:34扎克伯格等來爆款,騰訊跟著被重估。文 | 深流研究所,作者 | 吳絳楓Meta 在 AI 上砸下的重金,什麼時候才能拿到大結果?Muse 給出了第一個答案。

剛剛
鈦媒體生成式AI

Opus 5.5來了,性能趕超Fable,成本低至40%

字母AI2026.09.23 10:31 · 來自北京全文4730字00:00 / 12:20被Astra“逼出來”的全新系列。文 | 字母AI5.5!!是Opus 5.5!!就在三天前,路透社還報道稱,隨著GPT-6 Astra開始搶回企業市場,Anthropic正在考慮提前推出一款新模型應戰。社區一直猜測新模型會是Opus 5.2還是5.5,現在答案終於揭曉了。而且這不是一個小更新,Anthropic直接掀開了Claude 5.5這個新系列。Opus 5.5先打頭陣,Sonnet 5.5和Haiku 5.

剛剛