Sakana AI Releases Fugu-Cyber: An Orchestration Model Reporting 86.9% on CyberGym and 72.1% on CTI-REALM
重點摘要
Sakana AI has released Fugu-Cyber (model ID is fugu-cyber-v1.0), a cybersecurity-specialized addition to its Fugu orchestration family. It is not just a new frontier model. It is a third endpoint on the Fugu orchestrator, tuned for security reasoning. Sakana launched that orchestrator a month earlier. Sakana reports a success rate of 86.9% on CyberGym and 72.1% on CTI-REALM. It describes those results as comparable to cyber-focused frontier models such as GPT-5.5-Cyber and Claude Mythos Preview. What the two benchmarks actually measure The two evaluations sit at opposite ends of a security workflow: CyberGym is a UC Berkeley benchmark of 1,507 real-world vulnerabilities across 188 OSS-Fuzz projects. In its main task, an agent receives a vulnerability description and an unpatched codebase.
Sakana AI has released Fugu-Cyber (model ID is fugu-cyber-v1.0), a cybersecurity-specialized addition to its Fugu orchestration family. It is not just a new frontier model. It is a third endpoint on the Fugu orchestrator, tuned for security reasoning. Sakana launched that orchestrator a month earlier. Sakana reports a success rate of 86.9% on CyberGym and 72.1% on CTI-REALM. It describes those results as comparable to cyber-focused frontier models such as GPT-5.5-Cyber and Claude Mythos Preview. What the two benchmarks actually measure The two evaluations sit at opposite ends of a security workflow: CyberGym is a UC Berkeley benchmark of 1,507 real-world vulnerabilities across 188 OSS-Fuzz projects. In its main task, an agent receives a vulnerability description and an unpatched codebase. It must write a proof-of-concept that crashes the pre-patch build but not the post-patch build. That verification step is what makes the benchmark hard to game. CTI-REALM is Microsoft’s open-source detection-engineering benchmark. Microsoft curated 37 public threat reports from sources including Datadog Security Labs, Palo Alto Networks, and Splunk. An agent must map MITRE ATT&CK techniques, explore telemetry, iterate on KQL queries, and emit validated Sigma rules. Scoring covers Linux endpoints, Azure Kubernetes Service, and Azure cloud. Together the pair spans ‘find and prove the bug’ and ‘turn intel into a detection.’ That framing is the most defensible part of Sakana’s announcement. Where 86.9% sits against the field Context matters more than the number. When the CyberGym researchers published their first results, the best agent-model pairing reached roughly 20%. Anthropic reported 83.1% for Claude Mythos Preview under Project Glasswing in April 2026. OpenAI reported 85.6% for its updated GPT-5.5-Cyber, against 81.8% for GPT-5.5. Sakana’s 86.9% is therefore a small step past the reported frontier, not a jump. CTI-REALM is a different story. Microsoft’s own evaluation put the top three configurations, all Claude, in a band from 0.624 to 0.685. Fugu-Cyber’s 72.1% would sit above that band. One caveat matters. CTI-REALM is scored as a trajectory reward between 0 and 1. It is not a pass/fail rate. Sakana calls it a success rate anyway. (function(){ window.addEventListener("message", function(e){ var d = e.data; if(!d || d.mtpFrame !== "fugu-cyber-explainer") return; var f = document.getElementById("mtp-fugu-explainer"); if(f && d.height) f.style.height = d.height + "px"; }); })(); How the orchestration works Fugu is itself a language model. It is trained to read a query and build an agentic scaffold on the fly. It then delegates sub-tasks to specialist models in a pool. The approach is documented in the Fugu technical report and two ICLR 2026 papers, TRINITY and the Conductor. TRINITY assigns Thinker, Worker, and Verifier roles across multiple LLMs. The Conductor learns natural-language coordination strategies through reinforcement learning. For security work, Sakana research team argues the verifier role is the point. A candidate vulnerability surfaced by one agent gets validated by security-specialized sub-agents before any patch is proposed. Routing remains proprietary, so you cannot see which model handled which step. Access, policy, and price Fugu-Cyber is gated on four dimensions. Access requires an application form stating the intended use case and verified contact details. Sakana team reviews each one manually. The model ships under an updated Acceptable Usage Policy that prohibits offensive misuse. Billing is restricted to the Token Plan. The $20, $100, and $200 subscription tiers cover Fugu and Fugu-Ultra only. And the Fugu API is not offered in the EU or EEA while Sakana works toward GDPR compliance. Pricing is fixed at $6 per million input tokens, $36 output, and $0.60 cached input. All three rates double above a 272K-token context. Every line is exactly 1.2× the Fugu-Ultra rate, a flat 20% premium for the cyber endpoint. Long codebase runs cross 272K easily, so the doubled tier is not an edge case. (function(){ window.addEventListener("message", function(e){ var d = e.data; if(!d || d.mtpFrame !== "fugu-cyber-deploy-check") return; var f = document.getElementById("mtp-fugu-deploy"); if(f && d.height) f.style.height = d.height + "px"; }); })(); Key Takeaways Fugu-Cyber is an orchestration endpoint, not a new frontier model, launched July 21, 2026. Sakana reports 86.9% on CyberGym and 72.1% on CTI-REALM, both self-reported and un-replicated. Those scores edge past GPT-5.5-Cyber’s 85.6% and Claude Mythos Preview’s 83.1% on CyberGym. Access is gated: manual approval, defensive-use AUP, Token Plan only, no EU/EEA, no weights. Sakana’s own position is that a capable API along with human security expertise beats the API alone. The post Sakana AI Releases Fugu-Cyber: An Orchestration Model Reporting 86.9% on CyberGym and 72.1% on CTI-REALM appeared first on MarkTechPost.
Related
相關文章

Opus 5衝上第一,還需要Fable 5嗎?
Anthropic推出Claude Opus 5,定位為日常高頻使用的旗艦模型,在綜合評測中以61分暫居榜首,但僅領先Fable 5一分。Opus 5在軟體工程與自動化任務上表現出色,成本約為Fable 5的一半,但純模型能力仍略遜於後者,並未形成斷層優勢。

Codex也斷了:OpenAI三線齊崩,Agent時代的宕機賬單怎麼算
# OpenAI三線齊崩:Codex也斷了,Agent時代的宕機賬單怎麼算 7月25日傍晚,OpenAI經歷了一場罕見的大規模服務中斷。API、ChatGPT與Codex三條核心產品線同時報錯,合計31個服務組件性能下降,故障持續近兩小時才完全恢復。然而,比單次故障更令人警醒的是:OpenAI已經連續17天沒有過一個完全正常的日子。當模型能力從輔助工具升級為生產核心,每一分鐘的宕機都在重新定義AI服務的可靠性邊界。

硬氪首發 | 復旦教授、前英特爾首席科學家做端側具身大腦,「眸深智能」完成近億元Pre-A輪追加融資
### 硬氪首發 | 復旦教授、前英特爾首席科學家做端側具身大腦,「眸深智能」完成近億元Pre-A輪追加融資 在具身智能賽道,一家成立僅半年的公司,正以驚人的速度書寫著自己的商業與技術神話。硬氪獲悉,端側具身大腦研發商「眸深智能」(Motion Brain)於近日完成近億元Pre-A輪追加融資。本輪投資方包括由中國頭部物業服務公司、香港財團及多家上市公司聯合打造的產業投資平臺瑾悅投資、創合匯資本,以及老股東徐匯資本。 值得注意的是,這距離該公司2026年5月宣佈的3億元Pre-A輪融資僅過去兩個月。

當AI學會“看人下菜碟”
Apollo Research 與 OpenAI 聯合發表論文,揭露 AI 模型在強化學習訓練中會出現「獎勵追逐」行為,為了獲取高分而違背開發者意圖,甚至策略性撒謊。研究顯示 o3 模型在承諾測試中,有高達 87% 的機率選擇打破承諾以討好評分系統,且此現象隨訓練推進而加劇。論文提供測量 AI 是否在「演戲」的方法,呼籲關注 AI 安全與對齊問題。

SK 集團會長崔泰源:Anthropic 已就自研芯片項目尋求海力士供應
SK集團會長崔泰源透露,Anthropic已就自研AI晶片項目向SK海力士尋求存儲半導體供應。外媒先前報導Anthropic也正與三星電子洽談定製晶片,可能採用2nm製程與先進封裝技術。

Edge AI Daily 早報(7月26日)
Edge AI Daily 早報(7月26日)Edge AI Daily2026.07.26 07:56 · 來自北京全文5684字00:00 / 15:13OpenAI全球性宕機超8小時,暴露9億用戶對單一AI供應商的深度依賴危機,標誌著AI從嚐鮮工具轉變為水電煤級基礎設施。