PulseAugur
实时 01:10:40
English(EN) Sakana AI Releases Fugu-Cyber: An Orchestration Model Reporting 86.9% on CyberGym and 72.1% on CTI-REALM

Sakana AI 推出 Fugu-Cyber 用于网络安全任务 · 已追踪 2 个来源

Sakana AI 推出了 Fugu-Cyber,一个专门针对网络安全任务进行调优的编排模型。该模型在 CyberGym 基准测试中取得了 86.9% 的成绩,在 CTI-REALM 基准测试中取得了 72.1% 的成绩,使其成为与 GPT-5.5-CyberClaude Mythos Preview 等其他专业模型竞争的有力选项。Fugu-Cyber 在 Sakana 的 Fugu orchestrator 中运行,该 orchestrator 将任务委托给专业模型,并特别强调以安全为重点的验证步骤。 AI

影响 这个专门的网络安全模型设定了新的基准,可能会影响安全任务 AI 代理的开发。

排序理由 Sakana AI 是一个前沿实验室,发布了一个具有基准测试结果的新模型。[lever_c_demoted from frontier_release: ic=2 ai=1.0]

在 MarkTechPost 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Sakana AI 推出 Fugu-Cyber 用于网络安全任务 · 已追踪 2 个来源

报道来源 [2]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Sakana AI 发布 Fugu-Cyber:一个在 CyberGym 上报告 86.9%、在 CTI-REALM 上报告 72.1% 的编排模型

    <p>Sakana AI has released Fugu-Cyber, a security-tuned endpoint on its Fugu orchestration model. It reports 86.9% on CyberGym and 72.1% on CTI-REALM, edging past GPT-5.5-Cyber and Claude Mythos Preview. Access is gated behind manual approval, a defensive-use policy, and the Token…

  2. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Sakana AI 发布 Fugu-Cyber,一款专注于网络安全的编排模型,在 CyberGym 上达到 86.9%,在 CTI-REALM 上达到 72.1%,超越 GP

    Sakana AI has released Fugu-Cyber, a cybersecurity-focused orchestration model that achieves 86.9% on CyberGym and 72.1% on CTI-REALM benchmarks, edging past GPT-5.5-Cyber and Claude Mythos Preview. Access is gated behind manual approval. https://www. marktechpost.com/2026/07/25/…