PulseAugur
中
实时 16:30:02
实体 Sakana AI

Sakana AI

PulseAugur coverage of Sakana AI — every cluster mentioning Sakana AI across labs, papers, and developer communities, ranked by signal.

Show in brief
总计 · 30天
57
90 天内 58
发布 · 30天
0
90 天内 0
论文 · 30天
10
90 天内 10
层级分布 · 90 天
主题
关系
时间线
  1. 2026-09-16 product_launch Sakana AI launched two new commercial models, Fugu Max and Fugu Ultra v2, both with a 1 million token context window. 来源
  2. 2026-09-11 product_launch Sakana AI launched Fugu Max and Fugu Ultra v2, new models for multi-agent orchestration. 来源
  3. 2026-09-11 product_launch Sakana AI launched Fugu Max and Fugu Ultra v2, new models for multi-agent orchestration. 来源
  4. 2026-09-11 product_launch Sakana AI launched Fugu Max and Fugu Ultra v2, new models for multi-agent orchestration. 来源
  5. 2026-09-11 product_launch Sakana AI released Fugu Max and Fugu Ultra v2, updates to their multi-agent orchestration system. 来源
  6. 2026-07-16 product_launch Sakana AI integrated Nvidia's Nemotron models into its Fugu orchestrator. 来源
  7. 2026-07-16 product_launch Sakana AI integrated Nvidia's Nemotron models into its Fugu orchestrator. 来源
  8. 2026-06-27 product_launch Sakana AI launched its Fugu orchestrator, aiming to match Western AI models. 来源
  9. 2026-06-27 product_launch Sakana AI launched its Fugu AI model, positioning it as an alternative to U.S. models facing export controls. 来源
  10. 2026-06-15 product_launch Sakana AI launched its first commercial product, Sakana Marlin, an enterprise agent for autonomous research report generation. 来源
情绪 · 30 天

5 天有情绪数据

LAB BRAIN
hypothesis expired 置信度 0.55

Sakana AI to release open-source version of PC-ALM training method

Sakana AI's recent proposal of the layer-local training method PC-ALM for deep networks suggests a potential shift towards more accessible deep learning research. Given their release of Fugu models, it's plausible they will open-source PC-ALM to foster community development and adoption, similar to other foundational AI research releases.

observation expired 置信度 0.70

Sakana AI's Fugu Ultra v2 shows strong performance against top-tier models

The release of Fugu Ultra v2, which reportedly surpasses Opus 5 and Fable 5 in visual reasoning and outperforms GPT-6 Astra and Fable 5.1 on coding benchmarks, indicates Sakana AI is achieving competitive performance with leading proprietary models. This suggests their multi-agent orchestration models are maturing rapidly.

hypothesis expired 置信度 0.60

Sakana AI to integrate PC-ALM into future Fugu model releases

Sakana AI's development of PC-ALM, a method for training very deep networks, could be a foundational technology for their future commercial models. It's likely they will aim to integrate this efficient training method into upcoming Fugu releases to improve performance and potentially reduce training costs.

hypothesis expired 置信度 0.55

Sakana AI to offer Marlin as a managed service or API within 90 days

Sakana AI has launched Marlin as an enterprise agent. Given its capabilities in generating detailed reports and presentation slides, it's plausible they will soon offer it as a managed service or an API to a wider customer base beyond direct enterprise deployment, enabling broader adoption.

observation expired 置信度 0.70

Sakana AI is focusing on specialized AI models and enterprise solutions

The recent launches of Sakana Fugu and Sakana Marlin, with Marlin specifically targeting enterprise research, indicate a strategic focus. Marlin's ability to generate extensive reports autonomously highlights a move towards practical, high-value applications of AI for businesses.

查看全部假设 →

最近 · 第 1/4 页 · 共 64 条
  1. COMMENTARY · CL_274935 ·

    MIT博士候选人探索递归语言模型和智能体集群

    MIT的博士候选人Alex Zhang正在探索递归语言模型(RLM)和先进的智能体系统在释放更强大AI能力方面的潜力。他的研究在Latent Space播客中进行了详细介绍,深入探讨了AI生成的GPU内核、通过组合器实现的组合泛化,以及“语言模型”演变成一个看不见的智能体集群的概念。Zhang的研究还涉及OpenAI的大规模智能体实验,并比较了多智能体系统的不同方法,表明当前的尖端模型可能已经拥有巨大的未开发潜力。

  2. SIGNIFICANT · CL_256464 ·

    Sakana AI 发布 Fugu Max 和 Fugu Ultra v2,支持 100 万 token 上下文

    Sakana AI 发布了两款新的商业模型 Fugu Max 和 Fugu Ultra v2,均支持 100 万 token 的上下文窗口。Fugu Max 的定价为每 100 万 token 2 美元(输入)/ 6 美元(输出),旨在降低长上下文实验的成本。Fugu Ultra v2 的定价更高,为每 100 万 token 5 美元(输入)/ 30 美元(输出)。

  3. TOOL · CL_253629 ·

    Sakana AI 提出层局部训练方法以训练 1000 层网络

    Sakana AI 的研究人员开发了一种名为增强拉格朗日预测编码 (PC-ALM) 的新颖训练方法,它提供了一种传统的反向传播的层局部替代方法。这种新方法允许训练非常深(多达 1000 层)的神经网络,同时保持接近反向传播的性能。该方法已在 MNIST 和 Fashion-MNIST 等小型图像基准测试中得到验证,与标准的预测编码相比,在深层窄网络架构中显示出显著的改进。

  4. RESEARCH · CL_251911 ·

    Sakana AI 发布 Fugu Ultra v2;Cursor 推出 AI 代理管理工具

    Sakana AI 发布了 Fugu Ultra v2 模型,其视觉推理能力超越了 Opus 5 和 Fable 5。另外,Cursor 推出了名为“Projects”的新功能,允许用户管理数千个 AI 代理。

  5. SIGNIFICANT · CL_249286 ·

    Anthropic 寻求 2 万亿美元 IPO,AI 实验室推动放缓发展

    据报道,Anthropic 正在为 10 月份进行的大规模 2 万亿美元 IPO 做准备,这得益于强劲的收入增长和高利润率,英伟达(Nvidia)正考虑进行重大投资。与此同时,微软(Microsoft)继 Anthropic 和 OpenAI 之后,也主张放慢人工智能(AI)的发展速度,以解决对齐和安全问题,并为其 AI 模型引入了行为准则。AI 行业也日益出现分歧,数学家们对 AI 实验室在未经充分验证的情况下急于声称取得数学突破表…

  6. FRONTIER RELEASE · CL_247570 ·

    Sakana AI 发布 Fugu Max 和 Ultra v2,用于经济高效且高能力的多智能体任务 · 跟踪 4 个来源

    Sakana AI 推出了 Fugu Max 和 Fugu Ultra v2,这是其 Sakana Fugu 系列中用于多智能体编排的两个新模型。Fugu Max 针对成本效益进行了优化,其定价低于 Sonnet 5 和 GPT 5.6 Terra 等竞争对手,同时实现了强大的基准性能。另一方面,Fugu Ultra v2 专注于复杂、多步骤任务的高能力,据报道,在特定编码基准测试中,它在不依赖 OpenAI 的 GPT-6 Astr…

  7. TOOL · CL_253454 ·

    Sakana AI 推出 Fugu Max 和 Ultra v2 以实现高效的多智能体协调

    Sakana AI 发布了 Fugu Max 和 Fugu Ultra v2,这是一个更新的多智能体协调系统,旨在提高 AI 的性能和成本效益。Fugu Max 旨在通过协调大量开放和专业化的模型,以更低的 token 成本提供前沿级别的结果。Fugu Ultra v2 专注于在不依赖最先进模型的情况下,实现复杂、多步骤任务的峰值性能。该系统强调集体智能和可替换智能体池的好处,以对冲 AI 基础设施中的供应商限制和权力集中。

  8. RESEARCH · CL_231315 ·

    新框架ARISE-RL赋能AI智能体进行迭代式自我进化

    研究人员推出ARISE-RL,一个旨在增强强化学习训练的智能体自我进化能力的新型框架。该框架通过耦合任务/评分标准生成器和推理求解器来应对开放式任务中的挑战,从而实现迭代式改进。它还纳入了奖励门控自我进化蒸馏(RG-SED)来优化策略学习,并引入了ECR-Bench,一个用于评估智能体在复杂研究和规划任务上性能的新基准套件。

  9. TOOL · CL_208773 ·

    Yuanluo Technology deploys humanoid robots in national labs for AI-driven science

    Yuanluo Technology has introduced Monte2, a humanoid robot designed to automate laboratory tasks in national-level research platforms. These robots can perform complex procedures like reagent handling, cell toxicity tes…

  10. TOOL · CL_197779 ·

    Sakana AI 为更新后的 Sakana Fugu 会话式AI提供免费试用

    Sakana AI 已更新其会话式AI服务 Sakana Fugu,并开始提供免费试用。此举旨在提高日本AI初创公司会话模型和产品的可访问性。此次更新对于希望比较模型响应特性和产品用户体验的开发者尤其有用。

  11. TOOL · CL_182731 ·

    Bluesky、大和证券和 Sakana AI 推出 Jetstream AI 用于市场分析

    Bluesky 和大和证券正将其联合人工智能项目 Jetstream 推向全面生产。该代理式人工智能系统将部署到大和的财富管理团队,以增强复杂的市场分析,尤其是在动荡条件下。这标志着 Sakana AI 的一个重要里程碑。

  12. SIGNIFICANT · CL_181426 ·

    Sakana AI 发布日本专用 AI 模型“Sakana Namazu”

    Sakana AI 发布了一款名为“Sakana Namazu”的新 AI 模型,该模型专为日本市场设计。这款专用模型旨在满足日本独特的语言和文化细微差别。此次发布标志着为满足特定区域需求和语言而量身定制的 AI 开发趋势日益增长。

  13. TOOL · CL_178241 ·

    使用LLM同行评审对AI科学家系统进行基准测试

    一项新近发表在arXiv上的研究介绍了一个用于评估AI科学家系统的基准协议,这些系统旨在进行自主研究。该协议利用了GPT-5.4、Gemini和Claude等前沿大型语言模型来评估AI生成的论文在原创性、科学严谨性、清晰度和重要性方面的表现。在测试中,一家商业公司FARS的论文在Sakana AI、CycleResearcher和Data-to-Paper等竞争框架中表现出色,在1-5的评分尺度上获得了更高的分数。

  14. SIGNIFICANT · CL_177305 ·

    Sakana AI 发布 Fugu-Ultra v1.1,超越 "Fable 5"

    Sakana AI 发布了 Fugu-Ultra v1.1,这是一个先进的多智能体系统,可作为单一基础模型运行。据报道,新版本在性能上超越了“Fable 5”。此外,Fugu-Ultra v1.1 还包含一个兼容“Claude Code”的端点。

  15. TOOL · CL_174425 ·

    AI科学家框架旨在自动化完整的研发周期

    AI科学家是一个能够管理整个研发周期的自主人工智能框架。这些框架由Sakana AI等实验室开创,旨在构思假设、进行实验、分析数据,甚至撰写科学手稿。

  16. SIGNIFICANT · CL_163533 ·

    Sakana AI 推出 Fugu-Cyber 用于网络安全任务 · 已追踪 2 个来源

    Sakana AI 推出了 Fugu-Cyber,一个专门针对网络安全任务进行调优的编排模型。该模型在 CyberGym 基准测试中取得了 86.9% 的成绩,在 CTI-REALM 基准测试中取得了 72.1% 的成绩,使其成为与 GPT-5.5-Cyber 和 Claude Mythos Preview 等其他专业模型竞争的有力选项。Fugu-Cyber 在 Sakana 的 Fugu orchestrator 中运行,该 orc…

  17. TOOL · CL_162429 ·

    Sakana AI 发布 Fugu Ultra v1.1 支持 Claude Code

    Sakana AI 发布了 Fugu Ultra v1.1,这是其 AI 模型路由器的更新版本,与 v1.0 相比性能提升了 7.9 分。新版本现在包含一个 Claude Code 端点。然而,欧盟用户仍无法访问 Fugu Ultra。

  18. TOOL · CL_161790 ·

    Sakana AI的Fugu Ultra v1.1路由器声称性能超越Fable 5

    Sakana AI发布了其Fugu Ultra AI路由器v1.1的更新版本,声称性能有显著提升。该公司声称,即使Fable 5不被纳入路由池,新版本也能超越Fable 5模型。然而,该服务仍不对欧盟用户开放。

  19. TOOL · CL_157882 ·

    Fireworks AI:将任务分配给Kimi K3和Fable 5可将成本效率提高50倍

    Fireworks AI进行了一项实验,比较了其开源模型Kimi K3与闭源模型Fable 5在约1000项任务上的表现。结果表明,虽然两个模型总体表现相当,但它们各有优势,Kimi K3在符号计算和开发工具等领域表现出色,而Fable 5在Web开发和数据可视化方面表现更好。通过策略性地将任务分配给成本效益更高的模型,或结合使用,Fireworks AI在长时代理任务的成本效率方面取得了高达50倍的提升,理论上的最优分配策略准确率达到93%。

  20. TOOL · CL_156245 ·

    Sakana AI的Fugu模型在网络安全基准测试中取得最先进的性能

    总部位于日本的Sakana AI开发了一个名为Fugu的编排模型。该模型在真实世界网络安全基准测试中展现了最先进的性能。团队对他们的成就表示自豪。