PulseAugur
实时 18:01:01
English(EN) Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why

Kimi K3 在网络漏洞测试中落后于美国模型,蒸馏被怀疑 · 跟踪到 2 个来源

Moonshot AIKimi K3 模型在网络漏洞任务上的表现明显弱于领先的美国模型,在 ExploitBench 上的得分仅为 32%,而美国模型为 76%。该模型的安全防护措施在阻止模拟攻击方面也显得不足。研究人员推测,Kimi K3 在通用基准测试中得分较高,但在安全相关评估中表现不佳的差异,可能归因于蒸馏技术,其中可能涉及 Anthropic 的模型。 AI

影响 强调了 AI 模型中潜在的安全漏洞,并表明蒸馏技术可能会影响专业性能。

排序理由 该集群报告了 AI 模型在特定任务上的基准测试结果,属于研究范畴。

在 The Decoder 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Kimi K3 在网络漏洞测试中落后于美国模型,蒸馏被怀疑 · 跟踪到 2 个来源

报道来源 [2]

  1. The Decoder TIER_1 English(EN) · Matthias Bastian ·

    Kimi K3 在网络漏洞利用方面远落后于美国前沿模型,蒸馏法或可解释原因

    <p><img alt="" class="attachment-full size-full wp-post-image" height="1152" src="https://the-decoder.com/wp-content/uploads/2026/07/aisi_logo_pattern.png" style="height: auto; margin-bottom: 10px;" width="2048" /></p> <p> The British AI Security Institute and the U.S. Center for…

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Kimi K3 在 ExploitBench 上得分 32%,而顶级美国模型的得分为 76%,在防范网络漏洞方面存在薄弱的安全措施。蒸馏可能解释了其通用 b 之间的差距

    Kimi K3 scored 32% on ExploitBench vs 76% for top US models, with weak safeguards against cyber exploits. Distillation may explain the gap between its general benchmarks and security performance. # AI # Automation Source: The Decoder AI https:// the-decoder.com/kimi-k3-trails -fr…