PulseAugur
实时 17:44:43
English(EN) Grok 4.6 Released: Benchmarks, Pricing, and What It Means for Agent Builders

xAI 发布 Grok 4.6,目标是 Agent 任务和交易基准领先地位

xAI 发布了 Grok 4.6,该模型旨在用于长期 Agent 和复杂的交互式任务,而不是实现原始智能的飞跃。该公司声称其在 Artificial Analysis Intelligence Index 上可与 GPT-5.6 Sol 相媲美,并在其他基准测试中与 GPT-5.6 Sol 和 Fable 5 争夺领先地位,尽管它在实际终端工作等领域表现出弱点。Grok 4.6 的定价为每百万输入 token 2 美元起,每百万输出 token 6 美元,还有一个速度更快的版本,价格是其两倍。 AI

影响 此次发布标志着在 Agent 能力和前沿模型之间的基准测试竞争方面持续受到关注。

排序理由 前沿实验室模型发布,包含系统卡和基准测试数据。[lever_c_demoted from frontier_release: ic=1 ai=1.0]

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

xAI 发布 Grok 4.6,目标是 Agent 任务和交易基准领先地位

报道来源 [1]

  1. dev.to — LLM tag TIER_1 English(EN) · jamilxt ·

    Grok 4.6 发布:基准测试、定价以及对 Agent 构建者的意义

    <p>On August 12, 2026, xAI released Grok 4.6, the successor to Grok 4.5 that shipped in July. The positioning is different from the last release. This is not pitched as a raw intelligence jump. It is a model built for long-running agents and ambitious interactive and visual work:…