DeepSeek-V4-Flash-0731,来自 DeepSeek 的新模型,现已在 Fireworks 推理平台上线。据报道,该模型在九项 agentic 评估中优于其前身 V4 Pro,在 Terminal Bench 上取得了 82.7% 的分数。它还提供了更高的成本效益和 100 万个 token 的上下文窗口。 AI
影响 在 agentic 基准测试中设定了新的 SOTA,可能影响企业对 AI agent 的采用。
排序理由 Frontier-lab 模型发布,附带系统卡。[lever_c_demoted from frontier_release: ic=1 ai=1.0]
在 X — Fireworks (inference infra) 阅读 →
AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →