PulseAugur
实时 17:36:10
English(EN) Still early but some of first benchmarks for DeepSeek V4-Pro 0813 lands at 87.9, within a tenth of a point of Fable 5's 88.0, at roughly 57x cheaper on output pricing

DeepSeek V4-Pro 0813 基准测试接近 Fable 5,价格大幅降低

DeepSeek V4-Pro 0813 模型已显示出有希望的基准测试结果,得分达到 87.9。这一性能非常接近 An Ape and a Fox 的 Fable 5 模型,后者的得分是 88.0。值得注意的是,据报道 DeepSeek V4-Pro 0813 的输出定价大约便宜 57 倍。 AI

影响 此次基准测试表明 DeepSeek V4-Pro 0813 是一个具有竞争力且成本效益高的替代现有模型的选择。

排序理由 该条目报告了特定 AI 模型的基准测试结果,并将其与其他模型进行了比较。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/Anthropic 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

DeepSeek V4-Pro 0813 基准测试接近 Fable 5,价格大幅降低

报道来源 [1]

  1. r/Anthropic TIER_1 English(EN) · /u/ColdKiwi720 ·

    仍处于早期阶段,但 DeepSeek V4-Pro 0813 的首批基准测试得分达到 87.9,与 Fable 5 的 88.0 相差不到 0.1 分,而输出定价大约便宜 57 倍

    <table> <tr><td> <a href="https://www.reddit.com/r/Anthropic/comments/1vmmx0i/still_early_but_some_of_first_benchmarks_for/"> <img alt="Still early but some of first benchmarks for DeepSeek V4-Pro 0813 lands at 87.9, within a tenth of a point of Fable 5's 88.0, at roughly 57x che…