PulseAugur
实时 03:33:30
English(EN) ⚖️ Open vs proprietary, today Open-weight: Qwen3.8 2.4T A95B - 57.7 Proprietary: Claude Opus 5 - 63.1 Gap: 5.4 points · and 4x cheaper per 1M output tokens http

开放权重模型 Qwen3.8 在基准测试中挑战专有模型 Claude Opus

开放权重模型 Qwen3.8 在性能上已能与 Claude Opus 等专有模型相媲美。Claude Opus 在一项基准测试中得分 63.1,而 Qwen3.8 达到 57.7,差距为 5.4 分。此外,Qwen3.8 的成本效益也显著更高,每百万输出令牌的成本低四倍。 AI

影响 凸显了开放权重模型在性能和成本效益方面,与领先的专有系统相比,竞争力日益增强。

排序理由 开放权重模型与专有模型性能和成本的比较。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — sigmoid.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

开放权重模型 Qwen3.8 在基准测试中挑战专有模型 Claude Opus

报道来源 [1]

  1. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    ⚖️ 开放模型 vs 专有模型,今日开放权重:Qwen3.8 2.4T A95B - 57.7 专有模型:Claude Opus 5 - 63.1 差距:5.4分 · 且每100万输出token便宜4倍 http

    ⚖️ Open vs proprietary, today Open-weight: Qwen3.8 2.4T A95B - 57.7 Proprietary: Claude Opus 5 - 63.1 Gap: 5.4 points · and 4x cheaper per 1M output tokens https:// olud.ai/leaderboard.html # OpenSource # AI # LLM