PulseAugur
EN
LIVE 03:26:01

Open-weight Qwen3.8 challenges proprietary Claude Opus on benchmarks

Qwen3.8, an open-weight model, has demonstrated performance comparable to proprietary models like Claude Opus. While Claude Opus scored 63.1 on a benchmark, Qwen3.8 achieved 57.7, a gap of 5.4 points. Furthermore, Qwen3.8 is significantly more cost-effective, costing four times less per million output tokens. AI

IMPACT Highlights the increasing competitiveness of open-weight models against leading proprietary systems in terms of both performance and cost-efficiency.

RANK_REASON Comparison of open-weight vs proprietary model performance and cost. [lever_c_demoted from research: ic=1 ai=1.0]

Read on Mastodon — sigmoid.social →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Open-weight Qwen3.8 challenges proprietary Claude Opus on benchmarks

COVERAGE [1]

  1. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    ⚖️ Open vs proprietary, today Open-weight: Qwen3.8 2.4T A95B - 57.7 Proprietary: Claude Opus 5 - 63.1 Gap: 5.4 points · and 4x cheaper per 1M output tokens http

    ⚖️ Open vs proprietary, today Open-weight: Qwen3.8 2.4T A95B - 57.7 Proprietary: Claude Opus 5 - 63.1 Gap: 5.4 points · and 4x cheaper per 1M output tokens https:// olud.ai/leaderboard.html # OpenSource # AI # LLM