PulseAugur
EN
LIVE 08:28:03

Alibaba's Qwen3.8-Max claims SOTA over GPT-5.6, Fable 5, but faces scrutiny

Alibaba has released Qwen3.8-Max, a 2.4 trillion parameter model with a 1 million token context window, claiming it surpasses GPT-5.6 and Claude Fable 5 in agentic computer use benchmarks. However, the author urges caution, emphasizing the need to verify the model's license, distinguish between self-reported and independent benchmarks, and consider the practicalities of running such a large model. The author expresses particular interest in a smaller 27B version, which would be more feasible for self-hosting and potentially more impactful if released under a permissive license. AI

IMPACT If the Qwen3.8-27B model is released with a permissive license and performs as claimed, it could significantly boost indie builders and self-hosting capabilities.

RANK_REASON Frontier-lab model release with system card and benchmark claims. [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on dev.to — LLM tag →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Alibaba's Qwen3.8-Max claims SOTA over GPT-5.6, Fable 5, but faces scrutiny

COVERAGE [1]

  1. dev.to — LLM tag TIER_1 English(EN) · frank chu ·

    Qwen3.8-Max says it beats GPT-5.6 and Fable 5 at computer use. Here's my checklist before I believe any open-weights release

    <p>Alibaba released Qwen3.8-Max on August 3: a 2.4-trillion-parameter MoE with a 1M-token context window, priced at $2/$6 per million tokens. The claim that got everyone's attention: <strong>86.1 on OSWorld-Verified</strong>, ahead of GPT-5.6 Sol Max (83.2) and Claude Fable 5 (85…