Alibaba's Qwen3.8 Max has significantly improved its performance, achieving a score of 56 on the Artificial Analysis Intelligence Index. This marks a substantial leap from its previous version, Qwen3.7 Max, which scored 46. While Qwen3.8 Max now matches the performance of Anthropic's Claude Opus 4.8, the Kimi k3 model continues to offer superior results at a lower cost. AI
IMPACT Qwen3.8 Max's improved performance challenges established models like Claude Opus 4.8, potentially driving further competition and cost optimization in the LLM market.
RANK_REASON The cluster reports on benchmark scores for AI models, which falls under research.
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →