PulseAugur
EN
LIVE 21:38:55

Qwen3.8-Max scores 53 on AA index, lags Kimi K3

The Qwen3.8-Max model has achieved a score of 53 on the AA index, a benchmark for evaluating AI models. While this performance is considered decent, it is noted to be comparable to the glm 5.2 model on coding indices and less effective than the Kimi K3 model. The model's cost is also mentioned as being higher than some alternatives. AI

IMPACT This benchmark provides a data point for comparing AI model performance, particularly in coding-related tasks.

RANK_REASON The item discusses a specific benchmark score for an AI model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Qwen3.8-Max scores 53 on AA index, lags Kimi K3

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 (SO) · /u/_AnemicRoyalty_ ·

    Qwen3.8-max: 53 on AA index

    <!-- SC_OFF --><div class="md"><p>Not too shabby, was expecting it to be better.</p> <p>Does roughly the same as glm 5.2 on coding indices, worse than kimi k3. More expensive though.</p> <p><a href="https://artificialanalysis.ai/models/qwen3-8-max?intelligence=coding-index">https…