PulseAugur
实时 22:05:50
English(EN) Qwen3.8-27B scored 29/30 on AIME 2026 with FP8 + xhigh reasoning — BF16 vs FP8 results

Qwen3.8-27B 以 FP8 在 AIME 2026 数学基准测试中获得 29/30 分

Qwen3.8-27B 模型在 AIME 2026 数学数据集上的基准测试显示,其量化的 FP8 权重在设置为 xhigh 推理时,获得了 29/30 的分数。此性能在相同的 xhigh 推理设置下与 BF16 版本相当,但速度显著提高。FP8 xhigh 配置在速度更快的情况下,也达到了 BF16 medium 设置的分数。 AI

影响 在复杂推理任务上展现出强大性能,可能影响未来数学问题解决模型的开发。

排序理由 特定模型在数学数据集上的基准测试结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Qwen3.8-27B 以 FP8 在 AIME 2026 数学基准测试中获得 29/30 分

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/No_Run8812 ·

    Qwen3.8-27B 在 AIME 2026 中以 FP8 + xhigh 推理获得 29/30 分 — BF16 与 FP8 结果对比

    <!-- SC_OFF --><div class="md"><p>I benchmarked Qwen3.8-27B on <code>MathArena/aime_2026</code> dataset, comparing BF16 and FP8 weights at medium and xhigh reasoning effort.</p> <h1>Interesting findings are:</h1> <ol> <li>quantized FP8 xhigh is better than BF 16 medium equally go…