A benchmark test of the Qwen3.8-27B model on the AIME 2026 math dataset revealed that its quantized FP8 weights, when set to xhigh reasoning effort, achieved a score of 29/30. This performance was comparable to the BF16 version at the same xhigh reasoning setting, but with significantly improved speed. The FP8 xhigh configuration also matched the BF16 medium setting's score while being faster. AI
IMPACT Demonstrates strong performance on complex reasoning tasks, potentially influencing future model development for mathematical problem-solving.
RANK_REASON Benchmark results for a specific model on a math dataset. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →