A user tested the Qwen 3.8-27B model by having it take the ACT standardized test, achieving a composite score of 36 on one test and 34 on another. The model answered 326 out of 342 questions correctly, demonstrating strong performance, particularly in the reading section where it achieved a perfect score. The test took approximately 177 minutes to complete, with the user noting that the model's use of vision capabilities to process the PDF documents contributed to the extended duration. AI
IMPACT Demonstrates the capability of specific LLMs to perform on complex, multi-section standardized tests.
RANK_REASON User-driven benchmark of a specific model version on a standardized test. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →