PulseAugur
EN
LIVE 00:03:35

Qwen 3.8 27B model performance tested on MacBook Max

A user on Reddit shared their experience running the Qwen 3.8 27B model on a MacBook Max with 128GB of RAM. They reported initial speeds of around 30 tokens/second, which only slightly decreased to 27 tokens/second even with a context window usage of up to 50,000 tokens. The user is seeking advice on optimizing performance and confirming their quantization settings for the model, which is running via Unsloth Studio. AI

IMPACT Provides insights into the practical performance of large language models on high-end consumer hardware.

RANK_REASON User-generated performance test of a specific model on consumer hardware.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Qwen 3.8 27B model performance tested on MacBook Max

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Odd-Environment-7193 ·

    m5 Macbook Max 128gb. QWEN 3.8 27b Tk/s? Quick Test.

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vpg5h4/m5_macbook_max_128gb_qwen_38_27b_tks_quick_test/"> <img alt="m5 Macbook Max 128gb. QWEN 3.8 27b Tk/s? Quick Test." src="https://preview.redd.it/fxwst9c68mjh1.png?width=140&amp;height=81&amp;auto=webp&a…