PulseAugur
EN
LIVE 08:08:56

Qwen 3.8-27B model achieves impressive speeds on AMD hardware

A Reddit user shared performance metrics for the Qwen 3.8-27B model running on AMD hardware. The user reported achieving 24 tokens per second on a Ryzen AI MAX+ 395 and 51 tokens per second on a Radeon AI PRO R9700. These figures are presented as impressive for a dense model on AMD GPUs, and the user is soliciting further performance data and optimized command-line arguments from the community for various quantization levels and context lengths. AI

IMPACT Demonstrates potential for efficient local LLM deployment on AMD hardware.

RANK_REASON User-shared performance metrics for a specific model on consumer hardware.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Qwen 3.8-27B model achieves impressive speeds on AMD hardware

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/pmttyji ·

    Qwen 3.8 27B on AMD - 24 t/s on Ryzen AI Max+ 395 & 51 t/s on Radeon AI PRO R9700

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vouti3/qwen_38_27b_on_amd_24_ts_on_ryzen_ai_max_395_51/"> <img alt="Qwen 3.8 27B on AMD - 24 t/s on Ryzen AI Max+ 395 &amp; 51 t/s on Radeon AI PRO R9700" src="https://preview.redd.it/maecfved8hjh1.png?width=…