PulseAugur
EN
LIVE 08:21:15

User seeks fastest Qwen 3.8-27B for AMD GPU setup

A user on the r/LocalLLaMA subreddit is seeking the fastest Qwen 3.8-27B model version for their AMD GPU setup. They are specifically looking for GGUF versions that optimize prompt processing and token generation speeds. The user has a dual-GPU configuration with a W7900 and a W7800, each with 48GB of VRAM, totaling 96GB, but is encountering issues with tensor paralleling due to a bandwidth mismatch. AI

RANK_REASON This is a user query on a specific hardware configuration for a particular model, not a general industry announcement or development.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

User seeks fastest Qwen 3.8-27B for AMD GPU setup

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Gloomy_Letterhead395 ·

    Fastest qwen 3.8 27b for AMD gpu?

    <!-- SC_OFF --><div class="md"><p>Hey, just wondering if there are forks or exact gguf versions that give fastest prompt processing and token gen speeds for AMD gpu?<br /> Looking to run q8 or q6<br /> Vram 96gb<br /> W7900 + w7800 both 48gb<br /> With bandwidth mismatch, tensor …