PulseAugur
EN
LIVE 13:50:35

Vision Qwen 3.8 27B model runs on 16GB card with 85K context

A user on Reddit shared their configuration for running the Vision Qwen 3.8 27B model on a 16GB graphics card. The setup utilizes beellama.cpp and achieves an 85K context size with 45 tokens/second decode speed. The user also noted that moving the mmproj to the CPU could free up more VRAM for further optimization. AI

IMPACT Demonstrates efficient deployment of large models on limited hardware, potentially enabling wider access to advanced AI capabilities.

RANK_REASON User-shared configuration for running a specific LLM on consumer hardware.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Vision Qwen 3.8 27B model runs on 16GB card with 85K context

How we ranked this

Signal score
4 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
User-shared configuration for running a specific LLM on consumer hardware.
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

Full methodology in our editorial standards.

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/FerLuisxd ·

    Running Vision Qwen 3.8 27B on a 16GB Card, the config (45tks).

    <!-- SC_OFF --><div class="md"><p>I am just sharing my config for Qwen 3.8 27b that fits on a 5060TI, what is cool about this is that you can even get vision! and a 85K context (I have 1.5gb of headroom for more context or a better quant)</p> <p>Model: IQ3_XXS-mtp from <a href="h…