Alibaba releases Qwen3.8 multimodal model with broad hardware support · 10 sources tracked
ByPulseAugur Editorial·[65 sources]·
Alibaba's Qwen team has released Qwen3.8, a new multimodal dense model available in various sizes including 27B and 2.4T parameters. This release emphasizes day-zero support across diverse hardware like MediaTek Dimensity, Nvidia RTX Spark, and AMD Ryzen AI processors, enabling local deployment for developers. The model boasts a 262K native context window, extendable to 1M tokens, and shows improved performance in coding and office workflows compared to previous versions.
AI
IMPACT
Accelerates local AI development and deployment across diverse consumer hardware, lowering barriers to entry for multimodal AI applications.
RANK_REASON
Frontier-lab model release with system card.
<p>Qwen3.8-27B and Qwen3.8-2.4T can now be run locally in Unsloth!<br /> Run on 17GB RAM via Unsloth Dynamic GGUFs. You can also fine-tune Qwen3.8-27B in Unsloth.<br /> Qwen3.8-27B is by far the strongest model for its size. We also uploaded NVFP4 quants.</p> <p>Guide: <a href="h…
Heees my review of Qwen 3.8 27B - I can't remember the last time I've had this much fun playing with a local model that runs on my own computers simonwillison.net/2026/Aug/16/...
Qwen3.8-2.4T-A95B is finally out. Does anyone have a setup that can run this thing at home? lol # LLM # AI https:// huggingface.co/Qwen/Qwen3.8-2. 4T-A95B
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vx9302/this_is_what_qwen_38_27b_is_capable_of/"> <img alt="This is what Qwen 3.8 27b is capable of" src="https://preview.redd.it/56yzlgpvsclh1.png?width=640&crop=smart&auto=webp&s=28efa6e9c71a91d6…
<!-- SC_OFF --><div class="md"><p>Further to my last post, <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vldngi/tested_in_coding_bf16_muse_glimmer_vs_bf16_qwen36/">https://www.reddit.com/r/LocalLLaMA/comments/1vldngi/tested_in_coding_bf16_muse_glimmer_vs_bf16_qwen36/</a>…
<!-- SC_OFF --><div class="md"><p>So usually I avoid Q3 quants because I have had bad experiences with it, models were usually too degraded, so the smallest I normally do is Q4, since I only have rtx 4060 ti 16gb. But since there hasn't been a 35b-3ab released yet, I had to try i…
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vtozjq/qwenmix37_kept_seeing_posts_about_qwen38_and_36/"> <img alt="QwenMix-3.7: Kept seeing posts about Qwen3.8 and 3.6 sharing the same structure.. so I had Qwen3.8 combine them." src="https://external-prev…
<!-- SC_OFF --><div class="md"><p>Like many of you I've spent the last few days throwing Qwen3.8-27B against all of my usual use-cases and personal tasks/harnesses and workflows. It's great, phenomenal sometimes, but that's not what this post is about.</p> <p>One of my little per…
<!-- SC_OFF --><div class="md"><p>Hello everyone, tried so hard to optimize my config and finally I simply get up to 70 t/s with q6 variant. And wanted to share with you guys so that other people with the same setup can enjoy. Please check out and see if that improves your perfor…
<!-- SC_OFF --><div class="md"><p>Not sure if this is a stupid question but unsloth's models has the MTP built into the model right? I assume that is at the cost of some memory.</p> <p>If i want to use dflash, should i use a model that doesn't have MTP support then to save some v…
<!-- SC_OFF --><div class="md"><p>I'm trying to get a model to reverse engineer / decompile or otherwise reverse to source some of my old c,c++, pascal, and asm demo programs I made from decades ago and I'm constantly met with refusals. It is highly annoying. </p> <p>Does anybody…
Sidequest: Making Qwen 3.8 27B run as fast as possible for us VRAM-poor Nvidia 5060Ti 16GB owners. I don't know where this will end up, but I've spent a few nights on a fork of the Blackwell-only project NInfer but for low-VRAM cards like the 5060. NInfer is the fastest way to ru…
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vrabno/made_this_game_in_two_prompts_with_q4_qwen_38_is/"> <img alt="Made this game in two prompts with Q4, Qwen 3.8 is amazing" src="https://external-preview.redd.it/c3l4bmhqcWU1MWtoMXQju_xsxXc82LMLkgjpmBBzZ…
<!-- SC_OFF --><div class="md"><p>Artificial analysis index scores</p> <ul> <li>Qwen 3.5 27b: 35</li> <li>Qwen 3.6 27b: 38</li> <li>Qwen 3.8 27b: 52</li> </ul> <p>What the hell kind of a jump was that? Even if it is benchmaxxed, the jump is insane.</p> <ul> <li>Qwen3.6 35b A3b: 3…
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vqhw5s/share_your_favorite_thoughts_and_reasoning_from/"> <img alt="Share your favorite thoughts and reasoning from running Qwen 3.8 27b. This is mine." src="https://preview.redd.it/5t85ust41vjh1.png?width=14…
<table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vq9zc8/qwen_38_27b_vs_36_27b_how_good_is_with_a_turtle/"> <img alt="Qwen 3.8 27b vs 3.6 27b - how good is with a Turtle library." src="https://preview.redd.it/23dbjpl97tjh1.jpeg?width=640&crop=smart&a…
<!-- SC_OFF --><div class="md"><p>At hugging face, this is the model I use for most of my analysis that most of my services have to not be refused. previous versions are quite good. It seems this might have dropped today and pulling right now!</p> </div><!-- SC_ON -->   submi…
<!-- SC_OFF --><div class="md"><p>I sadly was not able to follow as much as I wanted the new advancements. Care to share your optimised setups ? Vllm, llama.cpp or any other engine? </p> </div><!-- SC_ON -->   submitted by   <a href="https://www.reddit.com/user/sagiroth">…
<!-- SC_OFF --><div class="md"><p>I see everyone gushing over 3.8, I get the impression people find it drastically better than previous Qwen models, but I can't believe it could be better than 3.5 122B. Is it?</p> </div><!-- SC_ON -->   submitted by   <a href="https://www…
<!-- SC_OFF --><div class="md"><p>I asked, "why don't you change to SCAN instead of KEYS for cache invalidation performance". </p> <p>Then, qwen 3.8 27b performed test code, regression tests, e2e tests, and even benchmarks for 2 hours.</p> <p>Here's final report from qw…
<p>The two models look identical on paper. Same dense 27B, same hybrid architecture, same 262,144-token context window, same Apache 2.0 license, same vision encoder. Qwen3.6-27B shipped in April and earned a nickname the community rarely gives out: "the sweet spot for local devel…
<!-- SC_OFF --><div class="md"><p>Megathread to help with the influx of duplicate / similar posts around the release of the Qwen 3.8 27B release.</p> <ul> <li>Quants</li> <li>Fine-Tunes & Abliterations</li> <li>Chat Templates</li> <li>Server Support & Configuration</li> <…
<!-- SC_OFF --><div class="md"><p>So a bit of context, I am a cybersecurity senior analyst<br /> I am interested in LLMs for that field especially with MCPs to connect them to the tools or for writing scripts </p> <p>I started this field by doing assembly language reading for hac…
<!-- SC_OFF --><div class="md"><p>I'm always testing an image prompt with a picture of a historic place in my hometown – a small but well known 250,000 people town in Germany. I'll just ask the model, in which City this photo has been taken.</p> <p>With the 3.6 generation of both…
<!-- SC_OFF --><div class="md"><p>Facing any issues?</p> <p>Chat Template is fine? </p> <p>Looping issue? </p> <p>Too much reasoning thing?</p> <p>How's MTP with this one? </p> <p>Any other issues faced by Qwen3.6-27B & Qwen3.5-27B during release time?</p> <p>If I missed any …
<!-- SC_OFF --><div class="md"><p>With your experiments, Qwen 3.8 27B most close which frontier model? And please specify which quantization you run. I will post to comments my tests and experience too.</p> </div><!-- SC_ON -->   submitted by   <a href="https://www.reddit…
<!-- SC_OFF --><div class="md"><p>So far, it seems like Qwen 3.8 might drop today as a 27B dense model. If that’s the case, no MoE offloading this time , I used to run Qwen 3.6 35B-A3B at around 70 tok/s on an RTX 3060, but offloading a dense model is a completely different story…
<!-- SC_OFF --><div class="md"><p>They highlighted 3 things on countdown page: VLM, Agentic Improvements, and Think mode. What improvements do you expect? Reply here!</p> <p>Personally, I want to meet a sage who has attained enlightenment. The first ASI was 27B. For example, I ho…
<!-- SC_OFF --><div class="md"><p>To my fellow crazies, the few. Those who dared wrestle with llama-70b, mistral-large, goliath, mistral8x22B, DeepSeekV2/3, wept when llama4 behemoth was announced, picked yourself up and are now wrestling with DeepSeekV4Pro, GLM5.2, MiMoV2.5Pro a…
<!-- SC_OFF --><div class="md"><p>i was hoping they would release both today but the big one just dropped now , and it says its been listed since 5 hours ago on huggingface. </p> <p>I guess the real date for 27b is <a href="https://modelscope.cn/models/Qwen/Qwen3.8-27B">https://m…
<!-- SC_OFF --><div class="md"><p>1) Are scaling laws dead ? If a model this small is so intelligent, then what's the key to intelligence ?</p> <p>2) Are the majority of today's biggest frontier models parameters just "fluff" that don't help a model reason, or don't hol…