Strata应用程序已更新,支持Qwen3.8-Flash-Next,使其能在消费级硬件上运行,并显著提升速度。这一进展使得拥有中等配置GPU和充足系统内存的用户能够运行先进的模型,可能取代更大、更难访问的版本。Strata通过优化模型加载和采用推测解码来实现这一点,使强大的AI模型更容易在本地使用。
AI
<!-- SC_OFF --><div class="md"><p>I've had one pull request merged into ds4 (DwarfStar), a tiny one. There are a few more still waiting in the queue. I’m not complaining. Antirez says it clearly in the README: with coding agents everyone can tune the engine for their hardware and…
<!-- SC_OFF --><div class="md"><p>It provides the speed and somewhat accessibility to low vram users to run Flash Next as their main. It's providing the speed of qwen 3.5-35-a3b but with higher intelligence over 3.8 27b. This will allow many people to replace 27b. </p> <p>I'm get…
<!-- SC_OFF --><div class="md"><p>IQ3_XXS weights are just under 80GB and my slowww DDR4+7900XTX is stabilizing around 45-70/s (sometimes higher while coding depending on mtp). Looking online I'm seeing similar results for users with 12GB and 16GB cards, and significantly faster …
<table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1wvxq9n/qwen38flashnext_on_strata_is_such_a_beast/"> <img alt="Qwen3.8-Flash-Next on Strata is such a beast" src="https://external-preview.redd.it/AbNvNzqvm0Ww5dJ46NocG0fLIaVid5gtnTwnGEDx9GI.png?width=140…