PulseAugur
中
实时 19:42:39
English(EN) Tuned/abliterated Qwen3.8-27b into a 24gb card 262k guff using the newest unreleased version of LexiPanel. It's fast with reliable draft acceptance. Made for 7900xtx but should work on whatever 24gb card with this setup and headless. Doesn't get dumber while coding like most of the other fine-tunes.

Qwen3.8-27B 模型针对 24GB 显卡进行了编码微调

Qwen3.8-27B 模型的一个微调版本 CODER 已发布,该版本针对编码任务进行了优化。此版本经过量化并使用 LexiPanel 进行适配,可在 24GB 显卡上运行,上下文窗口约为 262k token。它在草稿接受方面表现出快速且可靠的性能,即使在超过 100k token 的上下文中,测得速度也达到 36-41 token/s。 AI

影响 使得在消费级硬件上本地运行先进的编码模型成为可能,有望改善开发人员的工作流程。

排序理由 发布了一个用于本地部署的微调开源模型。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Qwen3.8-27B 模型针对 24GB 显卡进行了编码微调

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
发布了一个用于本地部署的微调开源模型。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
1 days old
Coverage has settled into its steady-state source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/W61k3r ·

    将 Qwen3.8-27b 微调/优化至 24gb 卡 262k guff,使用了最新未发布的 LexiPanel 版本。它速度快,草稿接受率可靠。专为 7900xtx 设计,但在此设置和无头模式下应适用于任何 24gb 卡。不像大多数其他微调模型那样,在编码时会变笨。

    <!-- SC_OFF --><div class="md"><p><a href="https://huggingface.co/Wa1k3r/Qwen3.8-27b-CODER-4q_xs-24GB-262k-Optimalcardfit">https://huggingface.co/Wa1k3r/Qwen3.8-27b-CODER-4q_xs-24GB-262k-Optimalcardfit</a> </p> <h1>Qwen3.8-27B CODER — IQ4_XS imatrix · 24 GB card fit · ~262k conte…