PulseAugur
实时 02:22:37
English(EN) We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

LLaMA 社区寻求新的 70-80B 参数模型竞争者

r/LocalLLaMA 子版块的一位用户正在寻求推荐,并希望看到 700-800 亿参数范围内的新大型语言模型。他们目前使用 DSV4 Flash 0731Inkling SmallLaguna S 2.1Step-3.7-Flash 等模型,但发现它们在代理编码任务上的性能有些慢。用户希望在这个规模级别上出现新的混合专家(MoE)模型,其性能优于当前的 Qwen 3.6 等选项,目标是在他们的硬件上实现显著更快的生成速度。 AI

影响 用户需求表明中等规模、高性能 LLM 可能存在市场空白。

排序理由 用户讨论和对新模型的要求,而非发布或官方公告。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

LLaMA 社区寻求新的 70-80B 参数模型竞争者

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/_TheWolfOfWalmart_ ·

    最近我们看到了一些很棒的中等规模模型(DSV4 Flash 0731、Inkling Small、Laguna S 2.1、Step 3.7 Flash),但还有没有人想看到一些新的 70-80b 竞争者?

    <!-- SC_OFF --><div class="md"><p>I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B.</p> <p>On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is not the worst but it does get a bit ann…