PulseAugur
实时 14:19:36
English(EN) Run Ling 3.0 Flash Locally: 124B of Knowledge on a 96 GB Machine

蚂蚁集团 Ling 3.0 Flash 模型发布,支持本地部署

蚂蚁集团 inclusionAI 推出的 Ling 3.0 Flash,一个拥有 124B 参数的混合专家模型,已获得 MIT 许可并在 Hugging Face 上发布。该模型专为本地运行而设计,由于其每个 token 只激活 5.1B 参数的架构,相比 Kimi K3 等其他大模型,所需的内存显著减少。它支持长上下文的混合线性注意力,并默认支持推理,也可选择禁用推理以获得更快的响应。该模型可以在拥有 96GB 统一内存的 Mac 上或在拥有 24GB GPU 和大量系统内存的系统上进行交互式运行,社区已提供 GGUF 转换。 AI

影响 支持大模型在本地运行,可能降低研究人员和开发者的门槛。

排序理由 来自主要 AI 实验室(蚂蚁集团 inclusionAI)的模型发布,包含详细的技术规格和可用性。[lever_c_demoted from frontier_release: ic=1 ai=1.0]

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

蚂蚁集团 Ling 3.0 Flash 模型发布,支持本地部署

报道来源 [1]

  1. dev.to — LLM tag TIER_1 English(EN) · David ·

    本地运行 Ling 3.0 Flash:96GB 机器上的 124B 知识

    <p>The two big open releases on everyone's feed right now are Kimi K3 (2.8 trillion parameters, the first open 3T-class model) and Ling 3.0 Flash from Ant Group's inclusionAI. Only one of them can live on hardware a person owns, and it is not the one with the bigger headline. Her…