PulseAugur
实时 16:46:12

inclusionAI 发布 Ling-3.0-flash 模型,提供官方 FP8 权重

inclusionAI 已在 Hugging Face 上发布其 Ling-3.0-flash 模型,提供 BF16 和官方 FP8 版本。该模型拥有 1275 亿个总参数和 51 亿个激活参数,采用具有 512 个专家的细粒度架构。FP8 版本体积显著减小,约为 128GB,对于拥有大量统一内存或多 GPU 设置的用户来说更加易于使用。 AI

影响 推出一款新的大型语言模型,其高效的 FP8 版本可供更广泛地使用。

排序理由 Frontier-lab 模型发布,附带系统卡。[lever_c_demoted from frontier_release: ic=2 ai=1.0]

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

inclusionAI 发布 Ling-3.0-flash 模型,提供官方 FP8 权重

报道来源 [2]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/derspenti ·

    inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vfdeek/inclusionailing30flash_weights_are_up_on_hugging/"> <img alt="inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8" src="https://external-preview.redd.it/N3g5MjI3N…

  2. r/LocalLLaMA TIER_1 English(EN) · /u/-Cubie- ·

    inclusionAI/Ling-3.0-flash · Hugging Face

    <!-- SC_OFF --><div class="md"><p>The Ling-3.0-flash MoE is now open-weighted at 124B A5B params. I know the original announcements were before the Kimi K3, DeepSeek-V4-Flash and Qwen3.8 hype, but this model might still have a good niche for itself due to its sizing. </p> <p>Disc…