PulseAugur
EN
LIVE 08:01:04

inclusionAI releases Ling-3.0-flash model with official FP8 weights

inclusionAI has released its Ling-3.0-flash model, available in both BF16 and an official FP8 version, on Hugging Face. The model boasts 127.5 billion total parameters with 5.1 billion active parameters, featuring a fine-grained architecture with 512 experts. The FP8 version is notably smaller, around 128GB, making it more accessible for users with significant unified memory or multi-GPU setups. AI

IMPACT Makes a new large language model with an efficient FP8 version available for broader use.

RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=2 ai=1.0]

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 3 sources. How we write summaries →

inclusionAI releases Ling-3.0-flash model with official FP8 weights

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
Source corroboration
3 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
57 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+1 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [3]

  1. Hugging Face Trending Models TIER_1 English(EN) · inclusionAI ·

    inclusionAI/Ling-3.0-flash

    text-generation · 25 downloads · 77 likes

  2. r/LocalLLaMA TIER_1 English(EN) · /u/derspenti ·

    inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vfdeek/inclusionailing30flash_weights_are_up_on_hugging/"> <img alt="inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8" src="https://external-preview.redd.it/N3g5MjI3N…

  3. r/LocalLLaMA TIER_1 English(EN) · /u/-Cubie- ·

    inclusionAI/Ling-3.0-flash · Hugging Face

    <!-- SC_OFF --><div class="md"><p>The Ling-3.0-flash MoE is now open-weighted at 124B A5B params. I know the original announcements were before the Kimi K3, DeepSeek-V4-Flash and Qwen3.8 hype, but this model might still have a good niche for itself due to its sizing. </p> <p>Disc…