PulseAugur
实时 19:43:27
Deutsch(DE) NVIDIA stellt Qwen3.8-2.4T-A95B als NVFP4-Quantisierung bereit. Das MoE-Modell (2,4T Parameter, 95B aktiv) nutzt 4-Bit-Präzision für vLLM/SGLang auf GB200/B300.

NVIDIA 发布 Qwen3.8-2.4T-A95B MoE 模型,支持 4 位量化

NVIDIA 发布了 Qwen3.8-2.4T-A95B,这是一个混合专家(MoE)模型,拥有 2.4 万亿参数和 950 亿活跃参数。该模型采用 NVFP4 量化和 4 位精度,针对 GB200/B300 硬件上的 vLLMSGLang 进行了优化。它支持 100 万个 token 的上下文长度。 AI

影响 此次发布提供了一个高效的 MoE 模型,具有大上下文窗口,有望提高复杂 AI 任务的性能并降低资源需求。

排序理由 来自主要 AI 实验室(NVIDIA)的模型发布。[lever_c_demoted from frontier_release: ic=1 ai=1.0]

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

NVIDIA 发布 Qwen3.8-2.4T-A95B MoE 模型,支持 4 位量化

本文如何被排名

Signal score
14 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
来自主要 AI 实验室(NVIDIA)的模型发布。[lever_c_demoted from frontier_release: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    NVIDIA 提供 Qwen3.8-2.4T-A95B 作为 NVFP4 量化。该 MoE 模型(2.4T 参数,95B 活跃)在 GB200/B300 上使用 4 位精度进行 vLLM/SGLang。

    NVIDIA stellt Qwen3.8-2.4T-A95B als NVFP4-Quantisierung bereit. Das MoE-Modell (2,4T Parameter, 95B aktiv) nutzt 4-Bit-Präzision für vLLM/SGLang auf GB200/B300. Kontextlänge: 1M Token. https:// huggingface.co/nvidia/Qwen3.8- 2.4T-A95B-NVFP4 # KI # AI # LLM # AISyndicate