PulseAugur
中
实时 10:24:42
Deutsch(DE) NVIDIA publiziert Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 für Blackwell. LatentMoE mit Mamba-Interleaving, Multi-Token Prediction (MTP) und NVFP4-Quantisierung die

NVIDIA 发布 Nemotron-Labs-3-Puzzle-75B 以支持 Blackwell 硬件

NVIDIA 发布了其 Nemotron-Labs-3-Puzzle-75B 模型,该模型已针对 Blackwell 硬件上的服务进行了优化。该模型集成了 LatentMoE、Mamba-Interleaving 和 Multi-Token Prediction (MTP) 以提高吞吐量。它在 OpenMDW-1.1 许可下可用,允许商业使用。 AI

影响 优化服务基础设施,并可能提高大型模型的吞吐量。

排序理由 NVIDIA 是一个前沿实验室,发布了具有特定技术细节的新模型。

在 Mastodon — mastodon.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 3 个来源。 我们如何撰写摘要 →

NVIDIA 发布 Nemotron-Labs-3-Puzzle-75B 以支持 Blackwell 硬件

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Frontier Release
NVIDIA 是一个前沿实验室,发布了具有特定技术细节的新模型。
Source corroboration
3 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
92 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [3]

  1. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    NVIDIA 发布 Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 以支持 Blackwell。具有 Mamba-Interleaving、多令牌预测 (MTP) 和 NVFP4 量化的 LatentMoE

    NVIDIA publiziert Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 für Blackwell. LatentMoE mit Mamba-Interleaving, Multi-Token Prediction (MTP) und NVFP4-Quantisierung dienen der Serving-Optimierung. OpenMDW-1.1 erlaubt kommerzielle Nutzung; AIME25: 89,9. https:// huggingface.co/nvidia/NVID…

  2. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Nemotron-Labs-3-Puzzle-75B-A9B:通过迭代式Puzzle压缩的具有Mamba和Attention层的LatentMoE。MTP-Heads将8×B200的吞吐量提高了约

    Nemotron-Labs-3-Puzzle-75B-A9B: LatentMoE mit Mamba- und Attention-Layern, komprimiert via Iterative Puzzle. MTP-Heads steigern den Durchsatz auf 8×B200 um ca. 2×. Lizenz OpenMDW-1.1, kommerzielle Nutzung erlaubt. https:// huggingface.co/nvidia/NVIDIA-N emotron-Labs-3-Puzzle-75B-…

  3. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Nemotron-Labs-3-Puzzle-75B-A9B-FP8 采用 LatentMoE、Mamba2 和 Multi-Token Prediction。在 8×B200 上,吞吐量增加 2 倍,H100 在 1M tokens 上竞争

    Nemotron-Labs-3-Puzzle-75B-A9B-FP8 nutzt LatentMoE, Mamba2 und Multi-Token Prediction. Auf 8×B200 steigt der Durchsatz um 2×, die H100-Konkurrenz bei 1M-Token-Kontext auf 8 Requests. Lizenz: OpenMDW-1.1. AIME25: 89.4, MMLU-Pro: 82.0. https:// huggingface.co/nvidia/NVIDIA-N emotro…