PulseAugur
实时 02:31:50
English(EN) Gemma 4’s 2 GB Mac Story Is Really SSD Streaming

Gemma 4 26B Mac 演示澄清是 SSD 流式传输,而非 2GB RAM 使用

最近在 Mac 上演示 Google 的 Gemma 4 26B-A4B 模型引发了关于其内存需求的讨论。虽然最初被宣传为 2 GB 模型,但仔细检查后发现它利用了 SSD 流式传输来实现其专家混合(Mixture-of-Experts)架构,需要大约 3 GB 的驻留 RAM 和显著更多的磁盘空间。这项技术通过按需从 SSD 加载模型组件,而不是将整个模型保存在 RAM 中,从而实现了较低的活动内存使用量。实际内存占用量可能因具体运行时和配置而异,某些设置甚至需要高达 26 GB 的 RAM。 AI

影响 澄清了在消费级硬件上运行大型语言模型的实际内存需求,并强调了 SSD 流式传输在管理资源限制中的作用。

排序理由 文章讨论了一个特定的技术演示及其在消费级硬件上运行特定 LLM 的影响,而不是一个新的模型发布或重大的行业事件。

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Gemma 4 26B Mac 演示澄清是 SSD 流式传输,而非 2GB RAM 使用

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
文章讨论了一个特定的技术演示及其在消费级硬件上运行特定 LLM 的影响,而不是一个新的模型发布或重大的行业事件。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
49 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [1]

  1. dev.to — LLM tag TIER_1 English(EN) · Simon Paxton ·

    Gemma 4 的 2 GB Mac 故事其实是 SSD 流式传输

    <p><strong>Gemma 4 26B-A4B does not, on the best citable evidence here, run as a true all-in-RAM 2 GB model on Macs.</strong> The strongest published figures from the project behind the claim put <strong>Gemma-only resident memory at about 3 GB with SSD-streamed expert loading</s…