PulseAugur
实时 04:50:09
English(EN) Qwen3.8-Max uses sparse mixture-of-experts routing: 2.4 trillion total parameters but only 95 billion active per query, meaning inference costs scale with activ

亚马逊增加 AI 基础设施支出;阿里巴巴发布 Qwen3.8-Max 模型 · 跟踪 2 个来源

亚马逊的首席财务官已将其公司 2026 年资本支出预测提高至 2200 亿美元,并预计这笔金额可能仍不足以满足需求,尤其是在 AI 基础设施的内存芯片可用性方面。与此同时,阿里巴巴发布了 Qwen3.8-Max,该模型采用了稀疏专家混合路由,总参数量为 2.4 万亿,但每次查询仅激活 950 亿,预计这将降低推理成本。一个较小的 Qwen3.8-27B 变体也计划发布,定价为每百万输入 token 2 美元。 AI

影响 亚马逊大幅增加资本支出的举动凸显了对 AI 基础设施日益增长的需求,而 Qwen3.8-Max 的高效路由有望降低 AI 推理成本。

排序理由 该集群涵盖了一家大型科技公司的重大资本支出公告以及一个重要 AI 实验室的新模型发布。

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

亚马逊增加 AI 基础设施支出;阿里巴巴发布 Qwen3.8-Max 模型 · 跟踪 2 个来源

报道来源 [2]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    亚马逊首席财务官将2026年资本支出提高至2200亿美元,并预计即便如此也无法满足需求。关注内存芯片短缺是否会成为制约因素

    Amazon's CFO raised 2026 capital spending to $220 billion and expects even that won't satisfy demand. Watch whether memory chip constraints become the binding constraint on AI infrastructure buildout across the industry. https://www. implicator.ai/amazon-tops-3-tr illion-market-c…

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Max 采用稀疏混合专家路由:总参数2.4万亿,但每次查询仅激活950亿,意味着推理成本随激活

    Qwen3.8-Max uses sparse mixture-of-experts routing: 2.4 trillion total parameters but only 95 billion active per query, meaning inference costs scale with active, not total, size. A smaller Qwen3.8-27B variant also launches next week. Pricing set at $2 per million input tokens. h…