PulseAugur
中
实时 19:51:55
English(EN) Hot Chips 2026: Nvidia presents Groq 3 LPX architecture and unveils its first third-party inference benchmark — LP30-based rack already in production, company says

NVIDIA Groq 3 LPX 推理芯片全面投入生产,目标是代理式 AI

NVIDIA 宣布其 Groq 3 LPX 推理芯片已全面投入生产,据称在代理式 AI 工作负载方面实现了每秒 3,400 个 token 的处理速度。尽管 NVIDIA 声称这比 Cerebras 快四倍,但由于所需的加速器数量不同,这种比较变得复杂,NVIDIA 需要多个单元才能匹配 Cerebras 的单加速器性能。Groq 3 LPX 旨在加速代理式 AI 推理,NVIDIA 还在探索将其 CUDA 支持扩展到 RISC-V 架构以增强 GPU 计算能力。 AI

影响 加速代理式 AI 推理,可能为 AI 工作负载的速度和效率设定新的基准。

排序理由 NVIDIA 宣布其 Groq 3 LPX 推理芯片全面投入生产。

在 Tom's Hardware 阅读 →

AI 生成摘要 · Google Gemini · 来自 10 个来源。 我们如何撰写摘要 →

NVIDIA Groq 3 LPX 推理芯片全面投入生产,目标是代理式 AI

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Frontier Release
NVIDIA 宣布其 Groq 3 LPX 推理芯片全面投入生产。
Source corroboration
10 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
45 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+2 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

完整方法见我们的编辑标准。

报道来源 [10]

  1. TLDR AI TIER_1 English(EN) · TLDR ·

    英伟达的 Groq 芯片 ⚡,前沿经济学 💰,Ox Alpha 之谜 🕵️

  2. The Decoder TIER_1 English(EN) · Maximilian Schreiner ·

    Nvidia称其Groq 3 LPX速度是Cerebras的四倍,但计算更复杂

    <p><img alt="" class="attachment-full size-full wp-post-image" height="1080" src="https://the-decoder.com/wp-content/uploads/2025/09/NVIDIA-Vera-Rubin-NVL144-with-Tray.jpg" style="height: auto; margin-bottom: 10px;" width="1920" /></p> <p> Nvidia is moving its Groq 3 LPX inferenc…

  3. Tom's Hardware TIER_1 English(EN) · Luke James ·

    Hot Chips 2026:Nvidia 展示 Groq 3 LPX 架构并公布首个第三方推理基准测试 — 公司称基于 LP30 的机架已投入生产

    Igor Arsovski, now Nvidia's VP of hardware, presented the Groq 3 LPX rack's architecture and published the first third-party benchmark of the hardware.

  4. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    NVIDIA Groq 3 LPX 全面投产,目标是 Agentic AI 推理 #AgenticAI #AgenticArtificialIntelligence #AI#

    https://www. europesays.com/3213703/ NVIDIA Groq 3 LPX enters full production, targeting agentic AI inference # AgenticAI # AgenticArtificialIntelligence # AI # ArtificialIntelligence

  5. The Register — AI TIER_1 English(EN) ·

    英伟达首款 Groq 3 LPU 基准测试揭示了什么,又隐藏了什么,关于其200亿美元的赌注

    Gemma 4 31B performance tests offer a best-case scenario for next-gen dataflow accelerators

  6. The Register — AI TIER_1 English(EN) ·

    英伟达首款 Groq 3 LPU 基准测试揭示其 200 亿美元赌局

    Gemma 4 31B performance tests offer a best-case scenario for next-gen dataflow accelerators

  7. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    英伟达公司表示其 Groq 3 LPX 人工智能推理加速器已全面投产。该芯片于 Hot Chips 2026 上发布。来源:SiliconANGLE https:// s

    Nvidia Corp. says its Groq 3 LPX AI inference accelerator has entered full production. The chip was announced at Hot Chips 2026. Source: SiliconANGLE https:// siliconangle.com/2026/08/24/nv idias-dedicated-inference-accelerator-groq-3-lpx-enters-full-production-to-supercharge-ai-…

  8. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    NVIDIA 的 Groq 3 LPU 加速器为 Heterogeneous AI Compute 在 Hot Chips 2026 上亮相

    NVIDIA’s Groq 3 LPU Accelerators for Heterogeneous AI Compute at Hot Chips 2026 https:// fed.brid.gy/r/https://www.serv ethehome.com/nvidias-groq-3-lpu-accelerators-for-heterogeneous-ai-compute-at-hot-chips-2026/

  9. Mastodon — mastodon.social TIER_1 日本語(JA) · [email protected] ·

    NVIDIA的“Groq 3 LPX”将Agent AI速度提升高达4倍

    NVIDIA、エージェント型AIを最大4倍高速化する「Groq 3 LPX」 https://www. watch.impress.co.jp/docs/news/ 2135283.html # watch_impress # テック # AI

  10. Mastodon — mastodon.social TIER_1 English(EN) · notatechguy ·

    NVIDIA Groq 3 LPX 全面量产,速度达每秒 3400 个 token NVIDIA 的 Groq 3 LPX 推理加速器现已全面量产,速度创下每秒 3400 个 token 的新纪录

    NVIDIA Groq 3 LPX hits full production at 3,400 tokens/second NVIDIA's Groq 3 LPX inference accelerator is now in full production, delivering a record 3,400 tokens/second for agentic AI workloads. https://www. notatechguy.com/nvidia-groq-3- lpx-hits-full-production-at-3-400-token…