NVL72
PulseAugur coverage of NVL72 — every cluster mentioning NVL72 across labs, papers, and developer communities, ranked by signal.
5 day(s) with sentiment data
-
Cursor releases open-source MoE training megakernel, Mixture-of-Kittens
Cursor Research has open-sourced Mixture-of-Kittens (MoK), a specialized training kernel designed for Mixture-of-Experts (MoE) models. This megakernel fuses MoE communication and computation into a single deterministic …
-
AI hardware advances with new clusters; open-source debated
Ineffable Labs has acquired its first Vera Rubin-NVL72 clusters, powered by Google Cloud and NVIDIA hardware. This move signifies a generational leap in AI hardware, accelerating the pace of advancement in the field. Se…
-
Kimi K3 model's scale and efficiency boost AI hardware demand, says SemiAnalysis
SemiAnalysis argues that the Kimi K3 model, despite its linear attention and lower KV-cache requirements, is beneficial for NVIDIA and the broader AI hardware ecosystem. The model's massive 2.8 trillion parameters neces…
-
SpaceX boosts AI Sat V1 power specs for enhanced compute capacity
SpaceX has increased the power specifications for its AI Sat V1 satellite. The peak power output has been raised to approximately 250kW with battery assistance, and the average power will be around 160kW. This upgraded …
-
Nvidia CEO denies Vera Rubin AI platform delays, cites 'giant amounts' in production
Nvidia CEO Jensen Huang has refuted claims of production delays for the company's upcoming Vera Rubin AI platform, stating that it is already in production and will be delivered in 'giant amounts.' While Huang confirmed…
-
AI Value Capture Shifts: From Infrastructure to Model Labs
SemiAnalysis reports that the AI industry's value capture is shifting from infrastructure providers to model labs. While 2023-2025 saw significant gains for companies like TSMC and NVIDIA due to infrastructure build-out…
-
NVIDIA's VR NVL72 System Features Cableless Design, Rubin GPUs
SemiAnalysis has provided a detailed breakdown of NVIDIA's upcoming VR NVL72 system, highlighting significant design changes from previous architectures like GB200. Key innovations include a cableless compute tray desig…
-
Anthropic's Claude models now run on NVIDIA Blackwell Ultra GPUs in Azure
Anthropic's Claude models are now generally available on Microsoft Azure, running on NVIDIA's GB300 Blackwell Ultra GPUs within Microsoft Foundry. This collaboration aims to provide enterprises with enhanced computing p…
-
NVIDIA Blackwell platform dominates MLPerf Training 6.0 benchmarks
NVIDIA's Blackwell platform has set new records in the MLPerf Training 6.0 benchmarks, achieving the fastest times across all seven tests. The platform demonstrated strong scaling, with clusters of up to 8,192 GPUs show…
-
Microsoft brings Nvidia Rubin GPU rack online with Foxconn
Microsoft has successfully brought online its first NVL72 rack, equipped with Nvidia's Rubin VR200 GPUs, in collaboration with ODM partner Foxconn. This marks a significant step in Microsoft's infrastructure build-out, …
-
Cerebras integrates AI rack onto single wafer, drawing Google interest
Cerebras has developed a novel approach to AI chip manufacturing, integrating an entire NVL72 rack onto a single wafer. This design bypasses traditional networking bottlenecks by routing around defects and keeping all c…
-
Nvidia blueprints AI factories as GPT-4.1 accuracy drops in real-world medical cases
Nvidia has released validated blueprints for AI data centers, detailing configurations for 4-node to 128-node clusters. These designs, named RTX PRO, HGX, and NVL72, are intended for advanced applications like agentic A…
-
Nvidia's GB300 GPU shows 2.7x faster inference than GB200
Nvidia's GB300 ultra NVL72 has demonstrated a 2.7x speed advantage over the GB200 NVL72 in inference tasks using the vLLM project's engine. This performance leap exceeds theoretical expectations based on the GB300's spe…