NVL72
PulseAugur coverage of NVL72 — every cluster mentioning NVL72 across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
AI interconnect demand surges, reshaping copper vs. optical debate
The demand for AI infrastructure is driving up the need for high-speed interconnects, pushing companies like Astera Labs to develop solutions that can handle increased data exchange between GPUs. This has led to a re-ev…
-
Together optimizes inference on Nvidia Rubin GPU with new features · 4 sources tracked
Together has announced advancements in their inference and OSS capabilities, leveraging new features from Nvidia's Rubin GPU. The company has updated its b200 gemms to incorporate Rubin's wider MMA steps, increased tens…
-
Nvidia and Palantir partner on AI-driven supply chain optimization · 4 sources tracked
Nvidia and Palantir have partnered to leverage AI for optimizing supply chain operations, with Nvidia's own complex, million-part supply chain serving as the initial test case. The collaboration utilizes Palantir Foundr…
-
OpenAI chip reportedly outperforms NVIDIA Rubin NVL72 on perf per watt
OpenAI has reportedly developed a new chip that offers superior performance per watt compared to NVIDIA's latest offerings. Despite lacking specific decoding capabilities, OpenAI's chip reportedly outperforms NVIDIA's R…
-
SemiAnalysis releases AgentX dataset for agentic inferencing
SemiAnalysis has released AgentX, an open-source dataset designed for agentic inferencing, featuring a 1 million token context length and multi-turn capabilities. The dataset aims to test the resilience of CUDA's domina…
-
TPU partners with Mooncake for inference optimization
SemiAnalysis reports that Tensor Processing Unit (TPU) is collaborating with the open-source inference optimization library Mooncake. This partnership aims to integrate TPU capabilities with Mooncake Store, enhancing pe…
-
AI Hardware Focus Shifts to Energy, Supply Chains, and Networking
The AI hardware landscape is undergoing a significant shift, moving beyond just GPUs to focus on energy efficiency, supply chain stability for critical minerals like copper, and advanced networking capabilities. Former …
-
Cursor releases open-source MoE training megakernel, Mixture-of-Kittens
Cursor Research has open-sourced Mixture-of-Kittens (MoK), a specialized training kernel designed for Mixture-of-Experts (MoE) models. This megakernel fuses MoE communication and computation into a single deterministic …
-
AI hardware advances with new clusters; open-source debated
Ineffable Labs has acquired its first Vera Rubin-NVL72 clusters, powered by Google Cloud and NVIDIA hardware. This move signifies a generational leap in AI hardware, accelerating the pace of advancement in the field. Se…
-
Kimi K3 model's scale and efficiency boost AI hardware demand, says SemiAnalysis
SemiAnalysis argues that the Kimi K3 model, despite its linear attention and lower KV-cache requirements, is beneficial for NVIDIA and the broader AI hardware ecosystem. The model's massive 2.8 trillion parameters neces…
-
SpaceX boosts AI Sat V1 power specs for enhanced compute capacity
SpaceX has increased the power specifications for its AI Sat V1 satellite. The peak power output has been raised to approximately 250kW with battery assistance, and the average power will be around 160kW. This upgraded …
-
Nvidia CEO denies Vera Rubin AI platform delays, cites 'giant amounts' in production
Nvidia CEO Jensen Huang has refuted claims of production delays for the company's upcoming Vera Rubin AI platform, stating that it is already in production and will be delivered in 'giant amounts.' While Huang confirmed…
-
AI Value Capture Shifts: From Infrastructure to Model Labs
SemiAnalysis reports that the AI industry's value capture is shifting from infrastructure providers to model labs. While 2023-2025 saw significant gains for companies like TSMC and NVIDIA due to infrastructure build-out…
-
NVIDIA's VR NVL72 System Features Cableless Design, Rubin GPUs
SemiAnalysis has provided a detailed breakdown of NVIDIA's upcoming VR NVL72 system, highlighting significant design changes from previous architectures like GB200. Key innovations include a cableless compute tray desig…
-
Anthropic's Claude models now run on NVIDIA Blackwell Ultra GPUs in Azure
Anthropic's Claude models are now generally available on Microsoft Azure, running on NVIDIA's GB300 Blackwell Ultra GPUs within Microsoft Foundry. This collaboration aims to provide enterprises with enhanced computing p…
-
NVIDIA Blackwell platform dominates MLPerf Training 6.0 benchmarks
NVIDIA's Blackwell platform has set new records in the MLPerf Training 6.0 benchmarks, achieving the fastest times across all seven tests. The platform demonstrated strong scaling, with clusters of up to 8,192 GPUs show…
-
Microsoft brings Nvidia Rubin GPU rack online with Foxconn
Microsoft has successfully brought online its first NVL72 rack, equipped with Nvidia's Rubin VR200 GPUs, in collaboration with ODM partner Foxconn. This marks a significant step in Microsoft's infrastructure build-out, …
-
Cerebras integrates AI rack onto single wafer, drawing Google interest
Cerebras has developed a novel approach to AI chip manufacturing, integrating an entire NVL72 rack onto a single wafer. This design bypasses traditional networking bottlenecks by routing around defects and keeping all c…
-
Nvidia blueprints AI factories as GPT-4.1 accuracy drops in real-world medical cases
Nvidia has released validated blueprints for AI data centers, detailing configurations for 4-node to 128-node clusters. These designs, named RTX PRO, HGX, and NVL72, are intended for advanced applications like agentic A…
-
Nvidia's GB300 GPU shows 2.7x faster inference than GB200
Nvidia's GB300 ultra NVL72 has demonstrated a 2.7x speed advantage over the GB200 NVL72 in inference tasks using the vLLM project's engine. This performance leap exceeds theoretical expectations based on the GB300's spe…