GB200
PulseAugur coverage of GB200 — every cluster mentioning GB200 across labs, papers, and developer communities, ranked by signal.
5 day(s) with sentiment data
-
NVIDIA open-sources OSMO for unified AI robotics development
NVIDIA has open-sourced OSMO, a Kubernetes-native workflow orchestrator designed to streamline AI development for robotics. OSMO allows developers to define training, simulation, and robot testing pipelines in a single …
-
NVIDIA vLLM supports DeepSeekv4.1 Flash on release; AMD vLLM lags
NVIDIA's vLLM software is functioning seamlessly with the new DeepSeekv4.1 Flash model across all six of its hardware SKUs, including H100, H200, B200, B300, GB200, and GB300. In contrast, AMD's vLLM implementation is e…
-
NVIDIA releases quantized DeepSeek and Qwen LLMs for Blackwell hardware
NVIDIA has released quantized versions of two large language models, DeepSeek-V4-Pro-0813 and Qwen3.8-2.4T-A95B, utilizing their NVFP4 quantization method. The DeepSeek model, with 1.65 trillion parameters, employs Hybr…
-
OpenAI unveils Jalapeño inference chip, challenging NVIDIA's dominance
OpenAI has revealed details about its custom inference chip, codenamed Jalapeño, which reportedly offers significant improvements in efficiency and latency compared to NVIDIA's Blackwell and Rubin-class systems. The chi…
-
OpenAI's custom 'jalapeño' chip benchmarks show it beating NVIDIA hardware
OpenAI has released benchmark results for its custom inference chip, codenamed "jalapeño," which it developed in collaboration with Broadcom. The chip reportedly outperforms NVIDIA's GB300 and GB200 systems in throughpu…
-
OpenAI's Jalapeño ASIC benchmarks show performance gains over Nvidia GPUs
OpenAI has developed its own 700W inference ASIC, codenamed Jalapeño, in collaboration with Broadcom. Benchmarks presented by OpenAI suggest that Jalapeño outperforms Nvidia's GB200 and GB300 GPUs in throughput per kilo…
-
OpenAI's Jalapeño chip shows superior inference performance over Nvidia
OpenAI has revealed initial performance data for its custom-designed "Jalapeño" inference chip, showcasing significant improvements in speed and power efficiency. Benchmarks indicate that Jalapeño outperforms competitor…
-
MiniMax H3 slashes video generation latency by 27x with NVIDIA Sol Engine
MiniMax AI has achieved a significant breakthrough in video generation latency using NVIDIA's Sol Engine and their MiniMax H3 model. By splitting the generation process into a low-resolution draft and a high-resolution …
-
NVIDIA releases Nemotron-Labs-Teacher models with 1M context · 4 sources tracked
NVIDIA has released a suite of Nemotron-Labs-Teacher models, each with 550 billion parameters, though only 55 billion are actively used. These models leverage a LatentMoE architecture incorporating Mamba-2, MoE, and Mul…
-
OpenAI's Jalapeño chip claims efficiency gains; agent systems evolve
OpenAI has released benchmark details for its custom inference chip, Jalapeño, claiming significant improvements in efficiency and latency over NVIDIA's GB200 and GB300 systems. The chip reportedly offers better perform…
-
Anyscale Ray optimizes NVIDIA GB300 NVL72 with NVLink Domain placement
Anyscale has introduced NVLink Domain-Aware Placement Groups for its Ray framework, designed to optimize performance on NVIDIA's GB300 NVL72 systems. These new placement groups ensure that tightly coupled actors are sch…
-
CoreWeave signs 2029 A100 GPU contracts, proving value of older AI hardware
CoreWeave, a cloud provider specializing in AI infrastructure, has signed a contract for NVIDIA A100 GPUs that extends into 2029, demonstrating the continued profitability of older AI hardware. Despite concerns about ra…
-
China's DFSX chip doubles memory bandwidth over NVIDIA's GB200
China's DFSX has developed a new chip that reportedly offers double the memory bandwidth of NVIDIA's GB200. This advancement utilizes a 14nm supernode process that bypasses microbumps for vertical compute memory towers.…
-
NVIDIA: AI cluster performance hinges on configuration, not just hardware
NVIDIA has highlighted that identical AI clusters using their H100, GB200, or GB300 systems can exhibit significant throughput differences. The company's analysis indicates that specific configuration choices have a gre…
-
Nota AI releases 4-bit quantized Solar Open2 250B model for NVIDIA Blackwell
Nota AI has released a 4-bit quantized version of Upstage's Solar Open2 250B model, named Solar Open2 250B — Nota NVFP4. This new version utilizes Nota AI's proprietary quantization technology, specifically designed for…
-
GMI Cloud unveils full-stack AI solutions at WAIC 2026
GMI Cloud showcased its comprehensive AI infrastructure solutions at WAIC 2026, highlighting its AI Cloud, MaaS, and Agentbox platforms. As a NVIDIA Cloud Partner, GMI Cloud offers GPU cloud services powered by high-per…
-
Kimi K3 model's scale and efficiency boost AI hardware demand, says SemiAnalysis
SemiAnalysis argues that the Kimi K3 model, despite its linear attention and lower KV-cache requirements, is beneficial for NVIDIA and the broader AI hardware ecosystem. The model's massive 2.8 trillion parameters neces…
-
NVIDIA's VR NVL72 System Features Cableless Design, Rubin GPUs
SemiAnalysis has provided a detailed breakdown of NVIDIA's upcoming VR NVL72 system, highlighting significant design changes from previous architectures like GB200. Key innovations include a cableless compute tray desig…
-
Liquid cooling industry enters order fulfillment phase, driven by GB200/GB300 demand
The liquid cooling industry is transitioning from speculative interest to tangible order fulfillment and performance realization. Companies with manufacturing capacity, client certifications, and key component expertise…
-
NVIDIA GPUs and Grace CPUs Power 81% of World's Fastest Supercomputers
NVIDIA technology dominates the latest TOP500 and Green500 supercomputer rankings, powering 81% of the TOP500 systems and the top eight on the Green500. The company's Grace CPU and GPUs are increasingly integrated into …