GB300
PulseAugur coverage of GB300 — every cluster mentioning GB300 across labs, papers, and developer communities, ranked by signal.
- 2026-07-23 product_launch NVIDIA partner Wistron opened a new plant in Texas to manufacture GB300 chips. source
3 day(s) with sentiment data
-
SemiAnalysis criticizes GTC VR performance, highlights AgentX at AI Infra Summit
SemiAnalysis claims Jensen presented misleading virtual reality performance metrics at Nvidia GTC, contrasting it with Ian Buck's upcoming presentation of AgentX at the AI Infra Summit in Santa Clara. The post highlight…
-
AMD MI355X closes performance gap with GB300 via SGLang and UMBP integration · 3 sources tracked
SemiAnalysis reports that AMD's MI355X is rapidly improving in agentic inference performance and total cost of ownership, nearing parity with GB300. This advancement is attributed to AMD's SGLang team and their MoRI lib…
-
OpenAI unveils Jalapeño AI chip, designed with LLM assistance
OpenAI has unveiled its first custom AI accelerator chip, named Jalapeño, which reportedly offers significant performance and power efficiency improvements over existing hardware like Nvidia's GB300. The chip was design…
-
HP ZGX Fury Workstation with GB300 Superchip Now Available
HP has made its ZGX Fury workstation available for order, featuring the GB300 Superchip and 748GB of unified memory. This workstation is designed to support AI workloads, with HP also planning an AI Factory for edge dep…
-
NVIDIA vLLM supports DeepSeekv4.1 Flash on release; AMD vLLM lags
NVIDIA's vLLM software is functioning seamlessly with the new DeepSeekv4.1 Flash model across all six of its hardware SKUs, including H100, H200, B200, B300, GB200, and GB300. In contrast, AMD's vLLM implementation is e…
-
OpenAI unveils Jalapeño inference chip, challenging NVIDIA's dominance
OpenAI has revealed details about its custom inference chip, codenamed Jalapeño, which reportedly offers significant improvements in efficiency and latency compared to NVIDIA's Blackwell and Rubin-class systems. The chi…
-
OpenAI's custom 'jalapeño' chip benchmarks show it beating NVIDIA hardware
OpenAI has released benchmark results for its custom inference chip, codenamed "jalapeño," which it developed in collaboration with Broadcom. The chip reportedly outperforms NVIDIA's GB300 and GB200 systems in throughpu…
-
OpenAI's Jalapeño ASIC benchmarks show performance gains over Nvidia GPUs
OpenAI has developed its own 700W inference ASIC, codenamed Jalapeño, in collaboration with Broadcom. Benchmarks presented by OpenAI suggest that Jalapeño outperforms Nvidia's GB200 and GB300 GPUs in throughput per kilo…
-
OpenAI's Jalapeño chip shows superior inference performance over Nvidia
OpenAI has revealed initial performance data for its custom-designed "Jalapeño" inference chip, showcasing significant improvements in speed and power efficiency. Benchmarks indicate that Jalapeño outperforms competitor…
-
SemiAnalysis releases AgentX dataset for agentic inferencing
SemiAnalysis has released AgentX, an open-source dataset designed for agentic inferencing, featuring a 1 million token context length and multi-turn capabilities. The dataset aims to test the resilience of CUDA's domina…
-
Nvidia's $100K DGX Station with GB300 Superchip listed online
Nvidia's new DGX Station desktop, powered by the GB300 Superchip, has been listed online for approximately $100,000. This high-performance machine features a Blackwell Ultra GPU with 252GB of HBM3e memory and a 72-core …
-
NVIDIA hires 109 Poolside AI staff in unique licensing deal
Poolside, a company known for its AI model factory, has undergone a significant shift involving NVIDIA and its founder Jensen Huang. Instead of a traditional acquisition, NVIDIA has licensed Poolside's technology and hi…
-
OpenAI's Jalapeño chip claims efficiency gains; agent systems evolve
OpenAI has released benchmark details for its custom inference chip, Jalapeño, claiming significant improvements in efficiency and latency over NVIDIA's GB200 and GB300 systems. The chip reportedly offers better perform…
-
SpaceX plans massive 2027 datacenter expansion, fueled by AI inference profits · 5 sources tracked
SemiAnalysis reports that SpaceX plans to build 6-8GW of incremental datacenters in 2027, potentially exceeding 10GW, with a path to $300 billion in annual recurring revenue. This ambitious expansion is driven by the hi…
-
CoreWeave signs 2029 A100 GPU contracts, proving value of older AI hardware
CoreWeave, a cloud provider specializing in AI infrastructure, has signed a contract for NVIDIA A100 GPUs that extends into 2029, demonstrating the continued profitability of older AI hardware. Despite concerns about ra…
-
NVIDIA DGX Station offers powerful Linux-based AI processing
NVIDIA's DGX Station, powered by its Grace CPU and Blackwell GPU, offers substantial local AI processing capabilities. This machine, which bears a resemblance to Apple's Mac Pro, currently operates exclusively on Linux.…
-
Attn-QAT enables stable 4-bit attention training for LLMs
Researchers have developed Attn-QAT, a novel method for 4-bit quantization-aware training of attention mechanisms in large language models. This approach addresses the challenges of low precision in FP4 computation, par…
-
NVIDIA: AI cluster performance hinges on configuration, not just hardware
NVIDIA has highlighted that identical AI clusters using their H100, GB200, or GB300 systems can exhibit significant throughput differences. The company's analysis indicates that specific configuration choices have a gre…
-
US Accuses China's Moonshot AI of Stealing Anthropic Model Tech
The U.S. government has accused Chinese AI firm Moonshot AI of stealing proprietary technology from Anthropic's Claude Fable-5 model to develop its Kimi K3 system. Officials allege Moonshot built a platform to covertly …
-
NVIDIA urges AI model co-design to boost GPU utilization
NVIDIA has released a technical blog post highlighting a critical issue in AI model design: poor hardware utilization due to models not being optimized for GPU architecture. The post explains that concepts like 'arithme…