H200
PulseAugur coverage of H200 — every cluster mentioning H200 across labs, papers, and developer communities, ranked by signal.
- instance of graphics processing unit 90%
- used by NVIDIA H100 70%
- competes with NVIDIA H100 70%
- used by Blackwell 70%
- affiliated with Supermicro 70%
- competes with H.1000 Gnome 70%
- used by Supermicro 70%
- competes with MI300X 70%
- used by vLLM 70%
- used by ByteDance 70%
- used by graphics processing unit 60%
- instance of NVIDIA H100 50%
- 2026-05-14 product_launch US government approves sale of NVIDIA H200 AI chips to ten Chinese companies.
7 day(s) with sentiment data
-
New RASP-QAOA system optimizes QAOA simulations
A new research paper introduces RASP-QAOA, a system designed to optimize the simulation of Quantum Approximate Optimization Algorithms (QAOA). RASP-QAOA intelligently selects the most suitable computational representati…
-
Moonshot AI releases Kimi K3 with 2.8T open weights, but only 1.8% are active
Moonshot AI has released the Kimi K3 model with 2.8 trillion open weights, making it the largest open-weight model to date. However, only a small fraction, approximately 1.8%, of these parameters are actively used for p…
-
Moonshot AI releases 2.8T Kimi K3 weights, largest ever, but impractical to run
Moonshot AI has released the full 2.8 trillion parameter weights for its Kimi K3 model, making it the largest open-weight model to date. Despite the massive size and open release, running Kimi K3 is practically impossib…
-
Induction Labs unveils Photon-1 imagination model outperforming Gemini
Induction Labs has introduced Photon-1, a 106-billion parameter mixture-of-experts model trained on raw video without action labels. This 'imagination model' architecture predicts future frames in a learned representati…
-
GMI Cloud unveils full-stack AI solutions at WAIC 2026
GMI Cloud showcased its comprehensive AI infrastructure solutions at WAIC 2026, highlighting its AI Cloud, MaaS, and Agentbox platforms. As a NVIDIA Cloud Partner, GMI Cloud offers GPU cloud services powered by high-per…
-
Cloud GPU rental guide for LLMs: Optimizing cost by model size
The optimal cloud GPU rental for running large language models (LLMs) in 2026 depends on the specific model size and workload, with a focus on tokens per dollar rather than hourly rates. For smaller models (7B-13B), bud…
-
NVIDIA H200 GPU liquid-cooling installation detailed in teardown video
A detailed video showcases the disassembly of an NVIDIA H200 NVL GPU and the installation of an EK-Pro H200 NVL water block. The process involves separating the PCB, cleaning and replacing thermal pads and paste, and ca…
-
US says NVIDIA H200 AI chip exports to China remain 'trivial'
The United States has stated that exports of NVIDIA's H200 AI chips to China have been minimal, despite recent approvals. A top Commerce Department official indicated that only a very small quantity has reached mainland…
-
US allows ZTE to buy Nvidia H200 AI chips, joining tech elite
The U.S. government has granted approval for Chinese telecommunications company ZTE to purchase Nvidia's H200 AI chips. This decision allows ZTE to join a select group of Chinese tech firms, including Alibaba, Tencent, …
-
Nvidia restricts AI chip sales in Asia to curb smuggling · 2 sources tracked
Nvidia is reportedly tightening its customer list in Asia to combat the smuggling of its AI chips, particularly into China. The company has allegedly halved its authorized client base after implementing stricter complia…
-
China's chip exports surge amid global AI demand · 1 source tracked
China's chip exports have nearly doubled in the first half of the year, driven by the global AI boom and strong demand for computing hardware. This surge, with integrated circuit exports increasing by over 96% year-on-y…
-
Tencent and VIDRAFT showcase sparse MoE models with reduced active parameters
Tencent has released Hy3, a 295-billion-parameter Mixture-of-Experts (MoE) model that utilizes only 21 billion active parameters per forward pass, significantly reducing inference costs. This MoE architecture, featuring…
-
Irish data centers consume 23% of national electricity due to AI growth
Irish data centers are consuming a significant portion of the country's electricity, reaching 23% of the national total. This surge in energy demand is largely driven by the expansion of AI infrastructure, with companie…
-
New CTA-pipelining method slashes multi-GPU latency for LLMs
Researchers have introduced CTA-pipelining, a novel execution paradigm for multi-GPU systems that optimizes for latency in serving large language models. This method exploits dependencies at the Cooperative Thread Array…
-
China Approves NVIDIA H200 AI Chip Imports for Domestic Firms
China has reportedly approved the import of NVIDIA's H200 AI chips for domestic AI companies. This move aims to support the development and deployment of advanced AI technologies within China. The approval signifies a p…
-
China Eases Nvidia H200 Chip Restrictions for Select AI Firms · 2 sources tracked
China is reportedly easing restrictions on the purchase of Nvidia H200 AI chips for select domestic companies like Alibaba and ByteDance. This policy shift aims to alleviate bottlenecks in AI training within China while…
-
Former GitHub CEO launches new AI development platform Revue de droit de l'Université de Sherbrooke
A new AI development platform, Revue de droit de l'Université de Sherbrooke, has been launched by the former CEO of GitHub. This platform aims to cater to the evolving needs of AI development, particularly in what the a…
-
GLM 5.2 achieves 79.8% on Terminal-Bench 2.1 with FP8 precision
A user on Reddit shared benchmark results for the GLM 5.2 model, achieving a score of 79.8% on the Terminal-Bench 2.1 test. The user specified that this score was achieved using FP8 precision for both the model weights …
-
HexGrid Cloud offers custom LLM GPU benchmarking for open-weight models
HexGrid Cloud is offering to benchmark open-weight LLMs on user-specified GPUs and configurations. They are seeking suggestions for models and hardware setups to test their deployment platform, focusing on chat/instruct…
-
New MoP training stack enables trillion-parameter MoE models with 1M context
Researchers have introduced a novel training stack called Mixture-of-Parallelisms (MoP) designed to enhance memory efficiency for Mixture-of-Experts (MoE) models. This approach integrates various existing and new parall…