H200 GPUs
PulseAugur coverage of H200 GPUs — every cluster mentioning H200 GPUs across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New Data-Centric Parallel method speeds up variable-length sequence training
Researchers have developed a new technique called Data-Centric Parallel (DCP) to address the computational challenges of training deep learning models on variable-length sequences. DCP dynamically adjusts runtime settin…
-
STAGE framework synthesizes LLM execution graphs for distributed workloads · 2 sources tracked
A new framework called STAGE has been developed to synthesize high-fidelity execution graphs for large language models (LLMs) and Mixture-of-Experts (MoEs). This framework aims to optimize distributed AI workloads by mo…
-
Prime Intellect releases open framework for training trillion-parameter MoE models
Prime Intellect has launched prime-rl 0.6.0, an open framework designed for training large Mixture-of-Experts (MoE) models using agentic reinforcement learning. This new system successfully trained the GLM-5 model on so…
-
Nvidia chips smuggled to China and Russia despite US export controls
U.S. authorities are investigating multiple cases of advanced Nvidia GPUs and other semiconductor technology being illegally smuggled to China and Russia, circumventing export controls. These efforts involve sophisticat…
-
Hugging Face and AWS Detail Foundation Model Infrastructure
Hugging Face and AWS have collaborated to detail the infrastructure required for training and running large foundation models. The blog post outlines a layered architecture, emphasizing the interplay between AWS's compu…