H200 GPU
PulseAugur coverage of H200 GPU — every cluster mentioning H200 GPU across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Perplexity launches Q2D-Web benchmark for retrieval in agentic RAG systems
Perplexity has introduced Q2D-Web, a new benchmark and leaderboard designed to evaluate retrieval performance in agentic RAG systems. The benchmark utilizes a large corpus of 190 million web documents and over 69,000 ag…
-
New SSRR loss boosts neural audio codec intelligibility and speed
A new research paper introduces a self-supervised representation reconstruction (SSRR) loss for neural audio codecs, aiming to improve intelligibility and reduce latency. This method accelerates training, allowing compe…
-
New methods slash LLM distillation costs and boost context length
A new paper from Multiverse Computing introduces two methods to make knowledge distillation for large language models more efficient. The first method, offline distillation, caches the teacher model's top-K logits, redu…
-
MegaSlide-DiT enables large video diffusion model adaptation on single GPU
Researchers have developed MegaSlide-DiT, a system enabling the adaptation of large video diffusion models on a single high-end GPU. This is achieved by keeping model weights and optimizer states in host RAM and streami…
-
China's synthetic diamond industry poised to supply AI chip cooling solutions
The increasing thermal demands of AI chips, which now exceed 1,000W, are creating a significant opportunity for China's synthetic diamond industry. Synthetic diamond is an ideal thermal solution due to its superior heat…