H200s
PulseAugur coverage of H200s — every cluster mentioning H200s across labs, papers, and developer communities, ranked by signal.
4 day(s) with sentiment data
SambaNova's strategy of optimizing existing hardware with new GPUs may lead to cost-effective AI solutions
SambaNova's focus on integrating NVIDIA H200 GPUs with their existing SN50 RDUs, rather than solely relying on new chip development, suggests a strategy aimed at providing more cost-effective AI inference solutions. This approach could appeal to businesses looking to enhance performance without the highest upfront hardware investment.
SambaNova's H200 integration yields significant benchmark improvements
SambaNova Systems has demonstrated substantial performance gains by integrating NVIDIA H200 GPUs with their SN50 RDUs. This heterogeneous approach achieved 763 tokens/sec on MiniMax M2.7 benchmark, and a subsequent partnership with MiniMax AI pushed this to 850 tokens/sec on short-context workloads. This indicates a strong synergy between SambaNova's architecture and NVIDIA's latest hardware for AI inference.
MiniMax AI M2.7 model shows record inference speeds with SambaNovaAI partnership
MiniMax AI's M2.7 model has achieved world-record inference speeds, reaching 850 tokens/sec on short-context and 450 tokens/sec on long-context workloads when paired with SambaNovaAI's H200s and SN50 RDUs. This performance, benchmarked by Artificial Analysis, highlights the effectiveness of this specific hardware combination for demanding AI tasks.
-
MiniMax AI M2.7 achieves record inference speeds with SambaNovaAI
MiniMax AI has announced a significant speed improvement for its M2.7 model, achieving new world records for inference speed. In partnership with SambaNovaAI, the models reached 850 tokens per second on short-context wo…
-
Red Hat offers 'forever' support for RHEL via new paid add-on
Red Hat is introducing a new 'Long-Life Add-On' for Red Hat Enterprise Linux (RHEL), offering extended support for specific releases indefinitely as long as customers are willing to pay. This move aims to provide stabil…
-
OpenAI enhances ChatGPT, SpaceXAI rebrands Grok, and AI hardware sees innovation · 1 source tracked
OpenAI has enhanced ChatGPT with a new feature called GPT-Live, enabling simultaneous talking, listening, and answer formulation for more natural conversations. Meanwhile, SpaceXAI, formerly known as X.AI, is rebranding…
-
Singapore's Temasek sees strong future in AI investments · 2 sources tracked
Singaporean sovereign wealth fund Temasek Holdings believes artificial intelligence will be a significant area for future investment and growth. The fund sees AI as a sector with strong potential for returns, despite th…
-
SambaNova integrates NVIDIA GPUs for AI benchmark boost · 1 source tracked
SambaNova Systems has demonstrated improved performance by integrating their SN50 RDUs with NVIDIA's H200 GPUs, achieving 763 tokens per second on the MiniMax M2.7 benchmark. This heterogeneous compute approach aims to …
-
GLM 5.2 model now runnable on consumer hardware with quantization
The GLM 5.2 model, a 753 billion parameter model with a 1 million token context window, is now available for local deployment on consumer hardware. While the full model requires over 1.5 TB of storage, quantized version…
-
Together Compute Expands GPU Offerings with H100, H200, and B200
Together, an inference and open-source AI company, has significantly expanded its on-demand compute platform. The company announced the addition of a substantial number of high-end GPUs, including H100s, H200s, and the …