DeepGEMM
PulseAugur coverage of DeepGEMM — every cluster mentioning DeepGEMM across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
vLLM v0.31.0 enhances serving with FlashMLA, DeepGEMM, and MXFP8
vLLM has released version 0.31.0, bringing significant enhancements to its serving capabilities. This update integrates FlashMLA with V4.1 NVFP4 compressed KV cache as the default for SM100, alongside the inclusion of D…
-
DeepSeek open-sources AI infrastructure for Huawei Ascend hardware
DeepSeek has released open-source Ascend-compatible versions of its TileLang, DeepGEMM, and DeepEP infrastructure components. These tools are designed to optimize AI model performance on Huawei Ascend hardware. The rele…
-
DeepSeek, Huawei Open-Source Tools to Challenge Nvidia's AI Dominance · 10 sources tracked
DeepSeek and Huawei have jointly released a suite of open-source tools designed to enhance the software ecosystem for Huawei's Ascend AI chips. This initiative aims to provide a viable alternative to Nvidia's dominant C…
-
Zhipu AI's GLM-5.2 model deployed on serverless GPUs
Zhipu AI has released GLM-5.2, a 700B Mixture-of-Experts (MoE) model that excels in complex reasoning and software engineering tasks, reportedly matching or surpassing proprietary models like Claude 3.5 Sonnet and GPT-4…
-
DeepSeek V4 achieves faster performance with custom kernels, replacing cuBLAS
DeepSeek has developed a custom kernel stack, DeepGEMM and TileLang, which not only matches but surpasses the performance of NVIDIA's cuBLAS. This custom implementation achieves bitwise determinism and batch invariance,…