GLM-5.2-FP8
PulseAugur coverage of GLM-5.2-FP8 — every cluster mentioning GLM-5.2-FP8 across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
Moonshot AI's Kimi K3 open-source model now on Telnyx API
Moonshot AI's Kimi K3, a 2.8 trillion parameter open-source model, is now accessible via the Telnyx Inference API. This model boasts a 1 million token context window, native vision capabilities, and configurable reasoni…
-
GLM-5.2-FP8-DSpark deployment shows mixed performance gains
A community member on GPUStack has deployed GLM-5.2-FP8-DSpark, an enhanced version of GLM-5.2-FP8 that incorporates speculative decoding with an external draft model from Red Hat AI. Performance tests yielded mixed res…
-
Zhipu AI's GLM-5.2 model deployed on serverless GPUs
Zhipu AI has released GLM-5.2, a 700B Mixture-of-Experts (MoE) model that excels in complex reasoning and software engineering tasks, reportedly matching or surpassing proprietary models like Claude 3.5 Sonnet and GPT-4…
-
GLM-5.2-FP8 deployed with 262k context on HGX-H200
A user shared their Docker deployment configuration for GLM-5.2-FP8 on an HGX-H200 system using SGLang. The configuration achieved a 262k context window and a throughput of 70 tokens per second. The user noted that cert…