zai-org/GLM-5.2-FP8
PulseAugur coverage of zai-org/GLM-5.2-FP8 — every cluster mentioning zai-org/GLM-5.2-FP8 across labs, papers, and developer communities, ranked by signal.
-
GLM-5.2 model updated for faster inference with colibri engine
A new version of the GLM-5.2 model, named "colibri int4 with int8 mtp", has been released on Hugging Face. This iteration is based on the original GLM-5.2 model and features int8 MTP heads designed to significantly boos…
-
GLM-5.2 model quantized for consumer hardware via Colibri engine
A quantized version of the GLM-5.2 model, named jlnsrk/GLM-5.2-colibri-int4, has been released on Hugging Face. This version is designed to run on consumer hardware by streaming experts from disk, requiring approximatel…
-
GLM-5.2 model speed boosted over 20x via custom hacks
A Reddit user detailed a method for significantly accelerating the GLM-5.2 large language model on a specialized GH200 system. By combining components from different repositories and patching the vLLM inference engine, …
-
Z.ai launches GLM-5.2 with 1M context and open-source license
Z.ai has released GLM-5.2, a new flagship model designed for long-horizon tasks, featuring a 1 million token context window. The model boasts improved coding capabilities with adjustable effort levels and an enhanced ar…