GLM 5.2 NVFP4
PulseAugur coverage of GLM 5.2 NVFP4 — every cluster mentioning GLM 5.2 NVFP4 across labs, papers, and developer communities, ranked by signal.
-
Kimi K3 and Inkling launch, intensifying open-model competition
The AI landscape is seeing intense competition, particularly with the release of Moonshot AI's Kimi K3, an open-weight model that rivals frontier-class closed models in coding and agentic tasks. This development is prom…
-
NVIDIA GLM-5.2-NVFP4 enables local AI on consumer hardware; Hermes agent guide updated
NVIDIA's GLM-5.2-NVFP4, a 4-bit FP4 quantized model, enables running large GLM-5 models on consumer hardware, marking a significant advancement for local AI computing and making advanced text generation more accessible …
-
GLM-5.2 NVFP4 achieves 24 tok/s at 128K context after bug fix · 1 source tracked
A user has resolved an issue with the GLM-5.2 NVFP4 model running on four DGX Sparks, achieving approximately 24 tokens per second at a 128K context length. The problem involved a bug in the speculative decoding configu…
-
NVIDIA releases quantized GLM-5.2 and MiniMax-M3 models
NVIDIA has released two new quantized text-generation models: GLM-5.2-NVFP4 and MiniMax-M3-NVFP4. The GLM-5.2-NVFP4 model, based on ZAI's GLM-5.2, is MIT-licensed and available for both commercial and non-commercial glo…