IQ1_M
PulseAugur coverage of IQ1_M — every cluster mentioning IQ1_M across labs, papers, and developer communities, ranked by signal.
-
Unsloth trims Kimi K3 LLM to under 500GB, boosting performance
Unsloth has released several versions of the Kimi K3 large language model, significantly reducing their file sizes. One version, IQ2_XXS, was trimmed from 711GB to 478GB by removing multilingual capabilities and retaini…
-
User asks about IQ1_M 342GB Pruned Kimi K3 model usability
A user on the r/LocalLLaMA subreddit is inquiring about the performance and usability of the IQ1_M 342GB Pruned Kimi K3 model. The post includes a link to a Hugging Face repository for the model, suggesting it is a comm…
-
AngelSlim/Hy3-GGUF model now supports multiple AI tools and libraries
The AngelSlim/Hy3-GGUF model is now available for use with various popular AI tools and libraries. Instructions are provided for integrating it with llama-cpp-python, llama.cpp, vLLM, Ollama, Unsloth Studio, and Pi. The…
-
GLM-5.2 model runs at 7.3 tok/s locally with 4x RTX 3090s
A user has detailed their experience running the GLM-5.2 UD-IQ2_M model locally, achieving approximately 7.3 tokens per second across four RTX 3090 GPUs and 192GB of RAM. They found that halving the quantization level (…