IQ1_M
PulseAugur coverage of IQ1_M — every cluster mentioning IQ1_M across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
Unsloth releases Kimi k3 model with multiple large-size versions
Unsloth has released a new model named Kimi k3, offering several versions with varying sizes. The models include Q1_0 at 466GB, TQ1_0 at 509GB, IQ1_M at 649GB, and TQ2_0 at 551GB. This release provides users with differ…
-
User asks about IQ1_M 342GB Pruned Kimi K3 model usability
A user on the r/LocalLLaMA subreddit is inquiring about the performance and usability of the IQ1_M 342GB Pruned Kimi K3 model. The post includes a link to a Hugging Face repository for the model, suggesting it is a comm…
-
AngelSlim/Hy3-GGUF model now supports multiple AI tools and libraries
The AngelSlim/Hy3-GGUF model is now available for use with various popular AI tools and libraries. Instructions are provided for integrating it with llama-cpp-python, llama.cpp, vLLM, Ollama, Unsloth Studio, and Pi. The…
-
GLM-5.2 model runs at 7.3 tok/s locally with 4x RTX 3090s
A user has detailed their experience running the GLM-5.2 UD-IQ2_M model locally, achieving approximately 7.3 tokens per second across four RTX 3090 GPUs and 192GB of RAM. They found that halving the quantization level (…