GGUFs
PulseAugur coverage of GGUFs — every cluster mentioning GGUFs across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Ornith 1.5 35B GGUF model receives subtle update with changed weights
A user on Reddit's r/LocalLLaMA subreddit has identified a subtle update to the Ornith 1.5 35B GGUF model files. The user discovered that the weights themselves have changed, along with the importance matrix and calibra…
-
Unsloth releases Dynamic 3.0 GGUFs for optimized LLM performance
Unsloth has released Dynamic 3.0 GGUFs, an update to their model optimization tools. This release focuses on enhancing the efficiency and performance of large language models. The GGUF format is widely used for running …
-
Unsloth releases Dynamic v3.0 quants for Qwen3.8-27B, boosting accuracy
Unsloth has released Dynamic v3.0 quantization for Qwen3.8-27B GGUFs, offering over 10% improved accuracy at the same model size compared to other providers. This new version utilizes an improved methodology with a high…
-
Alibaba's Qwen3.8-27B model released; AI aids GPU porting; LLM infra detailed
Alibaba's Qwen team has released Qwen3.8-27B, a dense 27-billion parameter model that fits on a single GPU and supports a 1 million token context window, with Day-0 integration in vLLM. Concurrently, research is explori…
-
Hy3 LLM integrated into llama.cpp with GGUF support
The Hy3 large language model has been released and is already being integrated into the llama.cpp project through a new pull request. Early users are reporting successful generation of coherent output, with performance …
-
Unsloth Studio boosts context length by 3x with GLM 5.2 support
Unsloth Studio has released version 0.1.47-beta, introducing support for GLM 5.2 GGUFs and an improved auto-fit algorithm that enables three times longer context lengths. This update also brings enhanced features such a…
-
Command A Plus language model integrated into llama.cpp
The Command A Plus language model has been integrated into llama.cpp, a popular inference engine for large language models. This update also includes support for North Mini Code. While GGUF quantized versions for North …
-
Qwen3.5 27B and 35B uncensored models released
The user LLMFan46 has released two uncensored versions of the Qwen3.5 model: a 27B parameter model and a 35B parameter model named A3B. These models are available in various formats including Safetensors, GGUFs, NVFP4, …