PulseAugur
EN
LIVE 14:34:38

Ollama v0.32.4 enhances local AI with Apple MLX support and multimodal vision

Ollama has released version v0.32.4, introducing significant enhancements for local AI inference, particularly for users with Apple Silicon hardware. This update brings support for the Laguna model family via Apple's MLX engine, enabling accelerated inference on integrated GPUs. Additionally, the release refines speculative decoding by quantizing draft-model output heads for improved efficiency and accuracy, and fixes decoding issues for Qwen3 MoE models. The update also includes preliminary vision support for the Minimax-M3 model in llama.cpp, expanding its multimodal capabilities. AI

IMPACT Enhances local AI inference capabilities for Apple Silicon users and expands multimodal support in llama.cpp.

RANK_REASON This is a software release for a tool that facilitates local AI model inference, not a frontier model release or core research.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 6 sources. How we write summaries →

Ollama v0.32.4 enhances local AI with Apple MLX support and multimodal vision

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
This is a software release for a tool that facilitates local AI model inference, not a frontier model release or core research.
Source corroboration
6 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
59 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+2 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [6]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🚀 Ollama v0.32.5 is out What's Changed · * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. · Full Changelog:… ⭐ 1

    🚀 Ollama v0.32.5 is out What's Changed · * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. · Full Changelog:… ⭐ 177,021 stars All open-source AI releases → https:// opensourceai.tech/releases.html # OpenSource # Release # DevTools # …

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.5 Release Notes: ## What's Changed * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particul

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.5 Release Notes: ## What's Changed * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. **Full Changelog**: https:// github.com/ollama/ollama/compa re/v0.32.4...v0.32.5 🔗 https:// github.com/olla…

  3. dev.to — LLM tag TIER_1 English(EN) · soy ·

    Ollama v0.32.4 Brings Apple MLX Support — Plus Rubin, llama.cpp & GPU Updates

    <p>Today's digest highlights Ollama v0.32.4 with new Apple MLX GPU support, alongside llama.cpp gaining preliminary vision capabilities. Additionally, we see releases for AMD ROCm 7.14.0 and TensorRT-LLM, plus details on NVIDIA's Rubin GPU architecture.</p> <h2> Local AI &amp; Op…

  4. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🚀 Ollama v0.32.4 is out What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating

    🚀 Ollama v0.32.4 is out What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating… ⭐ 176,886 stars All open-source AI releases → https:// opensourceai.tech/releases.html # OpenSource # Release # DevToo…

  5. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.4 Release Notes: ## What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output head

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.4 Release Notes: ## What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating speculative-decoding drafts. - Fixed Qwen3 MoE decoding for differently-quantized …

  6. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.3 Release Notes: ## What's Changed - Fixed model downloads that stall before sending data. - Improved integrations: res

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.3 Release Notes: ## What's Changed - Fixed model downloads that stall before sending data. - Improved integrations: restored Claude Code Channels, fixed Anthropic thinking streams, and made Hermes Desktop respect `--force-build`. - Expande…