PulseAugur
EN
LIVE 05:34:17

Ollama v0.32.4 enhances local AI with Apple MLX support and multimodal vision

Ollama has released version v0.32.4, introducing significant enhancements for local AI inference, particularly for users with Apple Silicon hardware. This update brings support for the Laguna model family via Apple's MLX engine, enabling accelerated inference on integrated GPUs. Additionally, the release refines speculative decoding by quantizing draft-model output heads for improved efficiency and accuracy, and fixes decoding issues for Qwen3 MoE models. The update also includes preliminary vision support for the Minimax-M3 model in llama.cpp, expanding its multimodal capabilities. AI

IMPACT Enhances local AI inference capabilities for Apple Silicon users and expands multimodal support in llama.cpp.

RANK_REASON This is a software release for a tool that facilitates local AI model inference, not a frontier model release or core research.

Read on Mastodon — fosstodon.org →

AI-generated summary · Google Gemini · from 6 sources. How we write summaries →

Ollama v0.32.4 enhances local AI with Apple MLX support and multimodal vision

COVERAGE [6]

  1. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🚀 Ollama v0.32.5 is out What's Changed · * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. · Full Changelog:… ⭐ 1

    🚀 Ollama v0.32.5 is out What's Changed · * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. · Full Changelog:… ⭐ 177,021 stars All open-source AI releases → https:// opensourceai.tech/releases.html # OpenSource # Release # DevTools # …

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.5 Release Notes: ## What's Changed * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particul

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.5 Release Notes: ## What's Changed * Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna. **Full Changelog**: https:// github.com/ollama/ollama/compa re/v0.32.4...v0.32.5 🔗 https:// github.com/olla…

  3. dev.to — LLM tag TIER_1 English(EN) · soy ·

    Ollama v0.32.4 Brings Apple MLX Support — Plus Rubin, llama.cpp & GPU Updates

    <p>Today's digest highlights Ollama v0.32.4 with new Apple MLX GPU support, alongside llama.cpp gaining preliminary vision capabilities. Additionally, we see releases for AMD ROCm 7.14.0 and TensorRT-LLM, plus details on NVIDIA's Rubin GPU architecture.</p> <h2> Local AI &amp; Op…

  4. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🚀 Ollama v0.32.4 is out What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating

    🚀 Ollama v0.32.4 is out What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating… ⭐ 176,886 stars All open-source AI releases → https:// opensourceai.tech/releases.html # OpenSource # Release # DevToo…

  5. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.4 Release Notes: ## What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output head

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.4 Release Notes: ## What's Changed - Support Laguna on Apple GPUs via the MLX engine - Quantize draft-model output heads at the requested type when creating speculative-decoding drafts. - Fixed Qwen3 MoE decoding for differently-quantized …

  6. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.3 Release Notes: ## What's Changed - Fixed model downloads that stall before sending data. - Improved integrations: res

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.3 Release Notes: ## What's Changed - Fixed model downloads that stall before sending data. - Improved integrations: restored Claude Code Channels, fixed Anthropic thinking streams, and made Hermes Desktop respect `--force-build`. - Expande…