PulseAugur
中
实时 19:40:40
English(EN) ⚙️ New Ollama Release! ⚙️ Version: v0.32.6 Release Notes: ## What's Changed - Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for

Ollama v0.32.6 提升 Qwen 3.5 在 Apple Silicon 上的速度,改进 OpenAI 兼容性 · 跟踪 4 个来源

Ollama 发布了 0.32.6 版本,通过 MLX 引擎和推测性解码显著提高了 Qwen 3.5 模型在 Apple Silicon Mac 上的性能。此次更新还通过对 /v1/chat/completions 端点的流式格式进行对齐,增强了与 OpenAI API 的兼容性。此外,KataGo 和 llama.cpp 等其他相关项目也收到了错误修复,KataGo 解决了 TensorRT 问题,llama.cpp 解决了 Vulkan 设备丢失错误。 AI

影响 提升了在 Apple Silicon 设备上运行 Qwen 3.5 的用户的本地 AI 推理性能和兼容性。

排序理由 这是本地 AI 推理工具的软件更新,而非主要实验室的前沿模型发布。

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 4 个来源。 我们如何撰写摘要 →

Ollama v0.32.6 提升 Qwen 3.5 在 Apple Silicon 上的速度,改进 OpenAI 兼容性 · 跟踪 4 个来源

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是本地 AI 推理工具的软件更新,而非主要实验室的前沿模型发布。
Source corroboration
4 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
61 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [4]

  1. dev.to — LLM tag TIER_1 English(EN) · soy ·

    Ollama v0.32.6 在 Apple Silicon 上加速 Qwen3.5 — 并支持 PyTorch、Mesa 26.2 及 GPU 修复

    <p>Today's digest brings significant updates for local AI, with Ollama v0.32.6 boosting Qwen3.5 performance on Apple Silicon and adding OpenAI streaming compatibility. Additionally, Mesa 26.2 adds NVK Mesh Shader support, llama.cpp fixes Vulkan errors, KataGo resolves critical Te…

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    ⚙️ Ollama 新版本发布!⚙️ 版本:v0.32.6 发布说明:## 变更内容 - Qwen3.5 在 Apple GPU 上速度更快:MLX 引擎现已使用模型的 MTP 头以

    ⚙️ New Ollama Release! ⚙️ Version: v0.32.6 Release Notes: ## What's Changed - Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically - `/v1/chat/completions` streaming now matches OpenAI's wire format: `role` only on t…

  3. dev.to — LLM tag TIER_1 English(EN) · LucioLiu ·

    Ollama 最新 RC 版本在 Apple GPU 上改进 Qwen3.5 并修复流式传输兼容性

    <p>Ollama v0.32.6-rc0 adds a useful path for people running Qwen3.5 on Apple GPUs: the MLX engine now uses the model's MTP head automatically for speculative decoding.</p> <p>The practical idea is straightforward. The MTP head predicts upcoming tokens, and the main model verifies…

  4. dev.to — LLM tag TIER_1 English(EN) · soy ·

    Ollama v0.32.6 提升 Qwen3.5 在 Apple GPU 上的性能 — 另有 AMD CDNA 5、FFmpeg 9.0 等

    <p>Ollama v0.32.6 ships with significant Qwen3.5 performance boosts for Apple GPUs, while AMD unveils its new CDNA 5 hardware and Helios Rackscale. This digest also covers the release of FFmpeg 9.0, NVIDIA's Alpamayo 2 Super for commercial use, and a trending Qwen3-VL model on Hu…