PulseAugur
中
实时 20:00:21
(AF) b11528: meta : handle host views (#30217)

llama.cpp b11528 发布修复 KV 缓存视图问题

llama.cpp 项目已发布 b11528 版本,其中包括对处理张量主机视图的修复。此更改由 Qwen3.8 Flash-Next 协助,解决了使用部分卸载进行张量拆分时 KV 缓存视图出现的断言错误。此次更新还重新启用了 K2 Horizon 的 SM 张量支持,并添加了一个 TODO 引用。 AI

影响 提高 llama.cpp 推理引擎用户的性能和稳定性,特别是那些使用部分卸载和 KV 缓存的用户。

排序理由 这是一个特定工具的软件发布,而不是前沿模型发布或重要的行业事件。

在 llama.cpp — Releases 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

llama.cpp b11528 发布修复 KV 缓存视图问题

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是一个特定工具的软件发布,而不是前沿模型发布或重要的行业事件。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
2 days old
Coverage has settled into its steady-state source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. llama.cpp — Releases TIER_1 (AF) · ggerganov ·

    b11528: meta: 处理主机视图 (#30217)

    <ul> <li>meta : handle views of tensors allocated on the host</li> </ul> <p>A view shares the memory of its view_src, so ggml-alloc never allocates a view in<br /> the buffer of the split it lands in - the scheduler copies the source into the<br /> split and the ops that use the …