PulseAugur
实时 05:19:51
English(EN) R9V Update: now ~100 tok/s in TG on Qwen3.8 Flash Next IQ4_XS on x2 R9700 + 128GB RAM. Fixed crashes with n-gram SSD streaming, improved diagnostics, plus pinned images. Q4_K_XL now supported, 50 tok/s TG.

R9V 软件更新提高了本地 LLM 推理速度和稳定性

R9V 软件已更新,提高了本地大型语言模型推理的性能和稳定性。此次更新提高了速度,在特定硬件配置下使用 Qwen3.8 Flash Next IQ4_XS 模型可实现约每秒 100 个 token 的速度。它还解决了与 SSD 流式传输相关的稳定性问题,并支持 Q4_K_XL 量化,尽管速度较慢。 AI

影响 提高了在消费级硬件上运行本地 LLM 的用户的性能和稳定性。

排序理由 特定本地 LLM 推理工具的软件更新。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

R9V 软件更新提高了本地 LLM 推理速度和稳定性

本文如何被排名

Signal score
6 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
特定本地 LLM 推理工具的软件更新。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Public_Umpire_1099 ·

    R9V 更新:现在 TG 上使用 Qwen3.8 Flash Next IQ4_XS 在 x2 R9700 + 128GB RAM 上速度约为 100 tok/s。修复了 n-gram SSD 流式传输崩溃问题,改进了诊断,并固定了图像。现已支持 Q4_K_XL,速度为 50 tok/s TG。

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1wfroih/r9v_update_now_100_toks_in_tg_on_qwen38_flash/"> <img alt="R9V Update: now ~100 tok/s in TG on Qwen3.8 Flash Next IQ4_XS on x2 R9700 + 128GB RAM. Fixed crashes with n-gram SSD streaming, improved diagn…