PulseAugur
中
实时 22:30:29
English(EN) Qwen3.8-27B: 159 tok/s on R9700, 64 tok/s on Strix Halo

Qwen3.8-27B 模型通过 iPad IDE 在 R9700 GPU 上实现每秒 159 个 token

LemonSeed Studio 是一个基于 iPad 的编辑器和 IDE,展示了使用 AMD GPU 的设备端推理能力。该系统在 R9700 GPU 上使用 Qwen3.8-27B 模型实现了每秒 159 个 token 的速度,在 Strix Halo GPU 上实现了每秒 64 个 token 的速度。该引擎 LSE 通过融合操作并为各种 GPU 生成自定义内核来优化模型性能,支持投机解码等功能以实现更快的推理。 AI

影响 展示了在消费级硬件上 LLM 的设备端推理能力的提升,可能为更强大的移动 AI 应用提供支持。

排序理由 这是特定软件工具在特定模型和硬件上的能力演示,而不是新的模型发布或重大的行业事件。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Qwen3.8-27B 模型通过 iPad IDE 在 R9700 GPU 上实现每秒 159 个 token

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是特定软件工具在特定模型和硬件上的能力演示,而不是新的模型发布或重大的行业事件。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
2 days old
Coverage has settled into its steady-state source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/TheOriginalG2 ·

    Qwen3.8-27B: R9700上159 tok/s,Strix Halo上64 tok/s

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1x18e95/qwen3827b_159_toks_on_r9700_64_toks_on_strix_halo/"> <img alt="Qwen3.8-27B: 159 tok/s on R9700, 64 tok/s on Strix Halo" src="https://preview.redd.it/3pwkx9zjfcuh1.jpg?width=140&amp;height=105&amp;auto=…