PulseAugur
实时 18:51:07
Nederlands(NL) GLM 5.2 speeds

GLM-5.2 用户在自定义硬件上分享性能基准

一位Reddit用户正在分享他们在自定义系统上运行GLM-5.2的经验。他们实现了20k上下文和44 tokens/秒的摄入速度,以及8 tokens/秒的生成速度。该用户希望与社区中的其他人比较其性能和成本效益。 AI

排序理由 关于特定模型版本的用户生成内容,而非主要来源发布或重大行业事件。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

GLM-5.2 用户在自定义硬件上分享性能基准

报道来源 [1]

  1. r/LocalLLaMA TIER_1 Nederlands(NL) · /u/naunen ·

    GLM 5.2 speeds

    <!-- SC_OFF --><div class="md"><p>Tell me, to get on 20k context and ingestion 44tks, generation 8tks is good numbers for 4x 8880 v4 cpus, 1tb 32channels ddr3 ram and 2x 3060 12gb. ? </p> <p>im running 3bit version</p> <p>whole system cost 900$/€</p> <p>who can beat me on tks/cos…