PulseAugur
EN
LIVE 18:50:59

GLM-5.2 user shares performance benchmarks on custom hardware

A user on Reddit is sharing their experience running GLM-5.2 on a custom-built system. They are achieving 20k context and ingestion speeds of 44 tokens/sec, with generation speeds of 8 tokens/sec. The user is seeking to compare their performance and cost-effectiveness against others in the community. AI

RANK_REASON User-generated content on a specific model version, not a primary source release or significant industry event.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

GLM-5.2 user shares performance benchmarks on custom hardware

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 Nederlands(NL) · /u/naunen ·

    GLM 5.2 speeds

    <!-- SC_OFF --><div class="md"><p>Tell me, to get on 20k context and ingestion 44tks, generation 8tks is good numbers for 4x 8880 v4 cpus, 1tb 32channels ddr3 ram and 2x 3060 12gb. ? </p> <p>im running 3bit version</p> <p>whole system cost 900$/€</p> <p>who can beat me on tks/cos…