PulseAugur
EN
LIVE 04:46:06

User tests CMP170HX GPUs for local LLM deployment

A user tested four CMP170HX graphics cards, each configured with 64GB of memory, totaling 256GB of VRAM. The tests focused on running various large language models, with results indicating that smaller models can fit entirely on a single card or run concurrently if their combined VRAM usage stays within the 64GB limit. The performance is comparable to NVIDIA's 30xx series, with potential for higher throughput if using cards with more memory. The user also noted that the PCIe Gen2 x4 interface did not significantly bottleneck performance, even when loading large models, suggesting a potential for scaling up to 16 cards in a single PC. AI

IMPACT Provides insights into the viability of using older or repurposed hardware for running large language models locally.

RANK_REASON User benchmark of consumer-grade hardware for LLM deployment.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

User tests CMP170HX GPUs for local LLM deployment

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/m94301 ·

    I tested the CMP170HX

    <!-- SC_OFF --><div class="md"><p>Lots of rumor and misinfo bouncing around, so I put some of these old mining cards to the test. I used 4 of the 8GB cards, set to 64GB each.</p> <p>Lots of models fit entirely on a single card, and you can also run several small models at the sam…