PulseAugur
实时 19:49:46
English(EN) DeepSeek-V4-Flash-0731-UD-Q3_K_XL 3x3090 test results

DeepSeek-V4-Flash 模型在 3x RTX 3090 GPU 上进行基准测试

一位用户分享了 DeepSeek-V4-Flash-0731-UD-Q3_K_XL 模型在配备三块 NVIDIA RTX 3090 GPU 的设备上的基准测试结果。结果显示,pp512 测试的每秒 token 数为 116.04,tg128 测试的每秒 token 数为 7.71。用户指出可能还有进一步优化的空间,但尚未取得更好的性能。 AI

影响 提供了特定模型配置的性能数据,对拥有类似硬件的用户有帮助。

排序理由 用户生成的特定模型量化基准测试结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

DeepSeek-V4-Flash 模型在 3x RTX 3090 GPU 上进行基准测试

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/consultkitapp ·

    DeepSeek-V4-Flash-0731-UD-Q3_K_XL 3x3090 test results

    <!-- SC_OFF --><div class="md"><p>For anyone interested, here are the llama-bench results on 3 bit K_XL quantization. I think this could be pushed further but no luck so far.</p> <h1>Command</h1> <p>./llama-bench -m /home/user/llamacpp/modelsmain/unsloth/ds4/DeepSeek-V4-Flash-073…