PulseAugur
实时 23:56:45

Unsloth 发布优化的 DeepSeek-V4-Flash-0731 GGUF 模型以供本地使用

Unsloth 发布了 DeepSeek-V4-Flash-0731 模型 GGUF 格式的优化版本,使其更易于在本地运行。这些模型与 llama.cppOllamaLM StudioUnsloth Studio 等各种流行推理工具兼容。性能基准测试表明,该模型可以在具有 40GB VRAM 的硬件上高效运行,生成速度超过每秒 16 个 token。 AI

影响 促进了更广泛的本地部署和对先进语言模型的实验。

排序理由 此次发布的是用于本地使用和现有推理工具的优化模型格式,而不是主要实验室发布的新前沿模型。

在 Hugging Face Trending Models 阅读 →

AI 生成摘要 · Google Gemini · 来自 3 个来源。 我们如何撰写摘要 →

Unsloth 发布优化的 DeepSeek-V4-Flash-0731 GGUF 模型以供本地使用

报道来源 [3]

  1. Hugging Face Trending Models TIER_1 English(EN) · unsloth ·

    unsloth/DeepSeek-V4-Flash-0731-GGUF

    0 downloads · 66 likes

  2. r/LocalLLaMA TIER_1 English(EN) · /u/Different-Pickle1021 ·

    DeepSeek-V4-Flash-0731 unsloth gguf on A100

    <!-- SC_OFF --><div class="md"><p>A100 with 40gb VRAM:</p> <ul> <li>162GB Q8_K_XL</li> <li>~16.1 tok/s generation</li> <li>Only 15.8GB of 40GB VRAM used with all experts on CPU</li> </ul> <p>NOTE just tested coding on linux box DeepSeek-V4-Flash-0731 runs losslessly on the single…

  3. r/LocalLLaMA TIER_1 English(EN) · /u/BlackBeardAI ·

    Unsloth Deepseek V4 0731 GGUF's are UP!

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbtdok/unsloth_deepseek_v4_0731_ggufs_are_up/"> <img alt="Unsloth Deepseek V4 0731 GGUF's are UP!" src="https://external-preview.redd.it/wMZUsudfTHG534WgBRWwhtsK20jasRrxNGwenbWxNxM.png?width=640&amp;crop=smar…