PulseAugur
实时 17:56:28
中文(ZH) DeepSeek低价风暴打服硅谷!海外平台争相倒贴V4 Flash

DeepSeek V4 Flash以低成本、高性能发布挑战顶级AI模型 · 追踪10个来源

DeepSeek发布了其V4 Flash模型,其性能可与OpenAI的GPT-5.6 Luna和Anthropic的Claude Opus 4.8等顶级模型相媲美,但成本却显著降低。这款新模型,特别是0731更新,由于重新进行了后训练,在智能体和编码基准测试中取得了 substantial 提升,使其成为寻求成本效益AI解决方案的开发人员和企业的极具吸引力的选择。V4 Flash的经济实惠性和强大性能正在重塑开发者的默认选择,并被视为中国AI发展的一个关键时刻,有可能使先进AI功能的访问民主化。 AI

影响 通过使高性能模型对开发人员和企业来说更易于获得和负担得起,加速了先进AI的采用。

排序理由 前沿实验室模型发布,包含性能和定价细节。

在 量子位 (QbitAI) 阅读 →

AI 生成摘要 · Google Gemini · 来自 113 个来源。 我们如何撰写摘要 →

DeepSeek V4 Flash以低成本、高性能发布挑战顶级AI模型 · 追踪10个来源

报道来源 [113]

  1. Unsloth — Releases TIER_1 English(EN) · danielhanchen ·

    更快的下载 + DeepSeek-V4 Flash 0731

    <p>Hey everyone! For folks who missed the news - Kimi K3 &amp; DeepSeek v4 Flash can now run locally with Unsloth Dynamic GGUFs! We now added more efficient and faster downloading for Colab, low memory systems and also high memory and CPU systems - we auto fallback to HTTP as wel…

  2. Unsloth — Releases TIER_1 English(EN) · danielhanchen ·

    更快的下载 + DeepSeek-V4 Flash 0731

    <p>Hey everyone! For folks who missed the news - Kimi K3 &amp; DeepSeek v4 Flash can now run locally with Unsloth Dynamic GGUFs! We now added more efficient and faster downloading for Colab, low memory systems and also high memory and CPU systems - we auto fallback to HTTP as wel…

  3. Unsloth — Releases TIER_1 English(EN) · danielhanchen ·

    更快的下载 + DeepSeek-V4 Flash 0731

    <p>Hey everyone! For folks who missed the news - Kimi K3 &amp; DeepSeek v4 Flash can now run locally with Unsloth Dynamic GGUFs! We now added more efficient and faster downloading for Colab, low memory systems and also high memory and CPU systems - we auto fallback to HTTP as wel…

  4. 量子位 (QbitAI) TIER_1 中文(ZH) · Jay ·

    DeepSeek低价风暴席卷硅谷!海外平台争相补贴V4 Flash

    那还说啥了梁圣,我订阅费全给你就是了呗

  5. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    DeepSeek V4 Flash 0731 现已在 Together AI 上开放微调。

    DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for production inference. https://t.co/MFkqLlOKAj https://t.co/TpTihq8vic

  6. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    触手可及的 Frontier 性能。很荣幸能为 Ollama 云上的 DeepSeek-V4-Flash 提供支持,提供最快的托管性能。开放权重,

    Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. &amp; EU hosting.

  7. X — SemiAnalysis TIER_1 English(EN) · SemiAnalysis_ ·

    祝贺 @deepseek_ai 发布 DeepSeekv4 Flash 0731 🔥 它在智能体任务上大幅超越 Nemotron3 Ultra,同时激活参数却少了 4.2 倍

    Congrats to @deepseek_ai on their release of DeepSeekv4 Flash 0731 🔥 It massively beats Nemotron3 Ultra on agentic tasks while having 4.2x fewer active parameters and close to 2x fewer total parameters! Committee-based model frontier development does not work. A focused team is …

  8. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    DeepSeek V4 Flash 0728 确实是当前性价比的顶峰!

    DeepSeek V4 Flash 0728 really is the current cost-performance frontier!

  9. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    我们使用 DeepSWE 分析了 DeepSeek-V4 Flash-0731 与 GPT-5.6 Luna 在软件工程任务上的表现。

    We analyzed DeepSeek-V4 Flash-0731 vs. GPT-5.6 Luna on software engineering tasks using DeepSWE. DeepSeek-V4 Flash-0731 delivers 80% of Luna’s performance at roughly 1/6 the cost. More insights in the thread! 👇

  10. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    🤯 DeepSeek 报告 V4 Flash 在 Terminal-Bench 2.1 上得分 82.7,领先 V4 Pro Preview 的 72.1,尽管参数总量仅使用了约五分之一。https://t.co/S

    🤯 DeepSeek reports V4 Flash at 82.7 on Terminal-Bench 2.1, ahead of V4 Pro Preview at 72.1, despite using roughly one-fifth the total parameters. https://t.co/SGRiiKCcu3

  11. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    RT @zainhas: DeepSeek-V4 Flash 对比 GPT 5.6 Luna 太疯狂了!https://t.co/HhTiVHnbgf

    RT @zainhas: DeepSeek-V4 Flash vs. GPT 5.6 Luna is crazy! https://t.co/HhTiVHnbgf

  12. 36氪 (36Kr) TIER_1 中文(ZH) ·

    DeepSeek-V4-Flash 官方版 API 登陆国家超算互联网

    36氪获悉,近日,DeepSeek-V4-Flash正式版API面向公众开启公测,国家超算互联网第一时间同步上线DeepSeek-V4-Flash模型API调用与下载服务。

  13. TLDR AI TIER_1 English(EN) · TLDR ·

    DeepSeek V4 Flash ⚡,OpenAI 数学突破 🔢,Qwen 3.8-Max 🤖

  14. 36氪 (36Kr) TIER_1 中文(ZH) ·

    DeepSeek:DeepSeek-V4-Pro 正式版即将发布

    36氪获悉,DeepSeek发布更新日志,DeepSeek-V4-Flash正式版API上线公测。DeepSeek-V4-Flash-0731的模型结构、尺寸和DeepSeek-V4-Flash-preview保持一致,仅重新进行了后训练。DeepSeek-V4-Pro正式版将会尽快发布。

  15. The Decoder TIER_1 English(EN) · Thomas Joos ·

    Deepseek新Flash模型成本低约60% 媲美OpenAI的GPT-5.6 Luna

    <p><img alt="" class="attachment-full size-full wp-post-image" height="1152" src="https://the-decoder.com/wp-content/uploads/2026/06/deepseek_red_whale.png" style="height: auto; margin-bottom: 10px;" width="2048" /></p> <p> Deepseek's budget model V4 Flash gets a major boost with…

  16. Pandaily TIER_1 English(EN) · [email protected] (Pandaily) ·

    DeepSeek V4-Flash 登顶全球代币消耗榜:中国模型包揽 OpenRouter 前五,性价比“杀手线”重塑开发者默认选项

    DeepSeek V4-Flash leads OpenRouter weekly rankings with 7.1 trillion tokens, Chinese models claim nine of the global top 10 slots, and global weekly AI token usage crosses 56.8 trillion.

  17. Pandaily TIER_1 English(EN) · [email protected] (Pandaily) ·

    DeepSeek-V4-Flash 正式版发布并开源:304B 轻量模型超越 V4-Pro 预览版,媲美 Claude Opus-4.8,被誉为 DeepSeek 第三个里程碑

    DeepSeek-V4-Flash-0731 official release outperforms V4-Pro preview with 82.7 performance score at 0.14/0.28 USD pricing, tops VulcanBench, and reaches HuggingFace trending second.

  18. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    DeepSeek 升级 DeepSeek-V4-Flash-0731,在 Agentic 和编码方面取得重大进展

    <p>DeepSeek published DeepSeek-V4-Flash-0731 on Hugging Face and moved the official V4-Flash API into public beta on July 31, 2026. The model card is explicit that this is the official release superseding the preview, and that the architecture and size are unchanged. The gains co…

  19. Towards AI TIER_1 English(EN) · Caspar Bannink - AI Engineer ·

    Deepseek 再创佳绩,V4 Flash GA 是最具性价比的模型

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://pub.towardsai.net/deepseek-did-it-again-v4-flash-ga-is-the-best-model-per-dollar-cae8661d962e?source=rss----98111c9905da---4"><img src="https://cdn-images-1.medium.com/max/1400/1*hgYzGY5jaXpgnUMfoZYVNw.jp…

  20. Towards AI TIER_1 English(EN) · allglenn ·

    DeepSeek V4 Flash 0731 成本降低 5 倍,性能超越 V4 Pro

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://pub.towardsai.net/deepseek-v4-flash-0731-outscores-v4-pro-at-5x-lower-cost-cd81d817e982?source=rss----98111c9905da---4"><img src="https://cdn-images-1.medium.com/max/1536/1*d2asC8IzuphoDaVbG9QEmw.png" wid…

  21. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    一个拥有3040亿参数的开源模型刚刚追平了Claude Opus-4.8。北京的DeepSeek正式发布V4-Flash,标志着中国AI效率的巨大飞跃

    A 304B parameter open-source model just matched Claude Opus-4.8. Beijing-based DeepSeek’s official release of V4-Flash marks a massive leap in Chinese AI efficiency, outperforming its own V4-Pro preview. This "third DeepSeek moment" gives global developers a formidable, highly ef…

  22. Medium — Claude tag TIER_1 English(EN) · Joe Njenga ·

    我试用了 DeepSeek V4-Flash 运行 Claude Code(成本降低 71 倍,性能超越 Fable 5)

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/@joe.njenga/i-tried-deepseek-v4-flash-on-claude-code-it-beats-fable-5-at-71x-lower-cost-fa5b74669412?source=rss------claude-5"><img src="https://cdn-images-1.medium.com/max/1280/1*mklI1uM1G6qwV…

  23. Mastodon — sigmoid.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @u1tra_instinct: 🚨🚨🚨🚨🚨: 自 DeepSeek-V4 Flash GA 于 7 月 31 日发布以来,呼声最高。昨天。现在已完全淘汰 32/32,100% 兼容 m

    RT @u1tra_instinct: 🚨🚨🚨🚨🚨: am meisten angefragt seit der DeepSeekV4-Flash-GA-Veröffentlichung vom 31.07. Gestern. Jetzt abliteriert 32/32, zu 100 % kompatibel mit DSpark. Wenn euch meine Arbeit gefällt und ihr beitragen möchtet, könnt ihr mir gerne über X-Money oder GoFundMe (Lin…

  24. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash 0731之所以有趣,不是因为它变大了,而是因为它通过后期训练变得强大得多。相同的架构现在...

    DeepSeek V4 Flash 0731 is interesting not because it became larger, but because it became much more capable through post-training. The same architecture now delivers stronger coding-agent and tool-use performance while remaining remarkably inexpensive through the API. https:// op…

  25. Medium — Claude tag TIER_1 English(EN) · Chimin ·

    DeepSeek-V4-Flash 公测:Claude Code 和 Codex 都能用,价格便宜到离谱…

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://githubdaily.medium.com/deepseek-v4-flash-public-beta-claude-code-and-codex-can-both-be-used-at-dirt-cheap-prices-29a955104c96?source=rss------claude-5"><img src="https://cdn-images-1.medium.com/max/1742/1…

  26. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    DeepSeek 升级 DeepSeek-V4-Flash-0731,在 Agentic 和编码方面取得重大进展 #AgenticAI #AgenticArtificialIntelligence#

    https://www. europesays.com/3166986/ DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains # AgenticAI # AgenticArtificialIntelligence # AI # ArtificialIntelligence

  27. Medium — Claude tag TIER_1 English(EN) · inprogrammer ·

    DeepSeek V4 Flash 比 Claude 便宜 30 倍。我还是换回去了。

    <div class="medium-feed-item"><p class="medium-feed-image"><a href="https://medium.com/data-science-collective/deepseek-v4-flash-is-30x-cheaper-than-claude-i-still-switched-back-2ac429bf4e14?source=rss------claude-5"><img src="https://cdn-images-1.medium.com/max/1024/1*ZKNqc_ARJ4…

  28. Towards AI TIER_1 Nederlands(NL) · Mia Efoxtech ·

    DeepSeek V4 对比 DeepSeek V4 Flash:开发者在 2026 年应选择哪个模型?

    <p>Choose DeepSeek V4 Flash for high-volume, latency-sensitive, and cost-controlled workloads; choose DeepSeek V4 Pro for difficult reasoning, coding, research, and long-horizon agents. Both offer a 1M-token context window, 384K maximum output, thinking and non-thinking modes, JS…

  29. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @jun_song: SuperDeepseek-V4-Flash 在 2xDGX Spark 上运行,峰值性能为每秒 122 个 token。该实现基于 R

    RT @jun_song: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einer Spitzenleistung von 122 Token pro Sekunde. Die Implementierung erfolgte basierend auf dem Rezept von @MiaAIlab, um DFlash MTP zu nutzen. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Das Modell wurd…

  30. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @pupposandro: Lucebox Engine 现已在 128 GB AMD Strix Halo 系统上,从单个 98.29 GB GGUF 运行 DeepSeek V4 Flash 0731。该模型

    RT @pupposandro: Lucebox Engine betreibt nun DeepSeek V4 Flash 0731 von einem einzelnen 98,29 GB großen GGUF auf einem 128 GB AMD Strix Halo System. Das Modell erreicht 82/92 Punkte auf ds4-eval-92 und erzielt 32,7 Token pro Sekunde mit DSpark. Über unsere festen Evaluierungen Hu…

  31. r/LocalLLaMA TIER_1 English(EN) · /u/FantasticNature7590 ·

    我用一块RTX PRO 6000运行了DeepSeek V4 Flash 284B + DSpark。草稿生成器在RAM中的速度比VRAM快。

    <!-- SC_OFF --><div class="md"><p>Hey guys,</p> <p>Just finished benchmarking <strong>DeepSeek V4 Flash 284B + DSpark on a single RTX PRO 6000 96GB</strong>.</p> <p>Short version:</p> <ul> <li><strong>DSpark: ~15–17% faster generation</strong> on my coding workload</li> <li><stro…

  32. r/LocalLLaMA TIER_1 English(EN) · /u/stereohype ·

    DeepSeek V4 Flash 0731 在 Strix Halo 上以 27+ t/s 解码 — Vulkan + DSpark 全指南

    <!-- SC_OFF --><div class="md"><p>Been benchmarking DSv4 Flash 0731 on a Flow Z13 (Ryzen AI MAX+ 395, Radeon 8060S / gfx1151, 128GB LPDDR5X) for the past week. Figured I'd share what actually works and what doesn't — there are a lot of gotchas on this hardware.</p> <h2>Results</h…

  33. r/LocalLLaMA TIER_1 English(EN) · /u/Striking-Swim6702 ·

    DeepSeek-V4-Flash-0731 (284B MoE) 在 2× DGX Spark 上达到 75 tok/s — 全套配置,11 个陷阱,防重启集群,Codex CLI 集成

    <!-- SC_OFF --><div class="md"><p>Spent two nights getting <code>deepseek-ai/DeepSeek-V4-Flash-0731</code> (284B MoE, 13B active, native FP4/FP8, 1M context) running production-grade on two DGX Sparks connected by one QSFP DAC cable. Everything — scripts, tuning data, raw benchma…

  34. r/LocalLLaMA TIER_1 English(EN) · /u/Porespellar ·

    DeepSeek V4 Flash 0731 是将大卖特卖 DGX Sparks 的‘杀手级应用’

    <!-- SC_OFF --><div class="md"><p>Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getting a lot of people to buy a couple of NVIDIA G…

  35. r/LocalLLaMA TIER_1 English(EN) · /u/Exciting-Camera3226 ·

    DeepSeek V4 Flash 0731 在独立公开评测中(445次试验)以82.7%的成绩登上Terminal-Bench 2.1榜首

    <!-- SC_OFF --><div class="md"><p>Disclosure: I’m the author of Ante.</p> <p>DeepSeek recently reported an 82.7% score on Terminal-Bench 2.1 for DeepSeek V4 Flash 0731. Its evaluation used “DeepSeek Harness minimal mode,” which hasn’t been released yet.</p> <p>We wanted to see wh…

  36. r/LocalLLaMA TIER_1 English(EN) · /u/corruptbytes ·

    更新基准:Deepseek V4 Flash 在 SlopCodeBench (本地) 上表现

    <!-- SC_OFF --><div class="md"><p>Howdy - I posted a benchmark here - <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbtiy7/deepseek_v4_flash_on_slopcodebench/">https://www.reddit.com/r/LocalLLaMA/comments/1vbtiy7/deepseek_v4_flash_on_slopcodebench/</a></p> <p>This was us…

  37. r/LocalLLaMA TIER_1 English(EN) · /u/nomorebuttsplz ·

    我应该在 deepseek flash 0731 的 quants 上运行哪些相对快速的公开基准测试?

    <!-- SC_OFF --><div class="md"><p>I have various quants of this model and am curious how they perform. can anyone recommend which benchmark would be a good test case for quantization effects? Maybe that can be completed with about 1 million tokens?</p> </div><!-- SC_ON --> &#32; …

  38. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash 价格飙升 56%,SGLang 新增 Kimi K3 支持,同一小时内的两个信号可能改变您的部署计算。https:// olud.ai

    DeepSeek V4 Flash just jumped 56% in price as SGLang adds Kimi K3 support, two signals from the same hour that can change your deployment math. https:// olud.ai/news/2026-08-08.html # AI # OpenSource # AINews

  39. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash 以低廉价格逼近前沿基准,OpenAI 因担忧自主黑客攻击而暂停其 Astra 模型,以及内存芯片短缺现已成为现实

    DeepSeek V4 Flash closes in on frontier benchmarks at bargain prices, OpenAI pauses its Astra model over autonomous hacking fears, and memory chip shortages now extend through 2027. https:// ai0.news/posts/2026-08-08-dail y-digest/ # AI # Cybersecurity # OpenAI # AiPolicy

  40. r/LocalLLaMA TIER_1 English(EN) · /u/koibKop4 ·

    DeepSeek V4 Flash 0731 赏析文

    <!-- SC_OFF --><div class="md"><p>I’m running DSV4F 0731 on dual spark, and honestly… wow. It’s an absolute workhorse, and the benchmarks are real.</p> <p>Everyday tasks with Hermes agent? Effortless.</p> <p>Coding tasks with OpenCode? I’m genuinely amazed at what it can handle. …

  41. r/LocalLLaMA TIER_1 English(EN) · /u/kuhunaxeyive ·

    有人发现 DeepSeek-V4-Flash 在非编码任务上不可靠吗?

    <!-- SC_OFF --><div class="md"><p><em>(I am not a native speaker, written by myself, so please bear with me)</em></p> <p>I really want to like DeepSeek-V4-Flash-0731. But it has serious flaws that don't align with the high score on intelligence benchmarks. And those flaws render …

  42. r/LocalLLaMA TIER_1 Nederlands(NL) · /u/johnnyApplePRNG ·

    DeepSeek V4 Flash 0731 - ARC-AGI 结果

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vi9zls/deepseek_v4_flash_0731_arcagi_results/"> <img alt="DeepSeek V4 Flash 0731 - ARC-AGI Results" src="https://external-preview.redd.it/sWhpb1GjRlbd3knWV_xC1C2WMQX5RRFImjvBgSF_7ZI.png?width=640&amp;crop=sma…

  43. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    MIT许可的开源模型DeepSeek-V4-Flash-0731,30天内下载量达61.79万次。对于没有API限制的模型来说,这才是真正的吸引力

    DeepSeek-V4-Flash-0731, an MIT-licensed open-weight model, pulled 617.9k downloads in 30 days. That's real traction for a model that's not locked behind any API. Click to see why it's climbing. https:// olud.ai/#leaderboard # OpenSource # AI # HuggingFace

  44. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @UnslothAI: DeepSeek-V4-Flash现可在本地运行,速度提升2倍!⚡️ DSpark使V4-Flash-0731 GGUFs运行速度提高约1.4–2倍

    RT @UnslothAI: DeepSeek-V4-Flash kann jetzt mit DSpark 2× schneller lokal ausgeführt werden! ⚡️ DSpark ermöglicht es V4-Flash-0731 GGUFs, ~1,4–2× schneller zu generieren, ohne Genauigkeitsverlust. DeepSeek-V4-Flash-0731 erreicht bis zu 120 Token/s. GGUFs: https:// huggingface.co/…

  45. r/LocalLLaMA TIER_1 English(EN) · /u/Ok_Ninja7526 ·

    最终优化:在 128K 上下文的 1 块 RTX 3090 上,DeepSeek-V4-Flash-0731 的速度从约 10 tok/s 提升至约 15 tok/s

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vh1qn3/final_optimization_from_10_toks_to_15_toks_on/"> <img alt="Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090" src="https://preview.redd.it/tj77v44nqqhh1…

  46. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    📦 DeepSeek V4 Flash 0731 登陆排行榜,开放权重:1M 上下文,每百万 token 输入 0.09 美元 / 输出 0.18 美元。长上下文任务的新选择,无需高昂成本

    📦 DeepSeek V4 Flash 0731 lands on the leaderboard with open weights: 1M context at $0.09 in / $0.18 out. A new option for long-context tasks without the high cost. https:// olud.ai/latest.html # AI # LLM # OpenSource

  47. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash 不再仅限于文本 👀 在我们的内部基准测试中,其价格性能显著优于同类产品。我们已添加

    DeepSeek V4 Flash is no longer text-only 👀 In our internal benchmarks, it delivered significantly better price-performance than others in its class. We’ve added vision for screen-level understanding, capabilities needed for WebBrain Available on HF: https:// huggingface.co/webbra…

  48. r/LocalLLaMA TIER_1 English(EN) · /u/dangerous_inference ·

    DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vfrjwl/deepseekv4flash_on_sm89_4x48gb_4090s_with_dspark/"> <img alt="DeepSeek-V4-Flash on SM89 4x48gb 4090s with DSpark" src="https://external-preview.redd.it/ODFwYWZ6b2M2Z2hoMVEbeTvpmi808aYcwKkY2JtviqN_2flas…

  49. r/LocalLLaMA TIER_1 Svenska(SV) · /u/giveen ·

    DeepSeek-v4-Flash-Mini 54GB GGUF 运行速度约 20.5 t/s

    <!-- SC_OFF --><div class="md"><p>Took the REAP adaptation of DeepSeek-V4-Flash (<code>0xSero/DeepSeek-V4-Flash-0731-REAP</code>) along with <code>antirez/deepseek-v4-gguf</code> as inspiration, and decided to see how aggressive we could get with standard quant tricks to create a…

  50. r/LocalLLaMA TIER_1 Deutsch(DE) · /u/returnity ·

    DeepSeek v4 Flash 对比 Qwen3.6-27B、3.5-122B 和 Gemma 4 31B 基准测试

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vfhqkm/deepseek_v4_flash_vs_qwen3627b_35122b_and_gemma_4/"> <img alt="DeepSeek v4 Flash vs. Qwen3.6-27B, 3.5-122B, and Gemma 4 31B Benchmark" src="https://preview.redd.it/xucnql0lcehh1.png?width=140&amp;heigh…

  51. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    🤖 DeepSeek 的新 V4 Flash 比 Claude Opus 4.8 便宜 99%:每百万输出 token 仅需 0.28 美元,而 Claude Opus 4.8 为 25 美元。它还在 Arena 的前端编码方面击败了 Claude

    🤖 DeepSeek's new V4 Flash really is 99% cheaper than Claude Opus 4.8: $0.28 vs $25 per million output tokens. It also beat Claude on Arena's front-end coding leaderboard and scores 82.7 on Terminal-Bench, ahead of some Claude tiers. That part is true. The nuance: on overall intel…

  52. r/LocalLLaMA TIER_1 Nederlands(NL) · /u/Gohab2001 ·

    Deepseek V4 flash 0731 在 Agent Arena 排名第 21

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vff300/deepseek_v4_flash_0731_ranks_21_on_agent_arena/"> <img alt="Deepseek V4 flash 0731 ranks #21 on Agent Arena" src="https://preview.redd.it/522fsdwvtdhh1.png?width=140&amp;height=140&amp;auto=webp&amp;s=…

  53. r/LocalLLaMA TIER_1 English(EN) · /u/mrstoatey ·

    DeepSeek V4 Flash 0731 (Q4) 在一台 RTX PRO 6000 上实现了 1,328 tok/s 的预填充和约 29 tok/s 的解码速度

    <!-- SC_OFF --><div class="md"><p>I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB.</p> <p>These are timing-disabled internal Krasis results using INT4 experts. They aren't …

  54. r/LocalLLaMA TIER_1 English(EN) · /u/grumd ·

    Deepseek V4 Flash 2位量化是我能本地运行的第一个在SQL基准测试中达到100%的模型

    <!-- SC_OFF --><div class="md"><p>I really like to use this one SQL benchmark when testing new models. I had another post some time ago with my benchmarks, but I decided to post a new one because of how well Deepseek did. I like the benchmark because it's quick to run, is pretty …

  55. r/LocalLLaMA TIER_1 English(EN) · /u/BlackBeardAI ·

    [Deepseek-V4-Flash-0731] RTX5090 + DDR5 台式机设置,全1M上下文,支持VLLM CPU/内存卸载,约800 tps pp & 15+ tps 解码 [Agentic Coding]

    <!-- SC_OFF --><div class="md"><p>First of all, obviously I took some help from AI to type this post and this is the topic that enabled me to accomplish all that:</p> <p><a href="https://old.reddit.com/r/LocalLLaMA/comments/1veow4b/deepseek_v4flash_284b_moe_at_33_toks_single_68/"…

  56. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek-V4-Flash-0731 每月下载量达 23.61 万次——真实需求,而非炒作。MIT 许可的文本模型获得 2.1k 星标。开源权重领导者

    DeepSeek-V4-Flash-0731 is pulling 236.1k monthly downloads—real demand, not just hype. That’s 2.1k stars on an MIT-licensed text model. The open-weight leaderboard shifts daily, and this one’s climbing for a reason. https:// olud.ai/#leaderboard # OpenSource # AI # HuggingFace

  57. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash on a Single AMD MI300X https:// github.com/ryanzhou/deepseek-v 4-flash-mi300x Comments: https:// news.ycombinator.com/item?id=4 9166386 # Hack

    DeepSeek V4 Flash on a Single AMD MI300X https:// github.com/ryanzhou/deepseek-v 4-flash-mi300x Comments: https:// news.ycombinator.com/item?id=4 9166386 # HackerNews # DeepSeek # V4 # Flash # AMD # MI300X # AI # Technology # MachineLearning # GPU

  58. r/LocalLLaMA TIER_1 English(EN) · /u/AbbreviationsSad5582 ·

    DeepSeek V4-Flash (284B MoE) 在 2× RTX 3090 + 一台二手四路 Xeon DDR4 服务器上实现单卡 33 tok/s / 聚合 68 tok/s — 全配置

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1veow4b/deepseek_v4flash_284b_moe_at_33_toks_single_68/"> <img alt="DeepSeek V4-Flash (284B MoE) at 33 tok/s single / 68 tok/s aggregate on 2× RTX 3090 + a used quad-Xeon DDR4 server — full config" src="https:…

  59. r/LocalLLaMA TIER_1 English(EN) · /u/LimpComedian1317 ·

    我们测试了 Deepseek v4 flash、GLM 5.2 和 Kimi K3 在困难的代理任务上的表现,DeepSeek 表现出色

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1venecp/we_tested_deepseek_v4_flash_glm_52_and_kimi_k3_on/"> <img alt="We tested Deepseek v4 flash, GLM 5.2, and Kimi K3 on hard agentic tasks, and DeepSeek just crushed" src="https://preview.redd.it/uybjxyypj…

  60. r/LocalLLaMA TIER_1 English(EN) · /u/mintybadgerme ·

    我简直不敢相信我能在我的家用电脑上运行 frontier model DeepSeek-V4-Flash-0731。太疯狂了!

    <!-- SC_OFF --><div class="md"><p>So this is the stuff of absolute insanity. In less than 20 months we've gone from super expensive cloud models only, to being able to run a Q3 quant of DeepSeek on an Intel Windows PC with a very average 24GB of VRAM. No wonder the big boys are p…

  61. r/LocalLLaMA TIER_1 English(EN) · /u/reto-wyss ·

    DeepSeek V4 Flash 0731 - Happy Numbers (700pp/18tg) and Thoughts

    <!-- SC_OFF --><div class="md"><p>Originally, I was only getting around 140pp/s and about 21tg/s, but the config with <code>-b 8192 -ub 8192 --cpu-moe</code> is vastly superior, let's say <strong>700pp/s</strong> and <strong>18tg/s</strong> in the most relevant range.</p> <p><str…

  62. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @runsonai: 我一整天都在使用 DeepSeek V4 Flash,印象非常深刻。作为一个 Hermes 代理,它能满足你的一切愿望:代理式的

    RT @runsonai: Ich habe den ganzen Tag über DeepSeek V4 Flash genutzt und bin sehr beeindruckt. Als Hermes-Agent erledigt er alles, was man sich wünscht: agentices Verhalten, Tool-Calling, Programmierung und Reaktionsfähigkeit. Die Schwäche von DS4F liegt im Harness; das Modell se…

  63. dev.to — LLM tag TIER_1 English(EN) · AI OpenFree ·

    DeepSeek-V4-Flash-0731 内部揭秘:72,317 个 harness 优化 MoE 的张量

    <p>On 2026-07-31, DeepSeek quietly shipped <strong>V4-Flash-0731</strong> under an MIT license. Instead of reading the model card, VIDRAFT's <strong>Darwin</strong> model-inspection platform read the actual <code>config.json</code> and every weight shard — <strong>72,317 tensors …

  64. r/LocalLLaMA TIER_1 English(EN) · /u/coder543 ·

    DeepSeek-V4-Flash-0731:低者高于高者

    <!-- SC_OFF --><div class="md"><p>I decided to test a few questions against DeepSeek-V4-Flash-0731. Locally, I was running Unsloth's UD-Q2_K_XL quant. After I saw the surprising shape of the results, I tested against DeepSeek's official API to confirm that I didn't do anything wr…

  65. r/LocalLLaMA TIER_1 English(EN) · /u/fragment_me ·

    Deepseek v4 flash - 预填充/并行处理速度提升 100-150 倍。

    <!-- SC_OFF --><div class="md"><p>You have two choices here (in order of pref):</p> <ol> <li><p>Downgrade CUDA from 13.3 to 13.1 (skip 13.2 due to bugs) &lt;- prefer this (thanks to <a href="/u/fairydreaming">u/fairydreaming</a> for pointing this out)</p></li> <li><p>Use this vib…

  66. r/LocalLLaMA TIER_1 English(EN) · /u/Puzzleheaded_Base302 ·

    在 DGX Spark 上使用 vLLM-Moet 2 位量化运行 DeepSeek-V4-Flash-0731 (155 GB MoE) - AI 的叙事

    <!-- SC_OFF --><div class="md"><p># Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization</p> <p>I used Deepseek-v4-Flash-0731 cloud API settig up vllm-moet to run deepseek-v4-flash with MTP locally on single DGX Spark at 2-bit quant. Though…

  67. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash (0731) 在智能-成本比方面非常接近 GPT-5.6 Luna。# ai # news

    DeepSeek V4 Flash (0731) is quite close to GPT-5.6 Luna in terms of intelligence-cost ratio. # ai # news

  68. r/LocalLLaMA TIER_1 English(EN) · /u/Blahblahblakha ·

    DeepSeek-V4-Flash 284B 内存占用 5.3GB

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vdbix4/deepseekv4flash_284b_on_53gb_of_memory/"> <img alt="DeepSeek-V4-Flash 284B on 5.3GB of memory" src="https://external-preview.redd.it/dmlxOHY4ZDh5d2doMTKeaIgiLAILwUoBdScnKpKTXpOnyWsUOZKXl9qB_yy6.png?wid…

  69. dev.to — LLM tag TIER_1 English(EN) · Hunter G ·

    DeepSeek V4-Flash 更新以超低价格提供顶级分数

    <p>DeepSeek shipped the official V4-Flash on July 31 — <strong>open weights, MIT license, and a technical report, all on day one.</strong></p> <p>The striking part isn't that the score is high. It's that <strong>the score and the price arrived together</strong>: 50 on the Artific…

  70. r/LocalLLaMA TIER_1 English(EN) · /u/Hyungsun ·

    DeepSeek-V4-Flash-0731 UD-IQ3_XXS 在 1x 7900 XTX 24GB + 3x MI60 32GB + 128GB DDR4 上约 11t/s

    <!-- SC_OFF --><div class="md"><p>Hello, Also I want to join the hype of posting token specs.</p> <p>CPU: 2x Intel Xeon CPU E5-2650 v4 @ 2.20GHz</p> <p>RAM: 2x 4 Channel 2400MHz DDR4</p> <p>GPU: 1x AMD Radeon 7900 XTX 24GB</p> <p>3x AMD Instinct MI60 32GB</p> <p>Strange GPU combi…

  71. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @deepseek_ai: 🚀 DeepSeek-V4 Flash API 正式上线公测!🔷 我们已大幅提升其代理能力——

    RT @deepseek_ai: 🚀 Die offizielle DeepSeek-V4-Flash-API ist jetzt in der öffentlichen Beta verfügbar! 🔷 Wir haben ihre Agent-Fähigkeiten massiv verbessert – die Benchmark-Werte übertreffen nun die V4-Pro-Preview bei weitem. Entdecken Sie den enormen Leistungssprung unten! 👇 🔷 Die…

  72. r/LocalLLaMA TIER_1 English(EN) · /u/USBhost ·

    DeepSeek-V4-Flash-0731 在 A6000 + 256GB DDR4 上实现 17.20~ t/s

    <!-- SC_OFF --><div class="md"><p>Hello everyone I want to join the hype of posting specs.</p> <p>CPU: AMD EPYC 74F3 24-Core</p> <p>RAM: 8 Channel 3200 DDR4</p> <p>GPU: RTX A6000 48GB</p> <p>Prompt processing is in the high 70t/s (got down to mid 30t/s at 300k context). Inference…

  73. r/LocalLLaMA TIER_1 English(EN) · /u/HockeyDadNinja ·

    DeepSeek-V4-Flash-0731 的专家级 IQ3 再量化:KLD 优于 UD-IQ3_S,CPU 溢出设备解码速度提升 1.4 倍

    <!-- SC_OFF --><div class="md"><p>Hey all,</p> <p>tldr / who this helps: you run a mixed multi-GPU box where the experts spill to RAM, and you want to stay in the 3-bit tier instead of dropping to Q2 to make it fit. </p> <p><a href="https://huggingface.co/TacoTakumi/DeepSeek-V4-F…

  74. r/LocalLLaMA TIER_1 English(EN) · /u/Ok_Ninja7526 ·

    DeepSeek-V4-Flash-0731 UD-IQ3_S 在RTX 3090 +128GB DDR5上实现12.5 tok/s

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vcz61x/deepseekv4flash0731_udiq3_s_125_toks_on_rtx_3090/"> <img alt="DeepSeek-V4-Flash-0731 UD-IQ3_S 12.5 tok/s on RTX 3090 +128GB DDR5" src="https://external-preview.redd.it/NHV0ZnJ5NHB5dGdoMVngZSdHf_oCzXnIj…

  75. r/LocalLLaMA TIER_1 English(EN) · /u/esw123 ·

    DeepSeek V4 Flash 0731 IQ2_M 基准测试,双 3060 和 96GB 内存 ≈ 3.5 token/s。

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vcrd6d/deepseek_v4_flash_0731_iq2_m_benchmark_for_dual/"> <img alt="DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM ≈ 3.5 tok/s." src="https://preview.redd.it/3mzcq9labsgh1.png?width=640&amp…

  76. dev.to — LLM tag TIER_1 English(EN) · OctoLab ·

    我如何使用 DeepSeek V4 Flash:为不确定性保留最强模型

    <p>I believe the group-chat comparison. I do not believe it reflects the model’s true value.</p> <p>One person used the same prompt with GLM 5.2 + Claude Code and DeepSeek V4 Flash + Codex. The first run produced a playable game in a little over twenty turns. Another similar test…

  77. Mastodon — fosstodon.org TIER_1 Italiano(IT) · [email protected] ·

    🧠 DeepSeek通过官方API公测发布DeepSeek V4 Flash 0731,更新主要侧重于代理能力。👉

    🧠 # DeepSeek ha rilasciato DeepSeek V4 Flash 0731 tramite API ufficiale in public beta, con un aggiornamento concentrato soprattutto sulle capacità agentiche. 👉 I dettagli: https://www. linkedin.com/posts/alessiopoma ro_deepseek-claude-opus-share-7489318228167397376-DKwe/ ___ ✉️ …

  78. dev.to — LLM tag TIER_1 English(EN) · lora ·

    我如何为DeepSeek V4 Flash添加低Token视觉能力

    <p>Instead of replacing the main model with a more expensive multimodal model, I gave a text-only agent an on-demand visual sensor.</p> <p>Many coding agents can already read repositories, write code, execute commands, and run tests. But in real development workflows, they often …

  79. dev.to — LLM tag TIER_1 English(EN) · TokenPAPA ·

    DeepSeek V4 Flash 正式发布 (0731):Agent 性能超越 V4 Pro — 已上线 TokenPAPA

    <h1> DeepSeek V4 Flash Official Release (0731): Agent Benchmarks Jump Past V4 Pro </h1> <p>On July 31, 2026, DeepSeek officially released <strong>DeepSeek-V4-Flash</strong> to public beta. The API calling convention is unchanged — set <code>model</code> to <code>deepseek-v4-flash…

  80. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4-Flash API发布,通过训练后更新提升代理能力,Terminal-Bench得分+25.8,并发布开源权重。来源:Latent Space

    DeepSeek's V4-Flash API launch improves agent capabilities with a post-training update, +25.8 on Terminal-Bench, and open-weights release. Source: Latent Space https://www. latent.space/p/ainews-not-much -happened-today-038 # AI # Automation

  81. Mastodon — fosstodon.org TIER_1 Italiano(IT) · [email protected] ·

    DeepSeek 发布 #V4FlashAPI 公测版,新模型展现出 Agent 能力,并在基准测试中优于 V4 Pro Preview,标志着

    DeepSeek ha rilasciato la beta pubblica della # V4FlashAPI . Il nuovo modello mostra capacità per agenti e benchmark superiori alla V4 Pro Preview, segnando un passo avanti per l'ecosistema # AI e lo sviluppo di agenti intelligenti.

  82. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek 已将其 V4-Flash 模型升级,在智能体和编码能力方面取得重大进展。该 284B 参数 MoE 模型现已包含 DSpark 投机解码

    DeepSeek has upgraded its V4-Flash model with major gains in agentic and coding capabilities. The 284B-parameter MoE model now includes DSpark speculative decoding and supports the Responses API format. Pricing starts at 0.14 USD per million input tokens. https://www. marktechpos…

  83. r/LocalLLaMA TIER_1 English(EN) · /u/challis88ocarina ·

    DeepSeek-V4-Flash-Q4KExperts-F16HC-F16Compressor-F16Indexer-Q8Attn-Q8Shared-Q8Out-chat-v2-imatrix-0731.gguf

    <!-- SC_OFF --><div class="md"><p>Antirez stealthily uploaded the new weights in the old folder... and there we were tapping our fingers.</p> <p><a href="https://huggingface.co/antirez/deepseek-v4-gguf/tree/main">https://huggingface.co/antirez/deepseek-v4-gguf/tree/main</a></p> <…

  84. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash现已默认使用AI Gateway上的更新权重,在Terminal-Bench上提升25.8分至82.7,以增强代理编码任务。无模式

    DeepSeek V4 Flash now defaults to updated weights on AI Gateway, delivering a 25.8-point Terminal-Bench boost to 82.7 for stronger agentic coding tasks. No model ID or code changes needed. Zero Data Retention providers arrive next week. Source: Vercel Blog https:// vercel.com/cha…

  85. r/LocalLLaMA TIER_1 English(EN) · /u/davidthesong ·

    Deepseek V4 Flash 现已成为 Kimi K3 的约第二大开源模型,且成本降低 50 倍以上

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vc4041/deepseek_v4_flash_is_now_2_open_weight_model_to/"> <img alt="Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and &gt;50x cheaper" src="https://preview.redd.it/h7zv5tb3tmgh1.png?width=140&amp;…

  86. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    DeepSeek的V4 Flash模型在AI Index上得分50,成本降低60%,几乎追平GPT-5.6 Luna。效率上的明显胜利。来源:The Decoder AI

    DeepSeek's V4 Flash model now scores 50 on the AI Index, nearly matching GPT-5.6 Luna at 60% lower cost. A clear win for efficiency. Source: The Decoder AI https:// the-decoder.com/new-deepseek-f lash-model-matches-openais-gpt-5-6-luna-at-roughly-60-percent-lower-cost/ # AI # Aut…

  87. r/LocalLLaMA TIER_1 English(EN) · /u/curiousily_ ·

    DeepSeek v4 Flash 的初步测试显示其在 UI/UX 设计能力方面有显著提升(尽管它很耗费 token)

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbxoi5/initial_testing_of_deepseek_v4_flash_shows/"> <img alt="Initial testing of DeepSeek v4 Flash shows significant improvements in UI/UX design capabilities (despite being token hungry)" src="https://exter…

  88. r/LocalLLaMA TIER_1 English(EN) · /u/RunawayPeeko ·

    DeepSeek V4 Flash unsloth quants 现已发布!

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbxk53/deepseek_v4_flash_unsloth_quants_are_out/"> <img alt="DeepSeek V4 Flash unsloth quants are out!" src="https://external-preview.redd.it/wMZUsudfTHG534WgBRWwhtsK20jasRrxNGwenbWxNxM.png?width=640&amp;crop…

  89. dev.to — LLM tag TIER_1 Nederlands(NL) · David ·

    DeepSeek V4 Flash 对比 V4 Pro:开发者决策指南

    <p>With today's V4 Flash 0731 release, DeepSeek's V4 family now has two very different members, and picking between them is a real decision. Short version: Flash for agents and code, Pro for the hardest reasoning, Flash by forfeit if you want to run it yourself.</p> <h2> The spec…

  90. r/LocalLLaMA TIER_1 Nederlands(NL) · /u/Nyghtbynger ·

    DeepSeek-V4-Flash 20260731 观点评测

    <!-- SC_OFF --><div class="md"><p>First of all, I want to apologize if it's off-topic or in the wrong format.</p> <p>Having tried Deepseek Flash with reasoning high on a conceptually difficult task, involving Machine Learning classifiers and graphs. I am extremely impressed. It's…

  91. r/LocalLLaMA TIER_1 English(EN) · /u/SnooBunnies8392 ·

    DeepSeek-V4-Flash-0731 在基准测试中已远超 DeepSeek-V4-Pro-Preview

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbkvau/deepseekv4flash0731_now_far_surpassing_the/"> <img alt="DeepSeek-V4-Flash-0731 now far surpassing the DeepSeek-V4-Pro-Preview in benchmarks" src="https://preview.redd.it/bq9d2c2vyigh1.jpeg?width=640&am…

  92. r/LocalLLaMA TIER_1 English(EN) · /u/MagicZhang ·

    New DeepSeek V4-Flash 在 ArtificalAnalysis 指数上取得 50 分,比 GLM-5.2 和 GPT-5.6 Luna 低 1 分

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbk5ob/new_deepseek_v4flash_achieves_50_on/"> <img alt="New DeepSeek V4-Flash achieves 50 on ArtificalAnalysis Index, 1 point below GLM-5.2 and GPT-5.6 Luna" src="https://preview.redd.it/mtmrp4lnrigh1.jpeg?wi…

  93. r/LocalLLaMA TIER_1 English(EN) · /u/Nunki08 ·

    DeepSeek-V4-Flash 已更新,“DeepSeek-V4-Pro 官方发布将很快跟进”

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vbidkp/deepseekv4flash_has_been_updated_the_official/"> <img alt="DeepSeek-V4-Flash has been updated, &quot;The official release of DeepSeek-V4-Pro will follow soon&quot;" src="https://preview.redd.it/mbz7sdw…

  94. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @jun_song: 翻译:SuperDeepseek-V4-Flash 在 2xDGX Spark 上运行,峰值速度为每秒 122 个 token。该工作基于...

    RT @jun_song: TRANSLASATION: SuperDeepseek-V4-Flash läuft auf 2xDGX Spark mit einem Spitzenwert von 122 Token pro Sekunde. Die Arbeit basiert auf dem Rezept von @MiaAIlab zur Ausführung mit DFlash MTP. 4-Bit-Quantisierung mit weniger als ~1 % Qualitätsverlust. Außerdem wurde es a…

  95. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @pupposandro: Lucebox Engine 现已在 128 GB AMD Strix Halo 系统上,通过单个 98.29 GB GGUF 模型运行 DeepSeek V4 Flash 0731。

    RT @pupposandro: Lucebox Engine betreibt nun DeepSeek V4 Flash 0731 von einem einzelnen 98,29 GB großen GGUF-Modell auf einem 128 GB AMD Strix Halo System. Das Modell erzielt 82/92 Punkte auf ds4-eval-92 und erreicht 32,7 Token pro Sekunde mit DSpark. Über unsere festgelegten Eva…

  96. Mastodon — mastodon.social TIER_1 English(EN) · sipirtu ·

    DeepSeek的V4 Flash模型每百万输入令牌定价0.14美元,每百万输出令牌定价0.28美元,但第三方价格更低。30倍的涨价可能仍

    DeepSeek’s V4 Flash model lists at $0.14 per million input tokens and $0.28 per million output tokens, yet third parties undercut this. A 30x price hike may still leave it the cheapest for users with stable prompts due to its cache architecture. Source: Pandaily https:// pandaily…

  97. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @MiaAI_lab: 给我一个更好的设置,可以在本地以 80 多个 token/s 的速度运行 DeepSeek v4 Flash。你只需要 2 台 DGX Sparks。我等着呢

    RT @MiaAI_lab: Gebt mir ein besseres Setup, das DeepSeek v4 Flash lokal mit 80+ Token/s ausführt. Alles was ihr braucht sind 2 DGX Sparks. Ich warte darauf, dass Mike (@mike64t) Nvidia lacht, während die Leute auf einen 4.000 USD teuren glorifizierten Raspberry Pi hereinfallen, d…

  98. Mastodon — mastodon.social TIER_1 English(EN) · aisyndicate ·

    284B 2位量化至128GB:DeepSeek V4 Flash 20 tok/s DeepSeek V4 Flash作为2位GGUF运行在128GB统一内存上:20 tok/s,工具评估81/100。一个有效

    284B in 2 Bit auf 128 GB: DeepSeek V4 Flash mit 20 tok/s DeepSeek V4 Flash als 2-Bit-GGUF auf 128 GB Unified Memory: 20 tok/s, 81/100 im Tool-Eval. Ein validiertes Laborobjekt für Offline-Betrieb, nicht für Effizienz. https:// aisyndicate.ch/deepseek-v4-fla sh-dgx-spark-test # AI…

  99. Mastodon — mastodon.social TIER_1 English(EN) · opensourceaitech ·

    DeepSeek V4 Flash 0731 价格30天内下跌36%:每100万输出令牌价格从0.28美元降至0.18美元。开放权重使其成为一个有趣的替代方案。https:// olud.ai/mode

    DeepSeek V4 Flash 0731’s price fell 36% in 30 days: $0.28→$0.18 per 1M output tokens. Open weights make it an interesting alternative now. https:// olud.ai/model/deepseek-v4-flas h-0731.html # AI # LLM # Pricing

  100. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @iam_elias1: Deepseek v4 Flash 因使用人数过多而过载。当你把一个模型做得太好时就会发生这种情况。OpenCode (@opencode) 拥有

    RT @iam_elias1: Deepseek v4 Flash ist überlastet, weil zu viele Menschen es nutzen. Das passiert, wenn man ein Modell zu gut macht. OpenCode (@opencode) hat derzeit Kapazitätsprobleme bei Deepseek Flash aufgrund des beispiellosen Volumens; Sie können Fehler sehen - wir arbeiten a…

  101. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    我只是在试用 DeepSeek v4 Flash 0731,天哪,这个模型这么小,需要的资源这么少,但它却令人难以置信。它只需要

    I am just trying out DeepSeek v4 Flash 0731 and holly shit this thing is incredible for how small this model is and how little resources it needs. It only needs 128 GB... which is at least 4 to 15 times less then the comparable new Models like GLM-5.2 or Kimi-K3... I guess China …

  102. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    DeepSeek V4 Flash on a Single AMD MI300X

    DeepSeek V4 Flash on a Single AMD MI300X Article URL: https:// github.com/ryanzhou/deepseek-v 4-flash-mi300x Comments URL: https:// news.ycombinator.com/item?id=4 9166386 Points: 12 # Comments: 0 https:// github.com/ryanzhou/deepseek-v 4-flash-mi300x # Tech # Technology # TechNew…

  103. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @runsonai: 我一整天都在使用 DeepSeek V4 Flash,印象非常深刻。作为一个 Hermes 代理,它能满足你的一切愿望:代理式

    RT @runsonai: Ich habe den ganzen Tag über DeepSeek V4 Flash genutzt und bin sehr beeindruckt. Als Hermes-Agent erledigt er alles, was man sich wünscht: agentices Verhalten, Tool-Calling, Programmierung und Reaktionsfähigkeit. Die Schwäche von DS4F liegt im Harness; das Modell se…

  104. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @Bassmaster187: 我现在已经试用了 DeepSeek V4 Flash 0731,因为我已经达到了 Qwen 3.6 27B 的极限。根据 Benc 的说法,新的 DeepSeek 0731 应该...

    RT @Bassmaster187: Ich habe jetzt DeepSeek V4 Flash 0731 ausprobiert, da ich mit Qwen 3.6 27B an die Grenzen gestoßen bin. Das neue DeepSeek 0731 soll laut Benchmarks extrem gut sein und extrem günstig. Ich hatte mal ein Angular-Plugin für Grafana geschrieben, und Angular ist jet…

  105. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @MiaAI_lab: DeepSeek v4 Flash 感觉就像 Qwen3.6 27B 的时刻重演,只是规模要大得多。以前从未有过如此低的成本实现如此多的可能性

    RT @MiaAI_lab: DeepSeek v4 Flash fühlt sich wie der Qwen3.6 27B-Moment noch einmal an, nur auf viel größerer Ebene. Noch nie konnte man so viel für so wenig Geld tun. Dies könnte die Wirtschaftlichkeit von KI verändern. OpenCode (@opencode) DeepSeek Flash hat am 1. August 8 Billi…

  106. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @__tinygrad__: 245 单用户令牌/秒,在 DeepSeek-V4-Flash-0731 上,仅使用 2 个 RTX 6000 Blackwell GPU!为纪念 DeepSe

    RT @__tinygrad__: 245 Single-User-Tokens pro Sekunde auf DeepSeek-V4-Flash-0731, und dabei werden nur 2 der RTX 6000 Blackwell-GPUs genutzt! Zu Ehren von DeepSeek starten wir die 2-GPU-Edition unserer tinybox. Alle Hardwarekomponenten für die Installation von 4 GPUs sind enthalte…

  107. r/Anthropic TIER_1 English(EN) · /u/hibzy7 ·

    DeepSeek V4 Flash API 输入便宜18倍,输出便宜28倍,性能媲美Opus 4.8。是时候让Claude至少降低Sonnet定价了

    <table> <tr><td> <a href="https://www.reddit.com/r/Anthropic/comments/1vcbac2/deepseek_v4_flash_api_is_18x_cheaper_on_input_28x/"> <img alt="DeepSeek V4 Flash API is 18x cheaper on input, 28x cheaper on output, and matches Opus 4.8. Time for Claude to atleast reduce sonnet pricin…

  108. r/Anthropic TIER_1 English(EN) · /u/Icy-Investment407 ·

    新发布的 Deepseek V4 Flash Official 评分接近 Claude Opus 4.8。价格:0.18美元/100万输出 token

    <table> <tr><td> <a href="https://www.reddit.com/r/Anthropic/comments/1vc9zj1/newly_released_deepseek_v4_flash_official_scores/"> <img alt="Newly released Deepseek V4 Flash Official scores close to Claude Opus 4.8. Price: 0.18$ / 1mil OUTPUT tokens" src="https://preview.redd.it/u…

  109. Mastodon — mastodon.social TIER_1 Italiano(IT) · AI_BEAR_NEWS ·

    🤖 DeepSeek 发布 V4 Flash 公测 API 中国顶级模型现已具备增强的代理能力,基准测试超越 V4-Pro-Pr

    🤖 DeepSeek lancia API beta pubblica per V4 Flash Il modello di punta cinese ora disponibile con capacità agentiche potenziate e benchmark che superano V4-Pro-Preview. Aperta la strada per sviluppatori e aziende che cercano alternative economicamente accessibili. Fonte: Bloomberg …

  110. r/OpenAI TIER_2 English(EN) · /u/sirMoped ·

    DeepSeek v4 flash 的新帖子训练已发布

    <table> <tr><td> <a href="https://www.reddit.com/r/OpenAI/comments/1vc527y/new_post_train_of_deepseek_v4_flash_is_out/"> <img alt="New post train of DeepSeek v4 flash is out" src="https://preview.redd.it/vsao0zgg3ngh1.png?width=640&amp;crop=smart&amp;auto=webp&amp;s=017d0be1788b6…

  111. r/singularity TIER_2 English(EN) · /u/DeArgonaut ·

    DeepSeek V4 Flash 0731 ARC-AGI-1 and 2

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1vj65p3/deepseek_v4_flash_0731_arcagi1_and_2/"> <img alt="DeepSeek V4 Flash 0731 ARC-AGI-1 and 2" src="https://preview.redd.it/o3yky8nrm7ih1.png?width=140&amp;height=65&amp;auto=webp&amp;s=c276ccfedaa1f9730ea…

  112. r/singularity TIER_2 English(EN) · /u/Hot_Example_4456 ·

    Deepseek v4 flash 0731 的权重已发布!!!

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1vbphpz/weights_of_deepseek_v4_flash_0731_have_been/"> <img alt="Weights of Deepseek v4 flash 0731 have been released!!!" src="https://external-preview.redd.it/4wuKMSgR8Kkk0pgA3-HhBYFoQX79t2M-87LJSpJS8lU.png?…

  113. r/singularity TIER_2 English(EN) · /u/Boring_Aioli7916 ·

    DeepSeek-V4-Flash 官方 API 公测上线!Flash 模型迎来重大升级。

    <table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1vbm5m8/deepseekv4flash_official_api_is_now_live_in/"> <img alt="DeepSeek-V4-Flash Official API is now LIVE in public beta! Massive upgrades for flash model." src="https://preview.redd.it/z35fca23cjgh1.jpeg?w…