PulseAugur
EN
LIVE 20:34:34

NVIDIA vLLM supports DeepSeekv4.1 Flash on release; AMD vLLM lags

NVIDIA's vLLM software is functioning seamlessly with the new DeepSeekv4.1 Flash model across all six of its hardware SKUs, including H100, H200, B200, B300, GB200, and GB300. In contrast, AMD's vLLM implementation is experiencing issues and has not yet been publicly released for the DeepSeekv4.1 Flash model, despite AMD's emphasis on speed. This disparity highlights a difference in immediate software support for the latest AI models between the two hardware manufacturers. AI

IMPACT Highlights the critical role of software ecosystem support for new AI model releases, potentially influencing hardware purchasing decisions.

RANK_REASON Comparison of immediate software support for a new model release across competing hardware vendors.

Read on X — SemiAnalysis →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

NVIDIA vLLM supports DeepSeekv4.1 Flash on release; AMD vLLM lags

How we ranked this

Signal score
21 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
Comparison of immediate software support for a new model release across competing hardware vendors.
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

Full methodology in our editorial standards.

COVERAGE [2]

  1. X — SemiAnalysis TIER_1 English(EN) · SemiAnalysis_ ·

    AMD's vLLM documentation points to using vllm/vllm-openai-rocm:deepseekv41-flash-0909, but from hour 0 of the model release to now, hour 23, AMD has still not p

    AMD's vLLM documentation points to using vllm/vllm-openai-rocm:deepseekv41-flash-0909, but from hour 0 of the model release to now, hour 23, AMD has still not publicly released the image. AMD claims, "SPEED IS THE MOAT," yet it has still not released it by the 23rd hour. We wish …

  2. X — SemiAnalysis TIER_1 English(EN) · SemiAnalysis_ ·

    On the Day 0 release of DeepSeekv4.1 Flash, NVIDIA vLLM works out of the box with zero issues across all 6 SKUs: H100, H200, B200, B300, GB200, GB300! Amazing w

    On the Day 0 release of DeepSeekv4.1 Flash, NVIDIA vLLM works out of the box with zero issues across all 6 SKUs: H100, H200, B200, B300, GB200, GB300! Amazing work by the NVIDIA &amp; Inferact teams! In comparison, AMD vLLM still does not work on DeepSeekv4.1 Flash, as we will ht…