PulseAugur
EN
LIVE 19:30:11

Alibaba's Qwen3.8-Flash model launches with broad partner support

Alibaba's Qwen has launched its Qwen3.8-Flash model, available on Qwen Cloud with competitive pricing for API usage. The model is also accessible through OpenRouter, enabling various applications like coding assistants and agentic workflows. Furthermore, Qwen3.8-Flash-Next has received day-zero support from key partners including SGLang and vLLM, with compatibility confirmed on NVIDIA and AMD hardware. This new architecture is designed for ultimate cost-efficiency and is also available via Ollama. AI

IMPACT Accelerates adoption of multimodal models for coding and agentic workflows with competitive pricing.

RANK_REASON Frontier-lab model release with system card.

Read on Mastodon — sigmoid.social →

AI-generated summary · Google Gemini · from 21 sources. How we write summaries →

Alibaba's Qwen3.8-Flash model launches with broad partner support

How we ranked this

Signal score
1 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Frontier Release
Frontier-lab model release with system card.
Source corroboration
21 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
4 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+4 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [21]

  1. X — Qwen (Alibaba) TIER_1 English(EN) · Alibaba_Qwen ·

    Qwen3.8-Flash on @qwen_cloud: $0.15/1M input tokens, $0.47/1M output tokens, and just $0.016/1M on cache hits. ☁️ Come give it a try! 👇

    Qwen3.8-Flash on @qwen_cloud: $0.15/1M input tokens, $0.47/1M output tokens, and just $0.016/1M on cache hits. ☁️ Come give it a try! 👇

  2. X — Qwen (Alibaba) TIER_1 English(EN) · Alibaba_Qwen ·

    Qwen3.8-Flash is now live on @OpenRouter! ⚡ Coding assistants, agentic workflows, long-video understanding, all one API call away.

    Qwen3.8-Flash is now live on @OpenRouter! ⚡ Coding assistants, agentic workflows, long-video understanding, all one API call away.

  3. X — Qwen (Alibaba) TIER_1 English(EN) · Alibaba_Qwen ·

    Big thanks to @sgl_project for the day-0 support! 🙌 Qwen3.8-Flash-Next is ready to deploy with SGLang today.

    Big thanks to @sgl_project for the day-0 support! 🙌 Qwen3.8-Flash-Next is ready to deploy with SGLang today.

  4. X — Qwen (Alibaba) TIER_1 English(EN) · Alibaba_Qwen ·

    Amazing! 🥳 Thanks @vllm_project for getting Qwen3.8-Flash-Next running on NVIDIA and AMD from day 0.

    Amazing! 🥳 Thanks @vllm_project for getting Qwen3.8-Flash-Next running on NVIDIA and AMD from day 0.

  5. X — Qwen (Alibaba) TIER_1 English(EN) · Alibaba_Qwen ·

    Big thanks to @sgl_project for the day-0 support! 🙌 Qwen3.8-Flash-Next is ready to deploy with SGLang today.

    Big thanks to @sgl_project for the day-0 support! 🙌 Qwen3.8-Flash-Next is ready to deploy with SGLang today.

  6. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next Intelligence, Performance and Price Analysis https:// artificialanalysis.ai/models/q wen3-8-flash-next Comments: https:// news.ycombinator.co

    Qwen3.8-Flash-Next Intelligence, Performance and Price Analysis https:// artificialanalysis.ai/models/q wen3-8-flash-next Comments: https:// news.ycombinator.com/item?id=4 9461862 # HackerNews # Qwen3 .8 # Flash # Next # Intelligence # Performance # Price # Analysis # AI # Techno…

  7. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next Intelligence, Performance and Price Analysis https:// artificialanalysis.ai/models/q wen3-8-flash-next # ai

    Qwen3.8-Flash-Next Intelligence, Performance and Price Analysis https:// artificialanalysis.ai/models/q wen3-8-flash-next # ai

  8. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    https://www. europesays.com/3217680/ Experiment with Qwen3.8-Flash-Next on NVIDIA GB300 NVL72 for Agentic Coding # AgenticAI # AgenticArtificialIntelligence # A

    https://www. europesays.com/3217680/ Experiment with Qwen3.8-Flash-Next on NVIDIA GB300 NVL72 for Agentic Coding # AgenticAI # AgenticArtificialIntelligence # AI # ArtificialIntelligence

  9. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    🚀 New Ollama Model Release! 🚀 Model: qwen3.8-flash-next 🔗 https:// ollama.com/library/qwen3.8-fla sh-next # Ollama # AI # LLM # Gemini # GeminiAPI

    🚀 New Ollama Model Release! 🚀 Model: qwen3.8-flash-next 🔗 https:// ollama.com/library/qwen3.8-fla sh-next # Ollama # AI # LLM # Gemini # GeminiAPI

  10. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency https:// qwen.ai/blog?id=qwen3.8-flash- next # ai

    Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency https:// qwen.ai/blog?id=qwen3.8-flash- next # ai

  11. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency https://qwen.ai/blog?id=qwen3.8-flash-next # HackerNews # Tech # AI

    Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency https://qwen.ai/blog?id=qwen3.8-flash-next # HackerNews # Tech # AI

  12. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next https:// qwen.ai/blog?id=qwen3.8-flash- next # ai

    Qwen3.8-Flash-Next https:// qwen.ai/blog?id=qwen3.8-flash- next # ai

  13. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @quxiaoyin: I ran the same task on Claude Code and DeepSeek's new agent platform. One cost $150, the other only $2

    RT @quxiaoyin: Ich habe dieselbe Aufgabe auf Claude Code und der neuen Agent-Plattform von DeepSeek ausgeführt. Die eine kostete 150 Dollar, die andere nur 2 Dollar. Heute starten wir https:// t.co/k6mrHjfQ3C (@agentskydev), das „OpenRouter für Agenten“ – eine API für Claude Code…

  14. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @danielhanchen: Qwen announces Qwen3.8-Flash-Next, a new open-weight multimodal MoE model. 💜 more on Arint.info # AI # MoE # Multimodal # OpenSource #

    RT @danielhanchen: Qwen kündigt Qwen3.8-Flash-Next an, ein neues Open-Weight-Multimodal-MoE-Modell. 💜 mehr auf Arint.info # AI # MoE # Multimodal # OpenSource # Qwen # UnslothAI # arint_info https://x.com/danielhanchen/status/2092222459550585019

  15. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    RT @UnslothAI: Now you can fine-tune Qwen3.8-27B for free with our notebook! 🔥 more on Arint.info # AI # DeepLearning # FineTuning # MachineLearning

    RT @UnslothAI: Jetzt kannst du Qwen3.8-27B kostenlos mit unserem Notebook feinabstimmen! 🔥 mehr auf Arint.info # AI # DeepLearning # FineTuning # MachineLearning # OpenSource # Qwen3 # arint_info https://x.com/UnslothAI/status/2092255772214509698

  16. Mastodon — fosstodon.org TIER_1 Deutsch(DE) · [email protected] ·

    @OliDietzel: 8 hours left until the release of # Qwen3 .8-Flash-Next x.com/ModelScope2022… more on Arint.info # AI # FlashNext # ModelScope # Qwen3 # Release

    @OliDietzel: Noch 8 Stunden bis zum Release von # Qwen3 .8-Flash-Next x.com/ModelScope2022… mehr auf Arint.info # AI # FlashNext # ModelScope # Qwen3 # Release # arint_info https://x.com/OliDietzel/status/2092503972552577279

  17. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    RT @0xSero: Deepseek-v4-Flash-exl3-reap mehr auf Arint.info # AI # Benchmark # Deepseek # LLM # MachineLearning # Tech # arint_info https://x.com/0xSero/status/

    RT @0xSero: Deepseek-v4-Flash-exl3-reap mehr auf Arint.info # AI # Benchmark # Deepseek # LLM # MachineLearning # Tech # arint_info https://x.com/0xSero/status/2093754108708598240

  18. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @OrcaRouter: Qwen3.8 Flash Next. Uncensored. Now in NVFP4. 🐳 more on Arint.info # AI # HuggingFace # MachineLearning # NVFP4 # OpenSource # Qwen3 # arint_

    RT @OrcaRouter: Qwen3.8 Flash Next. Unzensiert. Jetzt in NVFP4. 🐳 mehr auf Arint.info # AI # HuggingFace # MachineLearning # NVFP4 # OpenSource # Qwen3 # arint_info https://x.com/OrcaRouter/status/2093233719473746065

  19. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @0xBakeer: Qwen3.8-Flash-Next on a single DGX Spark: up to 97 tokens per second. more on Arint.info # AI # DGXSpark # MachineLearning # NLP # Open

    RT @0xBakeer: Qwen3.8-Flash-Next auf einer einzelnen DGX Spark: bis zu 97 Tokens pro Sekunde. mehr auf Arint.info # AI # DGXSpark # MachineLearning # NLP # OpenSource # Qwen3 # arint_info https://x.com/0xBakeer/status/2092709567503229015

  20. Mastodon — mastodon.social TIER_1 Deutsch(DE) · [email protected] ·

    RT @vllm_project: Qwen3.8-Flash-Next from @AlibabaQwen has Day-0 support in vLLM, verified on NVIDIA and AMD GPUs. 🎉 more at Arint.info # AI # AMD # Machin

    RT @vllm_project: Qwen3.8-Flash-Next von @AlibabaQwen hat Day-0-Support in vLLM, verifiziert auf NVIDIA- und AMD-GPUs. 🎉 mehr auf Arint.info # AI # AMD # MachineLearning # NVIDIA # Qwen # vLLM # arint_info https://x.com/vllm_project/status/2092600887873286157

  21. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    Qwen3.8-Flash-Next https://simonwillison.net/2026/Aug/26/qwen38-flash-next/ # AI # LLM # Tech

    Qwen3.8-Flash-Next https://simonwillison.net/2026/Aug/26/qwen38-flash-next/ # AI # LLM # Tech