Alibaba's Qwen has launched its Qwen3.8-Flash model, available on Qwen Cloud with competitive pricing for API usage. The model is also accessible through OpenRouter, enabling various applications like coding assistants and agentic workflows. Furthermore, Qwen3.8-Flash-Next has received day-zero support from key partners including SGLang and vLLM, with compatibility confirmed on NVIDIA and AMD hardware. This new architecture is designed for ultimate cost-efficiency and is also available via Ollama. AI
IMPACT Accelerates adoption of multimodal models for coding and agentic workflows with competitive pricing.
RANK_REASON Frontier-lab model release with system card.
Read on Mastodon — sigmoid.social →
- Claude Code
- DeepSeek
- modelscope
- Qwen
- Qwen3.8-27B
- Qwen3.8-Flash-Next
- UnslothAI
- Alibaba Group
- Alibaba Qwen
- AMD
- NVIDIA
- SGLang
- vLLM
- YaRN
- Ollama
- OpenRouter
- Qwen3.8-Flash
- Qwen Cloud
AI-generated summary · Google Gemini · from 21 sources. How we write summaries →