Alibaba's Qwen team has released Qwen3.8-2.4T-A95B, marking the first open-weights release of a Qwen-Max-class model. This large model, featuring a hybrid attention architecture and 2.4 trillion parameters with 95 billion activated per token, is designed for demanding agentic and reasoning tasks. The release includes open weights and demonstrates deployment on Amazon SageMaker HyperPod using vLLM with NVIDIA B300 Blackwell Ultra GPUs, offering an OpenAI-compatible endpoint. AI
IMPACT Sets a new benchmark for open-weight frontier models, potentially accelerating self-hosted AI deployments for complex agentic tasks.
RANK_REASON Frontier-lab model release with open weights and system card. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
Read on AWS Machine Learning Blog →
- Alibaba Group
- Amazon SageMaker HyperPod
- Kimi k3
- NVIDIA B300 Blackwell Ultra GPUs
- OpenAI
- Qwen
- Qwen3.8-2.4T-A95B
- vLLM
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →