Alibaba's Qwen team has released Qwen3.8-2.4T-A95B, the first open-weight model in their Qwen-Max class. This large model, featuring a hybrid attention architecture and 2.4 trillion parameters with 95 billion activated per token, is designed for demanding agentic and reasoning tasks. The release details how to deploy this model on Amazon SageMaker HyperPod using vLLM, enabling efficient inference with features like tool calling and speculative decoding. AI
IMPACT Sets a new benchmark for open-weight models, potentially accelerating enterprise adoption for complex agentic workloads.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
Read on Mastodon — mastodon.social →
- Alibaba Group
- Amazon SageMaker HyperPod
- Kimi k3
- NVIDIA B300 Blackwell Ultra GPUs
- OpenAI
- Qwen
- Qwen3.8-2.4T-A95B
- vLLM
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →