Zhipu AI has released GLM 5.3 Flash, a new multimodal mixture-of-experts model with 320 billion parameters. This model is designed for efficient serving, offering significantly reduced compute and KV cache size compared to its predecessor, GLM 5.3. GLM 5.3 Flash achieves performance comparable to Claude Opus 4.8 on certain benchmarks, particularly in agentic tasks, while being substantially more cost-effective. The model is available under an MIT license with weights published at launch, a departure from Zhipu AI's previous API-first releases. AI
IMPACT This release offers a cost-effective alternative to top-tier models, potentially lowering the barrier for multimodal AI applications.
RANK_REASON Frontier-lab model release with system card and benchmark data. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →