An open-weight model named GLM 5.3 has demonstrated performance approaching that of Anthropic's restricted Mythos Preview on the ExploitBench benchmark. While GLM 5.3's capabilities are nearing frontier models, estimates suggest it would cost approximately $1,200 to remove most of its safety restrictions. The US government has indicated that GLM 5.3 still trails behind the most advanced frontier models. AI
IMPACT This development highlights the rapid progress of open-weight models, potentially narrowing the gap with proprietary frontier models and raising new safety considerations.
RANK_REASON Research milestone for an open-weight model achieving near-frontier performance on a benchmark. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →