Meta's latest foundational model, Muse Spark 1.2, has achieved high scores in third-party performance analyses, demonstrating rapid improvement since the Muse series' debut four months ago. The model notably surpassed Grok 4.5 on the Artificial Analysis Intelligence Index and showed enhanced capabilities in agentic tasks, coding, and customer support benchmarks. Despite its strong performance and cost-effectiveness compared to competitors, Muse Spark 1.2 slightly trails behind top-tier models like Claude Opus 5 and GPT-5.6 Sol in certain advanced benchmarks. AI
IMPACT Sets a new bar for cost-effective AI performance, potentially pressuring competitors to optimize their models for efficiency.
RANK_REASON Meta's release of a new foundational model version with benchmark performance data. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
Read on Mastodon — fosstodon.org →
- AA-Omniscience Index
- Artificial Analysis Intelligence Index
- Claude Fable 5
- Claude Opus 5
- GDPval-AA v2
- GPT-5.5
- GPT-5.6 Sol
- Grok 4.5
- Humanity's Last Exam
- Kimi K3
- Meta
- Muse Spark 1.0
- Muse Spark 1.1
- Muse Spark 1.2
- SciCode
- SpaceXAI
- Terminal-Bench v2.1
- Vals AI
- Vals Index
- τ³-Banking
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →