Artificial Analysis Intelligence Index
PulseAugur coverage of Artificial Analysis Intelligence Index — every cluster mentioning Artificial Analysis Intelligence Index across labs, papers, and developer communities, ranked by signal.
- instance of Kimi k3 90%
- instance of GLM-5.2 90%
- instance of Opus 4.8 90%
- instance of Claude Opus-5 90%
- instance of SpaceXAI 90%
- instance of Z Air 90%
- competes with GPT 5.6 "Sol" 80%
- competes with Claude Opus 4-8 80%
- instance of Claude Fable-5 70%
- instance of GPT 5.6 "Sol" 70%
- competes with Claude Opus-5 70%
- instance of Claude Opus 4-8 70%
7 day(s) with sentiment data
-
K2 Horizon 7B model shows strong performance for its size
The K2 Horizon 7B model has demonstrated strong performance, positioning itself between the Qwen 3.6 27B and Qwen 3.6 35BA3b models on the Artificial Analysis Intelligence Index. Initial testing suggests the model is so…
-
Alibaba's Qwen3.8-27B model achieves competitive performance on Cerebras hardware
Alibaba's Qwen3.8-27B model is now operational on Cerebras hardware, offering rapid inference capabilities. This open-weight model achieves a score of 34 on the Artificial Analysis Intelligence Index. This performance p…
-
DeepSeek launches V4.1-Flash with novel architecture for efficient AI agents
DeepSeek has released V4.1-Flash, a new open-weight flagship model featuring a novel causal Encoder-Decoder architecture. This model emphasizes extreme inference efficiency and cost-effectiveness, with a 763 billion par…
-
2.5B parameter MiniCPM5 model challenges larger LLMs on benchmarks
The MiniCPM5–2B model, developed by OpenBMB, represents a significant advancement in smaller, highly capable language models. Despite its 2.5 billion parameters and a 1.04 GB file size, it outperforms larger models like…
-
OpenAI's GPT-6 Astra shows major gains, especially in graphics, per analysis
Sebastian Raschka's analysis suggests OpenAI's new GPT-6 Astra model demonstrates significant improvements over its predecessor, GPT-5.6 "Sol," particularly in graphical tasks and achieving a near-perfect score on the A…
-
Google releases rapid-fire Gemini Flash models amid flagship delay
Google has released four Gemini Flash models in rapid succession, with the latest, Gemini 3.8 Flash, excelling in coding and matching larger models' performance at a lower cost. This rapid release pace is highlighted by…
-
Multiverse Computing launches Quasar 438B, Europe's leading AI model
Multiverse Computing has launched Quasar 438B, a new large language model designed for enterprise-scale agents and coding tasks. The model boasts strong performance on the Artificial Analysis Intelligence Index, scoring…
-
Zhipu AI's GLM-5.3-Flash offers 1/40th Claude Opus cost for agents
GLM-5.3-Flash, an open-source model from Zhipu AI, is significantly cheaper than competitors like Claude Opus 4.8, costing approximately 1/40th per token. This cost-effectiveness is particularly impactful for agent work…
-
Zhipu AI launches GLM-5.3-Flash, supported by SenseTime's domestic compute infrastructure
Zhipu AI has launched its first native multimodal model, GLM-5.3-Flash (320B-A18B), which achieves a score of 57 on the Artificial Analysis Intelligence Index, matching Anthropic's Claude Opus 4.8. The model is designed…
-
Z AI releases GLM-5.3-Flash with 400k context window
GLM-5.3-Flash, a proprietary multimodal model developed by Z AI, has been released with a 400k token context window. It performs well on the Artificial Analysis Intelligence Index, scoring 57, which is significantly abo…
-
GLM-5.3 ties Kimi K3 on AI Index, awaits full release · 2 sources tracked
The GLM-5.3 model from Z.ai has achieved a score of 60 on the Artificial Analysis Intelligence Index, matching the performance of Kimi K3. This represents a 7-point increase from its predecessor, GLM-5.2. While the mode…
-
GLM-5.3 (max) scores 60 on AI Intelligence Index, offers 1M context window
GLM-5.3 (max) has been evaluated on the Artificial Analysis Intelligence Index, achieving a score of 60, which is significantly above the median of 35 for comparable models. This proprietary model, developed by Z AI, bo…
-
Alibaba's Qwen3.8-27B matches GPT 5.6 Luna performance, runs locally
Alibaba's Qwen team has released Qwen3.8-27B, an open-weight model that achieves a score of 52 on the Artificial Analysis Intelligence Index. This new model reportedly matches the performance of GPT 5.6 Luna while also …
-
Qwen 3.8 27B model scores 52 on AI Index, matching GPT 5.6 Luna · 8 sources tracked
The Qwen 3.8 27B model has achieved a score of 52 on the Artificial Analysis Intelligence Index, matching GPT 5.6 Luna and nearing GLM-5.2 and DeepSeek V4 Pro 0813. Developed by Alibaba Group, this open-weights model su…
-
SpaceXAI upgrades Grok 4.6 with post-training enhancements, not new foundation
SpaceXAI has released Grok 4.6, which is an upgraded version of Grok 4.5 rather than a new foundational model. This post-training enhancement has significantly improved its performance on benchmarks like the Artificial …
-
OpenAI's GPT-5.6 Sol excels at code generation but struggles with database population
OpenAI has released GPT-5.6 Sol, which demonstrates significant improvements in coding tasks and token efficiency, outperforming previous models like Claude Opus 4.8 in benchmark tests. However, the model struggles with…
-
xAI's Grok 4.6 matches GPT-5.6 Sol, emphasizes agent endurance and tool integration
xAI has released Grok 4.6, which matches OpenAI's GPT-5.6 Sol on the Artificial Analysis Intelligence Index. However, the key innovation lies not in the benchmark score, but in Grok 4.6's focus on long-running agents ca…
-
xAI's Grok 4.6 matches OpenAI's GPT-5.6 Sol and outperforms Claude Opus 5 on agentic tasks
xAI's Grok 4.6 has achieved parity with OpenAI's GPT-5.6 Sol on the Artificial Analysis Intelligence Index. Furthermore, Grok 4.6 demonstrated superior efficiency in agentic tasks, completing them in 53 steps compared t…
-
SpaceXAI's Grok 4.6 hits AI frontier with strong agentic performance and lower cost · 4 sources tracked
SpaceXAI's Grok 4.6 has achieved a score of 61 on the Artificial Analysis Intelligence Index, placing it among the top-tier AI models. This new version shows significant improvement in agentic performance and turn effic…
-
SpaceXAI releases Grok 4.6, matching GPT-5.6 Sol on key benchmarks
SpaceXAI has released Grok 4.6, a new frontier intelligence model that significantly improves upon Grok 4.5, particularly in handling long-running agents and complex interactive tasks. The model matches GPT-5.6 Sol on t…