Artificial Analysis Intelligence Index
PulseAugur coverage of Artificial Analysis Intelligence Index — every cluster mentioning Artificial Analysis Intelligence Index across labs, papers, and developer communities, ranked by signal.
- instance of Claude Fable-5 90%
- instance of Opus 4.8 90%
- instance of Claude Opus-5 90%
- instance of GLM-5.2 90%
- instance of GPT 5.6 "Sol" 90%
- competes with Kimi k3 70%
- competes with GPT 5.6 "Sol" 70%
- competes with Claude (Opus 4.8) 70%
- instance of Claude (Opus 4.8) 70%
- affiliated with GPT 5.6 "Sol" 50%
14 day(s) with sentiment data
-
SpaceXAI's Grok 4.6 hits AI frontier with strong agentic performance and lower cost · 4 sources tracked
SpaceXAI's Grok 4.6 has achieved a score of 61 on the Artificial Analysis Intelligence Index, placing it among the top-tier AI models. This new version shows significant improvement in agentic performance and turn effic…
-
Meta's Muse Spark 1.2 shows rapid performance gains, rivals top AI models
Meta's latest foundational model, Muse Spark 1.2, has achieved high scores in third-party performance analyses, demonstrating rapid improvement since the Muse series' debut four months ago. The model notably surpassed G…
-
Qwen3.8 Max matches Claude Opus 4.8, but Kimi k3 remains more cost-effective
Alibaba's Qwen3.8 Max has significantly improved its performance, achieving a score of 56 on the Artificial Analysis Intelligence Index. This marks a substantial leap from its previous version, Qwen3.7 Max, which scored…
-
Chinese AI market sees dual competition: high capability and cost-efficiency
Huatai Securities suggests focusing on AI applications and domestic models as key investment areas, driven by a shift in large model competition towards cost-effectiveness. OpenAI has reduced prices for its Terra and Lu…
-
Four Chinese AI Labs Release Frontier Models in Summer 2026 · 1 source tracked
In the summer of 2026, four major Chinese AI labs released frontier-scale models, marking a significant shift towards open-weight availability. Kimi K3 from Moonshot achieved top third-party verified scores on the Artif…
-
2026 LLM Benchmark: No Single Winner, Specialized Leaders Emerge · 1 source tracked
A comprehensive benchmark of 20 leading LLMs in 2026 reveals no single dominant model, but rather specialized leaders across different tasks. Claude Opus 5 leads the overall Artificial Analysis Intelligence Index, while…
-
Anthropic launches Claude Opus 5, outperforming Fable 5 and GPT-5.6 Sol
Anthropic has released its new flagship model, Claude Opus 5, which demonstrates superior performance on various benchmarks, including agentic coding and knowledge work, outperforming models like Fable 5 and GPT-5.6 Sol…
-
Anthropic's Claude Opus 5 claims top AI benchmark spots, undercuts pricing · 1 source tracked
Anthropic has released Claude Opus 5, a new flagship AI model that has achieved top rankings on independent benchmarks like the Artificial Analysis Intelligence Index and the Agentic Index. The model significantly outpe…
-
Anthropic's Claude Opus 5 doubles cost for max effort, offers 1M context
Anthropic has released Claude Opus 5, which offers a 1M context window and five effort levels, with 'thinking' enabled by default. While the base API pricing remains unchanged from Opus 4.8, the new 'max effort' setting…
-
Anthropic's Claude Opus 5 shows mixed results, leads intelligence benchmarks
Anthropic has released Claude Opus 5, which is being integrated into various products like Claude Code. Early users report mixed experiences, with some finding it highly agentic and capable of complex tasks, while other…
-
Anthropic launches Claude Opus 5 with aggressive pricing, rivals IPO-bound OpenAI
Anthropic has launched Claude Opus 5, a new AI model that rivals top-tier competitors in intelligence while offering significantly lower pricing, reportedly 26% of the cost of models like Fable 5. This move is seen as a…
-
Kimi K3 model shows strong performance but questions arise over its 'open-source' cost claims
Moonshot's Kimi K3 model has achieved a notable position on the Artificial Analysis Intelligence Index, ranking closely behind leading models like Claude Fable-5 and GPT-5.6 Sol max. Despite its strong performance, the …
-
Moonshot AI releases Kimi K3, challenging frontier AI models with open weights
Moonshot AI has announced Kimi K3, a new 2.8 trillion parameter open-weight model with a 1 million token context window, positioning it as a strong contender in the frontier AI space. While benchmarks suggest Kimi K3 pe…
-
OpenAI removes Codex cap, introduces new cost-effective models Sol, Terra, Luna
OpenAI has reportedly removed the five-hour usage cap for its Codex coding model, replacing it with a weekly limit for some users. This change is seen as a significant improvement for those who use Codex for extensive c…
-
Meta shifts from open-weight AI, frontier moves to China
Meta, a key player in open-weight LLMs, has shifted its focus away from releasing open models, with its last public releases being Llama 4 Scout and Maverick in April 2025. The company's more recent frontier model, Muse…
-
OpenBMB releases MiniCPM5-1B, an efficient edge AI model for older phones
OpenBMB, a lab from Tsinghua University, has released MiniCPM5-1B, a new edge AI model designed for efficiency on older devices. This 1.08 billion parameter model features a 131K context window and native tool calling c…
-
SpaceXAI's Grok 4.5 ranks fourth on AI benchmark
SpaceXAI's Grok 4.5 has achieved a score of 54 on the Artificial Analysis Intelligence Index, securing the fourth position. This performance places it behind other leading AI models in the benchmark.
-
Anthropic releases Claude Sonnet 5, nearing Opus performance at lower cost
Anthropic has released Claude Sonnet 5, a new mid-tier model designed for enhanced agentic capabilities, including planning, tool use, and autonomous operation. While priced comparably to its predecessor, Sonnet 4.6, an…
-
Anthropic's Fable 5 removed amid export controls; GLM-5.2 takes top open-weight spot
Anthropic's Fable 5 model was available for only 96 hours before being removed globally due to US export controls. Four days later, Z.ai released GLM-5.2 under an MIT license, which quickly rose to fourth overall in the…
-
Anthropic's Claude Opus 4.8 claims AI crown as OpenAI retires GPT-4.5
OpenAI is retiring several of its older AI models, including GPT-4.5 and o3, with GPT-4.5 being removed from ChatGPT on June 27, 2026. This move is seen as a strategic shift ahead of potential IPO plans and the release …