Artificial Analysis
PulseAugur coverage of Artificial Analysis — every cluster mentioning Artificial Analysis across labs, papers, and developer communities, ranked by signal.
- instance of Fun-Realtime-TTS 95%
- instance of Qwen-Audio-3.0-TTS-Plus 95%
- instance of GLM-5.2 90%
- used by Fireworks AI 90%
- instance of Qwen-Audio-3.0-TTS 90%
- developed Intelligence indexes generalist genes for cognitive abilities 90%
- instance of MiniMax M2.7 80%
- competes with Kimi k3 70%
- competes with Claude Fable-5 70%
- competes with GLM-5.2 70%
- competes with GPT 5.6 "Sol" 70%
- competes with Opus 4.8 70%
- 2026-06-16 research_milestone Artificial Analysis released an updated version of its Intelligence Index, version 4.1, which includes a greater emphasis on agentic workloads and improved benchmarks. source
26 day(s) with sentiment data
-
SpaceXAI's Grok 4.6 hits AI frontier with strong agentic performance and lower cost · 4 sources tracked
SpaceXAI's Grok 4.6 has achieved a score of 61 on the Artificial Analysis Intelligence Index, placing it among the top-tier AI models. This new version shows significant improvement in agentic performance and turn effic…
-
Meta releases open-weight Muse Glimmer; OpenAI, Anthropic, NVIDIA also launch new models · 2 sources tracked
Meta has re-entered the open-weight frontier with the release of Muse Glimmer, a 30B multimodal, agent-focused model optimized for local deployment and consumer hardware. Concurrently, OpenAI launched GPT-5.6-Cyber, a r…
-
DeepSeek's V4-Flash-0731 model achieves superior agent performance via post-training
DeepSeek has released V4-Flash-0731, an updated version of its 284 billion parameter model that outperforms its previous flagship, V4-Pro-Preview, on several agent benchmarks. The performance gains were achieved through…
-
Google's Gemini 3.6 Flash outscores 'Pro' version in new benchmarks
Google has updated its Gemini offerings, introducing Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, while Gemini 3.1 Pro remains the default 'Pro' option in the Gemini application. Independent benchmarks from Artificial An…
-
AI landscape analysis guide released by Artificial Analysis
Artificial Analysis has released a comprehensive guide to understanding the AI landscape, aiming to help users select the most suitable models and providers for their specific needs. This resource provides a centralized…
-
Anthropic's Claude Opus 5 leads agentic index, Qwen3.8 Max close behind · 1 source tracked
Artificial Analysis's Agentic Index shows Anthropic's Claude Opus 5 leading, with Alibaba Group's Qwen3.8 Max closely following. While some reports incorrectly declared Qwen the top model, the index actually places Clau…
-
Artificial Analysis launches agentic AI index
Artificial Analysis has launched a new initiative focused on agentic AI, aiming to create a comprehensive index of AI agents. This project seeks to catalog and understand the rapidly evolving landscape of autonomous AI …
-
Alibaba launches CosyVoice Studio AI voice platform
Alibaba has launched CosyVoice Studio, a comprehensive AI voice platform powered by their Qwen-Audio model. This platform offers three main features: a voice keyboard that refines spoken language into structured text by…
-
User alleges bias in AI intelligence index favoring Anthropic
A user on r/LocalLLaMA has raised concerns about the perceived bias in Artificial Analysis's intelligence index. The user alleges that the index was adjusted to lower the ranking of an open-source model, Qwen 3.8 max, s…
-
AMD acquires Taalas to boost AI inference hardware
AMD has acquired Taalas, a company focused on custom ASICs for AI inference, signaling a strategic move to bolster its AI hardware capabilities. This acquisition aligns with the trend of vertical integration in the AI s…
-
Qwen3.8 Max leads AI model rankings on Agentic Index · 3 sources tracked
The Qwen3.8 Max model has been ranked as the top-performing AI model according to the Artificial Analysis Agentic Index. This index evaluates models on various capabilities, including agentic real-world work tasks, codi…
-
Open AI models narrow capability gap but lag in enterprise adoption and serving stack performance
Open-weight AI models have significantly closed the capability gap with proprietary models, reaching within 6 points on the Intelligence Index by April 2026. Despite this, enterprise adoption of open models has lagged, …
-
Runway ML's video models drop in rankings; subscription adds competitors
Runway ML's video generation models, Gen-4 and Gen-4.5, have fallen out of the top 20 on the Artificial Analysis leaderboard, according to July 2026 data. This is a significant drop from their claimed performance eight …
-
Alibaba and DeepSeek launch new AI models, intensifying China's race for cost-efficiency
Alibaba has released its largest AI model to date, Qwen3.8-Max, featuring 2.4 trillion parameters and a mixture-of-experts architecture that activates approximately 95 billion parameters per request. Concurrently, DeepS…
-
Fireworks AI launches index to verify model accuracy on inference endpoints
Fireworks AI has introduced the Artificial Analysis Endpoint Accuracy Index to measure how well serverless API endpoints maintain the accuracy of open-weight models. The index initially covers GLM-5.2, GPT-OSS 120B, and…
-
Kandinsky 5.0 vs. GPT Image 2 & Nano Banana 2: Russian prompt understanding compared
A comparison of text-to-image models reveals that while GPT Image 2 and Nano Banana 2 lead in general image generation based on blind human ratings, they face accessibility issues for users in Russia. Kandinsky 5.0, dev…
-
Kling AI's video model ranked 11th despite high review score
Kling AI's latest model, Kling 3.0 Pro, has been ranked eleventh in an independent image-to-video leaderboard, falling behind competitors like Veo 3.1 and Gemini Omni Flash. Despite this ranking, a separate review from …
-
OpenAI's GPT-5.6 Sol cuts costs by 20%, Astra makes math breakthroughs
OpenAI has achieved significant advancements with its frontier models, including GPT-5.6 Sol which autonomously optimized production GPU kernels, reducing serving costs by 20%. Concurrently, an internal model named Astr…
-
DeepSeek V4-Flash AI model leads in affordability, study finds · 2 sources tracked
A new study by Artificial Analysis has found that DeepSeek's V4-Flash AI model is the most affordable to operate among major models. The V4-Flash model significantly undercuts competitors like Anthropic's Claude Fable 5…
-
DeepSeek V4 Flash-0731 introduces 'kill line' concept, reshaping AI model market
DeepSeek has released its V4 Flash-0731 model, a 284B-parameter AI that challenges top-tier models like Anthropic's Opus 4.8 and GLM-5.2 in performance while offering a significantly lower price point. This release has …