PulseAugur
EN
LIVE 22:32:21
ENTITY Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index

PulseAugur coverage of Artificial Analysis Intelligence Index — every cluster mentioning Artificial Analysis Intelligence Index across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
9
38 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

7 day(s) with sentiment data

RECENT · PAGE 1/3 · 46 TOTAL
  1. TOOL · CL_256352 ·

    K2 Horizon 7B model shows strong performance for its size

    The K2 Horizon 7B model has demonstrated strong performance, positioning itself between the Qwen 3.6 27B and Qwen 3.6 35BA3b models on the Artificial Analysis Intelligence Index. Initial testing suggests the model is so…

  2. SIGNIFICANT · CL_249499 ·

    Alibaba's Qwen3.8-27B model achieves competitive performance on Cerebras hardware

    Alibaba's Qwen3.8-27B model is now operational on Cerebras hardware, offering rapid inference capabilities. This open-weight model achieves a score of 34 on the Artificial Analysis Intelligence Index. This performance p…

  3. SIGNIFICANT · CL_249412 ·

    DeepSeek launches V4.1-Flash with novel architecture for efficient AI agents

    DeepSeek has released V4.1-Flash, a new open-weight flagship model featuring a novel causal Encoder-Decoder architecture. This model emphasizes extreme inference efficiency and cost-effectiveness, with a 763 billion par…

  4. TOOL · CL_247499 ·

    2.5B parameter MiniCPM5 model challenges larger LLMs on benchmarks

    The MiniCPM5–2B model, developed by OpenBMB, represents a significant advancement in smaller, highly capable language models. Despite its 2.5 billion parameters and a 1.04 GB file size, it outperforms larger models like…

  5. COMMENTARY · CL_243400 ·

    OpenAI's GPT-6 Astra shows major gains, especially in graphics, per analysis

    Sebastian Raschka's analysis suggests OpenAI's new GPT-6 Astra model demonstrates significant improvements over its predecessor, GPT-5.6 "Sol," particularly in graphical tasks and achieving a near-perfect score on the A…

  6. COMMENTARY · CL_234793 ·

    Google releases rapid-fire Gemini Flash models amid flagship delay

    Google has released four Gemini Flash models in rapid succession, with the latest, Gemini 3.8 Flash, excelling in coding and matching larger models' performance at a lower cost. This rapid release pace is highlighted by…

  7. SIGNIFICANT · CL_232270 ·

    Multiverse Computing launches Quasar 438B, Europe's leading AI model

    Multiverse Computing has launched Quasar 438B, a new large language model designed for enterprise-scale agents and coding tasks. The model boasts strong performance on the Artificial Analysis Intelligence Index, scoring…

  8. RESEARCH · CL_228519 ·

    Zhipu AI's GLM-5.3-Flash offers 1/40th Claude Opus cost for agents

    GLM-5.3-Flash, an open-source model from Zhipu AI, is significantly cheaper than competitors like Claude Opus 4.8, costing approximately 1/40th per token. This cost-effectiveness is particularly impactful for agent work…

  9. SIGNIFICANT · CL_222085 ·

    Zhipu AI launches GLM-5.3-Flash, supported by SenseTime's domestic compute infrastructure

    Zhipu AI has launched its first native multimodal model, GLM-5.3-Flash (320B-A18B), which achieves a score of 57 on the Artificial Analysis Intelligence Index, matching Anthropic's Claude Opus 4.8. The model is designed…

  10. SIGNIFICANT · CL_220329 ·

    Z AI releases GLM-5.3-Flash with 400k context window

    GLM-5.3-Flash, a proprietary multimodal model developed by Z AI, has been released with a 400k token context window. It performs well on the Artificial Analysis Intelligence Index, scoring 57, which is significantly abo…

  11. RESEARCH · CL_207960 ·

    GLM-5.3 ties Kimi K3 on AI Index, awaits full release · 2 sources tracked

    The GLM-5.3 model from Z.ai has achieved a score of 60 on the Artificial Analysis Intelligence Index, matching the performance of Kimi K3. This represents a 7-point increase from its predecessor, GLM-5.2. While the mode…

  12. RESEARCH · CL_207977 ·

    GLM-5.3 (max) scores 60 on AI Intelligence Index, offers 1M context window

    GLM-5.3 (max) has been evaluated on the Artificial Analysis Intelligence Index, achieving a score of 60, which is significantly above the median of 35 for comparable models. This proprietary model, developed by Z AI, bo…

  13. SIGNIFICANT · CL_205875 ·

    Alibaba's Qwen3.8-27B matches GPT 5.6 Luna performance, runs locally

    Alibaba's Qwen team has released Qwen3.8-27B, an open-weight model that achieves a score of 52 on the Artificial Analysis Intelligence Index. This new model reportedly matches the performance of GPT 5.6 Luna while also …

  14. RESEARCH · CL_205136 ·

    Qwen 3.8 27B model scores 52 on AI Index, matching GPT 5.6 Luna · 8 sources tracked

    The Qwen 3.8 27B model has achieved a score of 52 on the Artificial Analysis Intelligence Index, matching GPT 5.6 Luna and nearing GLM-5.2 and DeepSeek V4 Pro 0813. Developed by Alibaba Group, this open-weights model su…

  15. COMMENTARY · CL_203441 ·

    SpaceXAI upgrades Grok 4.6 with post-training enhancements, not new foundation

    SpaceXAI has released Grok 4.6, which is an upgraded version of Grok 4.5 rather than a new foundational model. This post-training enhancement has significantly improved its performance on benchmarks like the Artificial …

  16. SIGNIFICANT · CL_203047 ·

    OpenAI's GPT-5.6 Sol excels at code generation but struggles with database population

    OpenAI has released GPT-5.6 Sol, which demonstrates significant improvements in coding tasks and token efficiency, outperforming previous models like Claude Opus 4.8 in benchmark tests. However, the model struggles with…

  17. SIGNIFICANT · CL_198343 ·

    xAI's Grok 4.6 matches GPT-5.6 Sol, emphasizes agent endurance and tool integration

    xAI has released Grok 4.6, which matches OpenAI's GPT-5.6 Sol on the Artificial Analysis Intelligence Index. However, the key innovation lies not in the benchmark score, but in Grok 4.6's focus on long-running agents ca…

  18. SIGNIFICANT · CL_197428 ·

    xAI's Grok 4.6 matches OpenAI's GPT-5.6 Sol and outperforms Claude Opus 5 on agentic tasks

    xAI's Grok 4.6 has achieved parity with OpenAI's GPT-5.6 Sol on the Artificial Analysis Intelligence Index. Furthermore, Grok 4.6 demonstrated superior efficiency in agentic tasks, completing them in 53 steps compared t…

  19. FRONTIER RELEASE · CL_197109 ·

    SpaceXAI's Grok 4.6 hits AI frontier with strong agentic performance and lower cost · 4 sources tracked

    SpaceXAI's Grok 4.6 has achieved a score of 61 on the Artificial Analysis Intelligence Index, placing it among the top-tier AI models. This new version shows significant improvement in agentic performance and turn effic…

  20. FRONTIER RELEASE · CL_196976 ·

    SpaceXAI releases Grok 4.6, matching GPT-5.6 Sol on key benchmarks

    SpaceXAI has released Grok 4.6, a new frontier intelligence model that significantly improves upon Grok 4.5, particularly in handling long-running agents and complex interactive tasks. The model matches GPT-5.6 Sol on t…