PulseAugur
EN
LIVE 05:32:20
ENTITY Agent Arena

Agent Arena

PulseAugur coverage of Agent Arena — every cluster mentioning Agent Arena across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
5 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
0 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 5 TOTAL
  1. TOOL · CL_178153 ·

    Alibaba's Qwen ranks second on Text Arena leaderboard

    Alibaba's Qwen model has achieved the second position on the Text Arena leaderboard. This ranking highlights the model's performance in comparative evaluations against other AI systems.

  2. TOOL · CL_153341 ·

    Kimi k3 matches Opus on non-vision tasks per Agent Arena

    A recent evaluation on Agent Arena suggests that Kimi k3 performs at a similar level to Opus for non-vision tasks. However, one user's testing on Android indicated that Opus's vision capabilities place it ahead of Kimi …

  3. SIGNIFICANT · CL_147389 ·

    Moonshot AI releases Kimi K3, largest open model with 1M context · 1 source tracked

    Moonshot AI has launched Kimi K3, a new open-weights frontier-class model boasting 2.8 trillion parameters and a 1 million token context window. The model features native multimodal input capabilities and a novel Kimi D…

  4. RESEARCH · CL_88579 ·

    Anthropic suspends Fable/Mythos models citing US gov directive

    Anthropic has suspended access to its Fable 5 and Mythos 5 models for all customers worldwide following a directive from the U.S. government, citing national cybersecurity risks. This abrupt revocation has disrupted dow…

  5. SIGNIFICANT · CL_101512 ·

    Anthropic's Fable 5 suspended by US export controls; GLM-5.2 emerges as top open-weight coding model

    Anthropic's Claude Fable 5 and Mythos 5 models faced significant disruption due to a US government export control directive, leading to their suspension for foreign nationals and impacting broader access. This event spa…