PulseAugur
EN
LIVE 13:18:43
ENTITY SWE-Marathon

SWE-Marathon

PulseAugur coverage of SWE-Marathon — every cluster mentioning SWE-Marathon across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
RECENT · PAGE 1/1 · 4 TOTAL
  1. SIGNIFICANT · CL_134472 ·

    xAI unveils Grok 4.5, targeting programmers and office tasks

    xAI has launched Grok 4.5, its most advanced model to date, specifically designed for programming, agentic tasks, and knowledge work. Trained with Cursor on NVIDIA GB300 GPUs, Grok 4.5 shows strong performance in coding…

  2. TOOL · CL_121794 ·

    AI models struggle with marathon tasks, revealing benchmark limitations · 1 source tracked

    New benchmarks reveal a significant gap between AI models' performance on short, single-session tasks and their ability to handle long, multi-hour operations. While models like GLM-5.2 and GPT-5.5 excel on benchmarks li…

  3. SIGNIFICANT · CL_117035 ·

    Zhipu AI releases GLM-5.2 with 1M context window, challenging top proprietary models

    Zhipu AI has released GLM-5.2, a 744B-parameter Mixture-of-Experts model featuring a 1 million token context window and MIT-licensed weights. This model achieves a high ranking on the BenchLM leaderboard and demonstrate…

  4. FRONTIER RELEASE · CL_92810 ·

    Z.ai releases GLM-5.2, setting new open-source benchmark for long-context AI

    Z.ai has released GLM-5.2, an open-source language model with a 1 million token context window, positioning it as a strong contender in long-horizon tasks and coding benchmarks. The model features an improved architectu…