FrontierSWE
PulseAugur coverage of FrontierSWE — every cluster mentioning FrontierSWE across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
Bespoke Labs seeks researcher for long-horizon agent benchmarks
Bespoke Labs is seeking a researcher to develop and evaluate reinforcement learning environments and benchmarks for long-horizon agent tasks. The role requires demonstrated experience with multi-step reasoning agents, s…
-
China's Z.ai releases low-cost GLM 5.2, challenging US AI model dominance
Z.ai, a Beijing-based lab, has released GLM 5.2, an open-weights AI model that significantly undercuts the cost of leading U.S. frontier models. While GLM 5.2's API costs are a fraction of those for models like Anthropi…
-
Moonshot AI releases Kimi K3 with 1M context, open weights, and competitive pricing
Moonshot AI has released Kimi K3, an open-weights model with approximately 2.8 trillion parameters and a 1 million token context window. The model achieved strong performance on benchmarks like Terminal-Bench 2.0, Front…
-
Zhipu AI releases GLM-5.2 with 1M context window, challenging top proprietary models
Zhipu AI has released GLM-5.2, a 744B-parameter Mixture-of-Experts model featuring a 1 million token context window and MIT-licensed weights. This model achieves a high ranking on the BenchLM leaderboard and demonstrate…
-
China's GLM-5.2 challenges GPT-5.5 and Claude Opus on coding benchmarks
Zhipu AI's GLM-5.2, a Chinese frontier model, has reportedly achieved strong performance on coding benchmarks, surpassing OpenAI's GPT-5.5 and Anthropic's Claude Opus 4.7. On the FrontierSWE benchmark, GLM-5.2 scored 74…
-
Z.ai releases GLM-5.2, setting new open-source benchmark for long-context AI
Z.ai has released GLM-5.2, an open-source language model with a 1 million token context window, positioning it as a strong contender in long-horizon tasks and coding benchmarks. The model features an improved architectu…