GLM~5.1
PulseAugur coverage of GLM~5.1 — every cluster mentioning GLM~5.1 across labs, papers, and developer communities, ranked by signal.
- developed by Zhipu AI 95%
- instance of Zhipu AI 95%
- instance of GLM-5.2 90%
- developed GLM-5.2 90%
- developed by GLM-5.2 90%
- uses Fireworks AI 90%
- competes with Claude (Opus 4.8) 80%
- competes with Zhipu AI 70%
- affiliated with Zhipu AI 70%
- competes with DeepSeek V4-Pro 70%
- used by DeepSeek V4-Pro 70%
- competes with SWE Bench Pro 70%
- 2026-06-15 product_launch Together AI has launched GLM 5.1, an open-source inference model. source
- 2026-06-09 research_milestone GLM-5.1 achieved a leading score on the SWE-Bench Pro benchmark, surpassing proprietary models. source
- 2026-05-11 product_launch Zai.org open-sourced its GLM-5.1 large language model. source
8 day(s) with sentiment data
-
CI check manages Chinese LLM model names and token budgets
A developer has created a CI check to manage the rapidly changing landscape of Chinese LLM model names and their associated token budgets. This tool helps ensure production stability by treating model catalogs as deploy…
-
PIVOT method accelerates long-context AI models by optimizing sparse attention
A new method called PIVOT has been developed to optimize the performance of Dynamic Sparse Attention (DSA) models, particularly for handling long contexts. PIVOT addresses a bottleneck in DSA's indexer, which previously…
-
MindForge pipeline trains small LLMs for full software engineering lifecycle
Researchers have developed MindForge, an automated pipeline designed to train smaller language models in comprehensive software engineering tasks. This system converts open-source command-line programs into source-free …
-
PIVOT indexing method accelerates sparse attention in LLMs
Researchers have developed PIVOT, a novel indexing method designed to optimize token-level sparse attention in large language models. PIVOT addresses the bottleneck created by indexers in systems like DeepSeek Sparse At…
-
Zhipu AI's GLM-5.2 leads open-weight models with 1M context window · 1 source tracked
GLM-5.2, a new open-weight model from China's Zhipu AI, has been recognized as the top-performing open model as of July 2026. It achieved a score of 51 in the Intelligence Index v4.1 by Artificial Analysis, placing it f…
-
LLM benchmark results reveal performance across multiple models · 9 sources tracked
A recent independent benchmark evaluation has revealed performance metrics for several large language models, including Kimi K2, Sarvam Maya, NVIDIA Nemotron 3 Super 120B, DeepSeek V3.2, Falcon H1R-7B, GLM-5.2, GLM-5.1,…
-
New Red Teaming Protocol Finds Hidden Bugs in LLM-Generated Code
Researchers have developed a new protocol called Code Monitor Red Teaming to identify hidden bugs in LLM-generated code that has already passed public tests. The study, which used the CodeMonitorBench benchmark with ove…
-
Anomaly launches $10/month OpenCode Go for curated AI coding models
Anomaly has launched OpenCode Go, a $10/month subscription service that provides access to a curated selection of 13 open-source AI coding models. The service aims to offer a reliable and affordable way for developers t…
-
Tencent's Hy3 open MoE model beats GLM-5.1 in blind test
Tencent has released Hy3, a 21 billion active parameter open Mixture-of-Experts model that outperformed Zhipu's GLM-5.1 in a blind test. The model, which has 295 billion total parameters and a 256K context window, is av…
-
Tencent releases Hy3 model with 295B parameters and 256K context
Tencent has released Hy3, an open-weights AI model with 295 billion parameters, featuring a 21 billion active parameter inference and a 3.8 billion parameter prediction head. The model boasts a 256K context window and u…
-
Open-weight LLMs are free to access but costly to run, challenging developers
The article argues that while open-weight large language models (LLMs) are technically free to access, their immense size often makes them prohibitively expensive and difficult to run on standard hardware. Models from Q…
-
Tencent's Hy3 model shows strong performance in real-world tasks
The Hy3 benchmark results show the model performing comparably to other advanced models like DeepSeek v4 and GLM-5.1. Tencent's Hy3 achieved a score of 2.67/4 on 312 real-world workflow tasks, outperforming GLM-5.1's sc…
-
Tencent releases Hy3, an open 295B MoE model with 256K context
Tencent has released Hy3, an open-source 295 billion parameter Mixture-of-Experts (MoE) model designed for complex reasoning, agentic workflows, and long-context tasks. The model activates only 21 billion parameters per…
-
AI efficiency prompts layoffs at major Chinese tech firms
A wave of layoffs is impacting major Chinese tech companies, with AI efficiency gains cited as a primary driver. Junior engineers and high-performing employees alike are being let go, as companies seek to streamline ope…
-
11 LLMs evaluated on code refactoring and proposal evaluation
An experiment evaluated eleven large language models on their ability to refactor a complex "god node" within a LangGraph agent. The models were tasked with proposing solutions to untangle the node's logic and then eval…
-
Mastermind framework boosts AI agents' vulnerability reproduction success
Researchers have developed a new framework called Mastermind to improve the performance of AI agents in complex software engineering tasks, specifically vulnerability reproduction. This framework separates the learning …
-
Mastermind framework improves AI agents' vulnerability reproduction capabilities
A new dual-loop framework called Mastermind has been introduced to enhance the ability of software engineering agents to reproduce repository-scale vulnerabilities. This framework separates strategy learning from task-s…
-
China's GLM-5.2 open-source model challenges top closed-source AI · 1 source tracked
Chinese AI lab Zhipu AI has released GLM-5.2, an open-weight model that rivals top-tier closed-source models like Claude Opus and GPT-5.5 in performance, particularly in coding and long-context tasks. GLM-5.2 supports a…
-
Chinese AI models tested on coding tasks: MiniMax and Kimi lead
A comparative analysis of five Chinese AI models—MiniMax M3, Kimi K2.6, DeepSeek V4 Pro, Qwen 3.7 Max, and GLM 5.1—evaluated on real-world engineering tasks revealed significant differences in their coding capabilities.…
-
AI models struggle to manage virtual companies; Claude Fable 5 leads with $47M profit · 1 source tracked
A recent CEO-Bench competition, designed to test AI's ability to run a virtual SaaS startup, revealed mixed results. While many advanced AI models like GLM 5.1 and Gemini 3 Flash went bankrupt, Claude Fable 5 emerged as…