PulseAugur
中
实时 06:18:16
中文(ZH) 顶流里最快!智谱,你是在「喷」代码吧

智谱AI推出GLM-5.1-highspeed API,速度达400 tokens/s

智谱AI发布了GLM-5.1-highspeed,这是其GLM-5.1模型的新API,推理速度达到每秒400个token。该新产品被定位为全球领先的LLM提供商中最快的,并在实际测试中表现出色,包括快速的代码生成和内容摘要。速度的提升归功于推理引擎、调度系统和底层基础设施在系统工程方面的显著优化,旨在通过减少等待时间和提高反馈频率来改善AI代理的用户体验。 AI

影响 加速了跨各种应用的AI代理响应能力和实时交互能力。

排序理由 前沿实验室的模型发布,创下新的速度基准。[lever_c_demoted from frontier_release: ic=2 ai=1.0]

在 量子位 (QbitAI) 阅读 →

AI 生成摘要 · Google Gemini · 来自 3 个来源。 我们如何撰写摘要 →

智谱AI推出GLM-5.1-highspeed API,速度达400 tokens/s

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Significant
前沿实验室的模型发布,创下新的速度基准。[lever_c_demoted from frontier_release: ic=2 ai=1.0]
Source corroboration
3 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release, product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
131 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+1 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

完整方法见我们的编辑标准。

报道来源 [3]

  1. 量子位 (QbitAI) TIER_1 中文(ZH) · 十三 ·

    顶流中的最快!智谱,你在“喷”代码吗?

    400 tokens/s

  2. Pandaily TIER_1 English(EN) · [email protected] (Pandaily) ·

    智谱AI发布GLM-5.1高速API:400 Tokens/s创全球新标杆

    Zhipu AI has launched GLM-5.1-highspeed, an API variant of its GLM-5.1 model delivering 400 tokens per second — reportedly the fastest inference speed among major global LLM providers.

  3. Mastodon — mastodon.social TIER_1 English(EN) · [email protected] ·

    智谱AI发布GLM-5.1-highspeed,其GLM-5.1大语言模型的加速API版本,每秒可达400个token,并据称创下了一项

    Zhipu AI has launched GLM-5.1-highspeed, a high-speed API variant of its GLM-5.1 large language model, delivering 400 tokens per second and reportedly setting a new global benchmark for inference speed among major LLM providers. The API targets enterprise applications requiring r…