2026 年秋季,大型语言模型的发布时间表前所未有,四大主要实验室在一个月内推出了新迭代。Anthropic 发布了 Claude Fable 5.1 和 Claude Opus 5.5,OpenAI 推出了 GPT-6 Astra、GPT-6 Sol 和 GPT-6 Luna,DeepSeek 推出了 V4.1 Flash,xAI 推出了 Grok 4.7。性能指标,包括 token 生成速度和首次 token 时间,在这些模型之间差异显著,其中 DeepSeek V4.1 Flash 在原始速度方面领先,DeepSeek 模型显示出快速的响应时间。然而,BitsMinds 和 DataCamp 等基准测试揭示了质量、速度和成本之间的权衡,表明不同的模型在不同的任务和价格点上表现出色。 AI
影响 为 LLM 发展设定了新的步伐,迫使竞争对手在性能和成本效益方面快速迭代。
排序理由 多个前沿实验室(Anthropic、OpenAI、DeepSeek、xAI)发布了具有特定名称和性能数据的新 LLM 版本。[lever_c_demoted from frontier_release: ic=1 ai=1.0]
- Anthropic
- Artificial Analysis
- BitsMinds
- Claude Fable 5.1
- Claude Opus 5.5
- Claude Sonnet 5.5
- DeepSeek
- DeepSeek V4.1 Flash
- DeepSeek V4-Pro
- GPT-6.1 Sol
- GPT-6 Astra
- GPT-6 Luna
- GPT-6 Sol
- Grok 4.7
- OpenAI
AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →