PulseAugur
实时 00:23:00
English(EN) RT @nutlope: Built a visual benchmark where I asked closed and open source models to build small games.

开源AI模型在成本和速度基准测试中表现优于闭源模型

由nutlope开发并由Together AI在X上分享的一个新的视觉基准测试,突显了AI模型之间显著的成本和速度差异。该基准测试测试了Opus 4.8和GPT-5.5等闭源模型与开源替代品,发现开源模型速度更快、成本更低。具体而言,Opus 4.8的成本是MiniMax M3的15倍,GPT-5.5的成本是MiniMax M3的10倍,而两种类型的模型产生的游戏质量相当。 AI

影响 开源模型展示了成本和速度优势,可能加速其相对于专有替代品的采用。

排序理由 该集群描述了一个比较AI模型的新基准测试,属于研究范畴。[lever_c_demoted from research: ic=1 ai=1.0]

在 X — Together (inference / OSS) 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

开源AI模型在成本和速度基准测试中表现优于闭源模型

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一个比较AI模型的新基准测试,属于研究范畴。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
87 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [1]

  1. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    RT @nutlope: 我构建了一个视觉基准测试,让闭源和开源模型构建小型游戏。

    RT @nutlope: Built a visual benchmark where I asked closed and open source models to build small games. Main takeaway: OSS models were a l…