PulseAugur
中
实时 10:01:58
Dansk(DA) Read our blog: https://t.co/oZ4le7JlpY

Together AI 为开放模型推出预留吞吐量

Together AI 推出了预留吞吐量(Provisioned Throughput)服务,该服务提供开源前沿模型的预留推理容量。该服务采用基于 token 的定价和 99% 的正常运行时间服务水平协议(SLA)。它旨在提供无服务器的简便性和保证的容量,同时成本显著低于 Opus 4.8 等现有解决方案,并初步支持 MiniMax M3 和 GLM-5.2 等模型。 AI

影响 为推理访问前沿开源模型提供了更具成本效益和可靠性的方式。

排序理由 AI基础设施提供商推出新服务。

在 X — Together (inference / OSS) 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Together AI 为开放模型推出预留吞吐量

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
AI基础设施提供商推出新服务。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
93 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [2]

  1. X — Together (inference / OSS) TIER_1 Dansk(DA) · togethercompute ·

    阅读我们的博客:https://t.co/oZ4le7JlpY

    Read our blog: https://t.co/oZ4le7JlpY

  2. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    我们推出预留吞吐量:面向前沿开放模型的预留推理容量,采用按令牌计费和99%的正常运行时间服务水平协议。

    We're introducing Provisioned Throughput: reserved inference capacity for frontier open models, with token-based pricing and a 99% uptime SLA. Serverless simplicity, guaranteed capacity, up to 90% lower cost vs. Opus 4.8. Get started with MiniMax M3 + GLM-5.2, read more 🧵 https…