PulseAugur
EN
LIVE 11:09:26

Together AI launches Provisioned Throughput for open models

Together AI has launched Provisioned Throughput, a new service offering reserved inference capacity for open-source frontier models. This service features token-based pricing and a 99% uptime Service Level Agreement (SLA). It aims to provide serverless simplicity and guaranteed capacity at a significantly lower cost compared to existing solutions like Opus 4.8, with initial support for models such as MiniMax M3 and GLM-5.2. AI

IMPACT Offers a more cost-effective and reliable way to access frontier open-source models for inference.

RANK_REASON Launch of a new service offering by an AI infrastructure provider.

Read on X — Together (inference / OSS) →

AI-generated summary · Google Gemini · from 2 sources. How we write summaries →

Together AI launches Provisioned Throughput for open models

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
Launch of a new service offering by an AI infrastructure provider.
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
product, infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
93 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [2]

  1. X — Together (inference / OSS) TIER_1 Dansk(DA) · togethercompute ·

    Read our blog: https://t.co/oZ4le7JlpY

    Read our blog: https://t.co/oZ4le7JlpY

  2. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    We're introducing Provisioned Throughput: reserved inference capacity for frontier open models, with token-based pricing and a 99% uptime SLA.

    We're introducing Provisioned Throughput: reserved inference capacity for frontier open models, with token-based pricing and a 99% uptime SLA. Serverless simplicity, guaranteed capacity, up to 90% lower cost vs. Opus 4.8. Get started with MiniMax M3 + GLM-5.2, read more 🧵 https…