PulseAugur
EN
LIVE 01:33:01

Together cuts H100 inference prices to $3.99/hr for September

Together, an inference and open-source AI platform, has announced a price reduction for its dedicated H100 GPU instances. Starting in September, the hourly rate for these instances will decrease from $5.49 to $3.99. This new pricing applies automatically to both new and existing deployments, allowing users to deploy a variety of models including Gemma 4, Qwen3/3.5, gpt-oss, llama, and Nemotron 3.5 Lightning, or to bring their own LoRA fine-tuned models. AI

IMPACT Reduces the cost of running large AI models on dedicated hardware, potentially lowering barriers for developers and researchers.

RANK_REASON This is a price reduction announcement for a cloud inference provider, not a new model release or core research.

Read on X — Together (inference / OSS) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Together cuts H100 inference prices to $3.99/hr for September

How we ranked this

Signal score
9 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
This is a price reduction announcement for a cloud inference provider, not a new model release or core research.
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
infra
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

Full methodology in our editorial standards.

COVERAGE [1]

  1. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    for september, we’re cutting the price of Dedicated Inference on H100s from $5.49/hr to $3.99/hr

    for september, we’re cutting the price of Dedicated Inference on H100s from $5.49/hr to $3.99/hr new + existing deployments get the lower price automatically deploy gemma 4, qwen3/3.5, gpt-oss, llama, nemotron 3.5 lightning models, or bring your own lora for a fine-tuned model …