Together, an inference and open-source AI platform, has announced a price reduction for its dedicated H100 GPU instances. Starting in September, the hourly rate for these instances will decrease from $5.49 to $3.99. This new pricing applies automatically to both new and existing deployments, allowing users to deploy a variety of models including Gemma 4, Qwen3/3.5, gpt-oss, llama, and Nemotron 3.5 Lightning, or to bring their own LoRA fine-tuned models. AI
IMPACT Reduces the cost of running large AI models on dedicated hardware, potentially lowering barriers for developers and researchers.
RANK_REASON This is a price reduction announcement for a cloud inference provider, not a new model release or core research.
Read on X — Together (inference / OSS) →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →