Together has launched a new preemptible compute offering for its GPU clusters, available in public preview. This service utilizes the same Nvidia GPU infrastructure as their on-demand options but is priced at 50% of the on-demand cost. It is designed for tasks such as evaluations, fine-tuning, batch inference, and short experiments, with a grace period of up to five minutes for checkpointing before compute resources are reclaimed. AI
IMPACT Provides a more cost-effective option for AI training and inference tasks, potentially lowering the barrier to entry for experimentation.
RANK_REASON This is a product launch for a compute service, not a core AI model release or research milestone.
Read on X — Together (inference / OSS) →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →