Together AI has launched a public preview of preemptible compute for its GPU Clusters, available on Kubernetes. This new offering allows users to access the same GPU infrastructure at half the on-demand price for interruption-tolerant tasks like short experiments, inference bursts, and batch jobs. Preemptible nodes are billed sub-hourly and can be reclaimed with a five-minute warning, making them suitable for workloads that can checkpoint or retry. AI
IMPACT Reduces the cost of running AI workloads, potentially accelerating experimentation and inference for users of Together AI's platform.
RANK_REASON This is a new product offering from an AI infrastructure provider, not a frontier model release or core research.
- GLM5.3 Flash
- Kubernetes
- Nvidia
- PyTorch Lightning
- Ray Train
- SIGTERM
- Together AI
- Together GPU Clusters
- TogetherPreempted
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →