Introducing preemptible compute: the same compute, half the price

Together AI launched a public preview of preemptible GPU compute for its GPU Clusters, letting teams add interruptible nodes billed at half the on‑demand rate. Nodes use the same hardware, are reclaimed with a five‑minute drain, and automatically refill toward the target, offering cheaper capacity for tolerant workloads.

Cover image for Introducing preemptible compute: the same compute, half the price