Cloud providers sell unused GPU capacity at 60–90% discounts as spot or preemptible instances, but can terminate them with as little as 30 seconds' notice. Training jobs that checkpoint frequently can resume on new spot instances, making large training runs dramatically cheaper. Inference workloads generally avoid spot instances because the interruption would drop live user requests. Major providers (AWS, GCP, Azure, Lambda Labs) all offer spot GPU instances.