introducing preemptible compute for together gpu clusters
same nvidia gpu infrastructure, 50% of the on-demand price
built for evals, fine-tuning, batch inference + short experiments, with up to 5 minutes to checkpoint before reclamation
now in public preview