Check these fields
- VRAM per GPU: allow space for model weights, activations, context, and your application’s overhead.
- GPU quantity: multiple GPUs require software that can use them. Their memory is not automatically one shared pool.
- Whole-instance hourly price: compare the configuration, not just the model name.
- Maximum duration and budget: make sure the available runtime covers your job.
- Storage: copy important results off the instance before it ends.
