Use cases
GPU cloud use cases: right GPU, right mode, right price
Eight workload playbooks. Each one pairs the job with the best-value cards from our catalogue, the billing mode that fits its failure tolerance, and honest cost math you can rerun yourself.
- LLM training Pre-train and continue-train transformer models on HBM GPUs with NVLink and InfiniBand clusters. H100 SXM from 30%+ under market Top pick: H100 SXM $1.587/hr Open the playbook
- LLM inference Serve 7B–70B models with vLLM or TensorRT-LLM at the lowest $ per million tokens. Best $/token in class Top pick: RTX 5090 $0.318/hr Open the playbook
- Fine-tuning LoRA, QLoRA and full fine-tunes — from a single 24 GB card to 8× A100 nodes. QLoRA 70B on one GPU Top pick: A100 SXM4 $0.583/hr Open the playbook
- Image generation SDXL, Flux and ComfyUI pipelines — interactive sessions or thousand-image batch runs. Flux dev ~2 s/image on 4090 Top pick: RTX 4090 $0.262/hr Open the playbook
- Video generation Wan, HunyuanVideo and image-to-video models need VRAM headroom — rent it by the second. 32–141 GB VRAM on tap Top pick: RTX 5090 $0.318/hr Open the playbook
- 3D rendering Blender, Octane, Redshift and V-Ray on RTX hardware — per-second billing fits render farms perfectly. Per-second render billing Top pick: RTX PRO 6000 WS $1.097/hr Open the playbook
- Computer vision Train YOLO and detection models, run batch inference over image and video archives. YOLO11 epochs from $0.09 Top pick: RTX 4090 $0.262/hr Open the playbook
- Scientific computing FP64, huge memory bandwidth and MPI clusters for simulation, genomics and quantitative research. Up to 4.8 TB/s per GPU Top pick: H200 $3.058/hr Open the playbook
Not sure where your job fits?
Two shortcuts: the VRAM sizing guide answers "which card can even run this?", and the price sheet answers "what will it cost per hour?". Everything deploys from the same console in about 30 seconds.