Platform 1 min read updated 2026-09-03
Quotas & limits
Defaults exist to keep the platform healthy against abuse, not to upsell you a tier. Most raise automatically as an account builds history; anything else is one support ticket.
Account defaults
| Limit | New account | Established* |
|---|---|---|
| Concurrent instances | 8 | 32 |
| GPUs per instance | 8 (hardware max) | 8 |
| Concurrent GPUs, total | 16 | 64 |
| Volumes | 10 · 4 TB total | 50 · 32 TB |
| Open top-ups per hour | 8 | 8 |
| API requests | 120/min | 600/min |
| Support tickets open | 10 | unlimited |
*Established = roughly two weeks of normal usage and settled top-ups; raises apply automatically, no request needed.
Hardware-shaped limits
- GPUs per machine — 1× to 8× depending on the chassis; the offer card is the truth.
- Ports per machine — 8–100 mapped ports, shown per offer, filterable (portsMin).
- Disk per machine — up to the free NVMe on that host, shown at deploy.
API rate behaviour
Past the per-minute budget the API answers 429 with a Retry-After header — back off and retry; nothing is queued or dropped silently. Polling advice: instance state changes also land as webhooks, which do not count against the budget.
Need more, sooner?
Fleet-scale needs (64+ GPUs, multi-node reservations) skip the automatic ladder — open a ticket with the shape of the workload, or read the enterprise page. Capacity plans answer within a business day.