All clusters are vetted and meet minimum technical requirements. Deploy on-demand or reserve instances with no long-term contracts.
- 24/7 support (Slack, phone, email)
- 1.5TB+ NVMe storage
- 1TB+ of RAM
- No ingress/egress fees
Everything you need to know about running on SF Compute. Still have questions? Reach out— we're quick to respond.
All of the above. Bare metal and managed Slurm for the lowest-level access, VMs for flexibility, and a managed Kubernetes offering on top.
Most VMs are ready in well under a minute. Cold starts on uncommon configurations can take a little longer.
Contact us for a BM or managed Slurm cluster which both support InfiniBand today. We'll support InfiniBand on VMs in Q3 2026.
Yes. Hardware failure rates are much higher on GPU clusters than on web servers. At certain scales, they're guaranteed, so we've designed for failure. We have strict hardware requirements and have seen just about everything that can go wrong. Unlike other providers, we refund for failed nodes and can repack your nodes to ones with healthy hardware.
Shared Slack channels with our engineers, plus on-call coverage for production clusters. Enterprise plans include a dedicated solutions engineer.
