<!-- GPU clusters -->

# GPU clusters

All clusters are vetted and meet [minimum technical requirements](https://sfcompute.com/requirements). Deploy on-demand or reserve instances with no long-term contracts.

- 24/7 support (Slack, phone, email)
- 1.5TB+ NVMe storage
- 1TB+ of RAM
- No ingress/egress fees

<!-- Frequently asked questions -->

# Frequently asked questions

Everything you need to know about running on SF Compute. Still have questions? [Reach out](https://sfcompute.com/contact) — we're quick to respond.

## Is it bare metal? Containers? VMs? Kubernetes?

All of the above. Bare metal and managed Slurm for the lowest-level access, VMs for flexibility, and a managed Kubernetes offering on top.

## How fast do VMs spin up?

Most VMs are ready in well under a minute. Cold starts on uncommon configurations can take a little longer.

## Are nodes fully-interconnected with InfiniBand?

Contact us for a BM or managed Slurm cluster which both support InfiniBand today. We'll support InfiniBand on VMs in Q3 2026.

## Will the nodes go down?

Yes. Hardware failure rates are much higher on GPU clusters than on web servers. At certain scales, they're guaranteed, so we've designed for failure. We have strict hardware requirements and have seen just about everything that can go wrong. Unlike other providers, we refund for failed nodes and can repack your nodes to ones with healthy hardware.

## What support do you have?

Shared Slack channels with our engineers, plus on-call coverage for production clusters. Enterprise plans include a dedicated solutions engineer.
