Skip to content

Clusters

Instant Clusters are multi-node GPU setups where multiple machines are connected together via high-speed InfiniBand interconnect. Unlike single GPU instances, Instant Clusters are designed for large-scale distributed training and inference jobs that require multiple nodes connected with fast Infiniband connection. Verda clusters are designed for distributed GPU workloads that need high-speed interconnect, coordinated scheduling, shared storage, and operational visibility.

You can now deploy high-performance GPU cluster with Infiniband interconnect from your Verda Cloud Console, the same way you would deploy a single GPU instance.

The only available contract length is: Pay As You Go.

Instant clusters are available with Nvidia H200 SXM5, Nvidia B200 SXM6, or Nvidia B300 GPUs. Each worker node has eight InfiniBand links, 400 Gb/s each on H200 and B200 (3.2 Tb/s per node) or 800 Gb/s each on B300 (6.4 Tb/s per node), plus a 100 Gb/s Ethernet network connecting all nodes in all our cluster product. The uplink to the Internet is symmetric 2 Gb/s.

Our instant clusters range from 16 to 128 GPUs. Each cluster in addition to worker nodes, has one jump host and one service node. Each worker node has 7TB of local NVMe storage and access to a configurable shared filesystem with up to 50 TiB of storage.

For larger or specialized configurations, set up a customized GPU cluster, or contact our support.

Clusters have Kubernetes or Slinky (Slurm on Kubernetes) pre-installed for easy job management and Grafana dashboard for monitoring and alerts. The instant clusters are currently available in FIN-03 location.