Server & GPU Rental

GPU Cluster Rental, 8-GPU Nodes

  • New
  • Multi-year savings
From $19,000.00 /mo per node or $228,000/yr

Pay monthly or annually — the contract value is the same.

Save 5% with a 2-year term or 10% with a 3-year term.

Dedicated 8-GPU NVIDIA H100, H200 or B200 nodes on a non-blocking InfiniBand fabric, with Slurm or Kubernetes and shared storage.

About this service

  • Dedicated nodes with 8 NVIDIA H100, H200 or B200 GPUs connected by NVLink and NVSwitch
  • Non-blocking InfiniBand fabric with 400 Gb/s per GPU (3.2 Tb/s per node) between nodes
  • Slurm or Kubernetes with the NVIDIA GPU Operator, CUDA and NCCL installed and validated
  • Shared parallel file system capacity included with every node
  • Burn-in and multi-node performance tests before handover; GPU health monitoring and part replacement included

Plans at a glance

Compare all features

About GPU Cluster Rental, 8-GPU Nodes

GPU Cluster Rental provides dedicated NVIDIA HGX nodes with eight GPUs each, connected by a non-blocking InfiniBand fabric, for distributed training, fine-tuning and high-throughput inference. You reserve the nodes for the full term: no spot interruptions, no shared GPUs and no per-hour metering.

Inside each node, NVLink and NVSwitch connect all eight GPUs to one another; between nodes, eight 400 Gb/s InfiniBand ports carry GPU-to-GPU traffic. Every node also has local NVMe scratch space and capacity on a shared parallel file system. We deliver the cluster with Slurm or Kubernetes, NVIDIA drivers, CUDA, NCCL and container tooling, validated by burn-in and multi-node performance tests before handover.

Our engineers monitor GPU health, replace failed components, keep firmware and drivers current in agreed maintenance windows and help your team tune multi-node jobs. You keep full administrative access to your nodes and your data.

Compare plans

All prices in USD, per node at the 1-year rate. Pay monthly or annually — the contract value is the same. Swipe the table sideways to see every plan.

Plan comparison for GPU Cluster Rental, 8-GPU Nodes
Feature H100 node H200 node Recommended B200 node
Price $19,000.00 /mo per node or $228,000.00/yr $34,000.00 /mo per node or $408,000.00/yr
With a 3-year term $17,100.00/mo Save 10% or $205,200.00/yr $30,600.00/mo Save 10% or $367,200.00/yr
Best for Hopper-generation nodes for large-scale training and fine-tuning. Blackwell-generation nodes for the largest training and inference workloads.
What’s included
  • 8× NVIDIA H100 80 GB SXM per node (640 GB HBM3)
  • NVLink and NVSwitch at 900 GB/s per GPU
  • 8× 400 Gb/s InfiniBand per node
  • 2 TB DDR5 RAM and 30.72 TB local NVMe per node
  • 100 TB of shared parallel file system per node
  • 4-hour hardware replacement target, 24/7
  • 8× NVIDIA B200 180 GB per node (1,440 GB HBM3e)
  • Fifth-generation NVLink at 1.8 TB/s per GPU
  • 8× 400 Gb/s InfiniBand per node
  • 4 TB DDR5 RAM and 30.72 TB local NVMe per node
  • 200 TB of shared parallel file system per node
  • Named cluster engineer and monthly performance review
  • 1-hour response for critical incidents, 24/7
Order

Billed monthly · 1-year term

Customize billing, term & quantity (Plan: H100 node)

Billed monthly · 1-year term

Customize billing, term & quantity (Plan: B200 node)

Specifications

Specifications for GPU Cluster Rental, 8-GPU Nodes
SKUTS-RNT-304
CategoryServer & GPU Rental
PricingPer node, per year; payable monthly or annually
ContractAnnual commitment; 1-, 2- or 3-year term
BillingMonthly or annual invoicing; same contract value
GPUs per node8× H100 80 GB / 8× H200 141 GB / 8× B200 180 GB depending on tier
GPU interconnectNVLink with NVSwitch: 900 GB/s per GPU (H100, H200), 1.8 TB/s per GPU (B200)
Cluster fabric8× 400 Gb/s NDR InfiniBand per node, non-blocking; redundant Ethernet front end
Node memory2 TB DDR5 (H100, H200) or 4 TB DDR5 (B200); 30.72 TB local NVMe
Shared storageParallel file system: 100 / 150 / 200 TB usable per node
OrchestrationSlurm or Kubernetes; root access on every node
Lead timeTypical delivery in 4–8 weeks, subject to hardware availability

Frequently asked questions

How many nodes can we rent, and are they on the same fabric?

You can order 1 to 16 nodes (8 to 128 GPUs) per order. Nodes ordered together are placed on the same non-blocking InfiniBand fabric so multi-node jobs run at full bandwidth. Larger clusters are quoted on request.

Can we pay monthly, and can we end the rental early?

You can choose monthly invoices, which spread the same annual commitment over 12 payments. The rental cannot be ended early: nodes are reserved for your company for the full term. You can add nodes at any time, prorated to your renewal date and subject to hardware availability.

Is support for our models and training code included?

We support the full infrastructure stack: hardware, firmware, drivers, CUDA, NCCL, the scheduler and the file system. Your model code and framework-level debugging stay with your team, although H200 and B200 tiers include regular tuning sessions with our engineers.

What happens to our data at the end of the term?

Your data stays on your nodes and file system allocation until the last day of the term; we recommend copying it off during the final two weeks. Drives are then securely wiped before the hardware is reassigned.

Have another question? Ask a solutions architect