| Training fabric — Hopper | 8×400 Gb/s InfiniBand NDR per node (3.2 Tbps), rail-optimized non-blocking fat-tree, SHARP in-network reduction |
|---|
| Training fabric — Blackwell | Quantum-X800 InfiniBand XDR, 800G per link; GB200 racks add a 72-GPU coherent NVLink 5 domain at 1.8 TB/s per GPU |
|---|
| Oversubscription | 1:1, all tiers. No blended east-west fabric, no shared spine with other tenants |
|---|
| Local scratch | 30 TB NVMe Gen5 per node, ~55 GB/s read — checkpoint staging without touching the network |
|---|
| Parallel filesystem | Managed WEKA, dedicated per cluster: up to 720 GB/s aggregate read per pod, POSIX + GPUDirect Storage |
|---|
| Object storage | S3-compatible, NVMe-cached, co-located with compute — $0.055/GB-month hot tier |
|---|
| Data transfer | Ingress $0 · 20 TB/node-month egress included, then $1.25/TB · inter-node $0. Free 100G Direct Connect on yearly reservations |
|---|
| Front-end network | Dual 100 GbE per node, DDoS-protected, BYO-IP supported |
|---|