| Compute nodes |
8–16 nodes, one rack |
48–128 nodes, several racks |
300+ nodes across islands |
| CPU partition |
Dual-socket, all memory channels populated |
As departmental, plus a high-memory tier |
Multiple generations, partitioned by microarchitecture |
| Accelerator partition |
1–2 nodes, only if a GPU code path exists |
8–24 GPU nodes with in-node peer links |
Dedicated GPU island with GPUDirect RDMA |
| Fabric |
100 GbE with RoCEv2, single leaf |
InfiniBand HDR, two-level fat tree |
InfiniBand NDR, topology-aware placement |
| Blocking ratio |
Non-blocking within the leaf |
2:1 taper at the spine |
Non-blocking inside islands, tapered between |
| Parallel filesystem |
NFS over RDMA on all-flash |
BeeGFS or Lustre, flash metadata tier |
Lustre with separate metadata and capacity tiers |
| Node-local scratch |
2 TB NVMe per node |
4 TB NVMe per node |
8 TB NVMe per node, plus a burst tier |
| Cooling |
Contained air in an existing room |
Air with rear-door heat exchangers |
Direct liquid cooling on dense nodes |
| Rack power envelope |
Roughly 8–15 kW |
Roughly 15–35 kW |
40 kW and above |
| Scheduling |
Slurm, two partitions, simple fair-share |
Slurm with QoS tiers, preemption and accounting |
Slurm with declared topology and federation |
| What we commit to |
Time-to-result on named cases |
Plus queue-wait targets per project |
Plus energy per completed run |