-
Eleven of the fifteen Nutanix NX models take cards, and the per-node ceiling is four H100 or L40S in the NX-9151-G9. We went through which nodes support GPUs, how many VMs a card yields via MIG and vGPU, and why the L40S cannot be cut into hardware slices at all. Plus what the presentations leave out: MIG disables NVLink, and live migration demands identical hardware on both hosts.
-
MIG partitions a card in hardware, vGPU slices time and bills a license per user. We work out how many virtual machines fit on an RTX PRO 6000, an RTX PRO 5000 and an L40S, which profiles and licenses that takes, which hypervisors run MIG-backed vGPU, and why a card listing MIG on its datasheet may still refuse to enable it.
-
The 600 W on a card's data sheet is a requirement on the power supply, not the draw of the machine. We collected measured figures for one, two and four card builds, went through the sense pins of 12VHPWR and 12V-2x6, and explained why an RTX PRO 6000 in a server often caps at 450 W instead of 600. Plus the arithmetic for a 16 amp circuit.
-
Model weights are one multiplication, and the KV cache eats as much again or more: a 70B with a 128K context is 180 GB for a single user. We walk through the official NVIDIA formulas and show how many parallel sessions a 32, 48, 72, 96 or 141 GB card will hold. Plus two traps that keep a model from starting on a card the arithmetic says it fits.