Planning a Dell VxRail or vSAN Ready Node HCI Cluster: A Sizing Guide

Hyperconverged infrastructure (HCI) collapses compute, storage, and virtualization into a cluster of standardized x86 nodes, and Dell Technologies offers two well-supported paths to it: Dell VxRail, a jointly engineered, lifecycle-managed appliance, and vSAN Ready Nodes built on PowerEdge servers. Both run VMware vSphere and vSAN. The difference is operational: VxRail adds VxRail Manager and a fully validated, single-vendor upgrade train, while vSAN Ready Nodes give you a pre-certified PowerEdge configuration you manage with your own vSphere lifecycle tooling. Either way, the engineering question is the same — how do you size nodes so the cluster survives failures, hits performance targets, and grows without a forklift? This guide walks through that decision.
Start With the Failure Domain, Not the Spec Sheet
The most common sizing mistake is picking CPU and capacity first and treating fault tolerance as an afterthought. Reverse it. Decide how many simultaneous failures the cluster must absorb, then size everything else around that.
- Three nodes is the practical floor for production vSAN, and it tolerates exactly one node or disk-group failure (FTT=1 with RAID-1 mirroring). During maintenance on one node, you have zero remaining headroom — a second fault means data unavailability.
- Four nodes gives you a true maintenance margin: patch one node while still tolerating a failure on another.
- Five or six nodes unlocks RAID-5/6 erasure coding (FTT=1 needs four hosts, FTT=2 needs six), which reclaims the capacity overhead that mirroring costs.
Always size usable capacity, not raw. With RAID-1 mirroring you lose roughly half of raw capacity to the protection copy; erasure coding is more efficient but needs more nodes. Then add slack space (VMware recommends keeping meaningful free capacity per host so the cluster can rebuild after a failure) plus room for snapshots and swap. A node that is 90% full has nowhere to rebuild to.
For ROBO and edge sites where three nodes is too many, a two-node cluster with a witness is a legitimate Dell-supported design — just account for the witness appliance and the WAN link it depends on.
Right-Sizing the Node: CPU, Memory, and the Storage Tier
With your node count fixed, size each node for the workload — and remember every node also runs the vSAN data services, which consume CPU and RAM you must not forget to budget.
Platform. The PowerEdge R660 (1U) and PowerEdge R760 (2U) are the mainstream two-socket building blocks for VxRail and vSAN Ready Nodes. The R760's extra drive bays and PCIe slots make it the better choice when you need dense all-flash capacity or GPUs; the R660 wins on rack density for compute-heavy, lower-capacity tiers.
Size each resource with explicit overhead:
- CPU: Count physical cores against your real vCPU demand and a sane consolidation ratio, then add overhead for vSAN itself. Avoid sizing so tight that losing one node oversubscribes the survivors past your performance SLA.
- Memory: RAM is usually the first constraint to bind in virtualization. Size so that N-1 nodes can still hold the full working set — if a node fails or is in maintenance, vSphere HA has to restart those VMs somewhere.
- Storage: Go all-flash. NVMe or SAS/SATA SSD is the standard for new HCI; hybrid (flash cache over spinning disk) is effectively legacy. Plan disk groups deliberately — more, smaller disk groups generally outperform fewer large ones and shrink the blast radius of a cache-device failure.
- Network: vSAN is sensitive to network quality. Use redundant 25GbE (or higher) per node, isolate vSAN traffic, and validate latency. Underprovisioned networking is a frequent root cause of "the cluster feels slow."
A clean discipline: write down the per-resource demand, multiply by your growth horizon, then confirm the cluster still satisfies it with one node removed. If it doesn't, you've sized for the happy path only.
Plan Growth as Scale-Out, and Keep the Cluster Homogeneous
HCI's advantage is incremental scale-out — add a node, get more compute and capacity at once. Two rules keep that promise intact:
- Keep nodes homogeneous. Mixing wildly different node specs in one cluster complicates vSAN storage policies and HA admission control. If you must mix, add capacity in matched groups rather than one-off SKUs.
- Don't fill the cluster to the edge before you expand. Order the next node while you still have rebuild headroom, not after you've run out.
Where HCI is the wrong tool, say so. If your storage growth massively outpaces compute, you're buying CPU and licenses you don't need just to add capacity — a disaggregated array such as PowerStore, PowerMax, or PowerScale is the better economic fit, and PowerProtect remains your backup target regardless. Good sizing means knowing when not to put a workload on HCI.
Operations, Lifecycle, and Federal Acquisition
Day-2 operations are where VxRail and vSAN Ready Nodes diverge most. VxRail's value is the validated lifecycle: VxRail Manager orchestrates firmware, driver, BIOS, ESXi, and vSAN updates as one tested bundle, reducing the version-matrix risk that bites self-managed clusters. vSAN Ready Nodes give you certified hardware but leave that orchestration to your team and your OpenManage plus vSphere Lifecycle Manager tooling. Both rely on iDRAC for out-of-band management — provision iDRAC networking and credentials as part of the build, not afterward.
For public-sector and regulated buyers, fold compliance into sizing:
- Acquire through a pre-competed federal contract and confirm TAA-compliant configurations where required.
- Where workloads demand it, specify FIPS 140-3 validated cryptographic modules and align the build to NIST 800-171 controls.
- Document the failure-tolerance and capacity assumptions in your acquisition package; contracting officers and auditors will ask how the cluster meets availability requirements.
Takeaway
Size the failure domain first, then CPU, memory, and all-flash capacity so the cluster meets its SLA with one node down — and add nodes before you run out of rebuild headroom. Choose VxRail when you want a managed lifecycle, vSAN Ready Nodes on PowerEdge R660/R760 when you want hardware certainty with your own tooling.
Want a node count and configuration sized to your workload and availability target? Talk to a Uniqcli Dell specialist for a tailored quote.
