Welcome to Microsoft Q&A!
Thank you for taking the time to review the documentation so carefully and for highlighting the apparent contradiction. Your concern is valid, and your interpretation of the published guidance for a 2+2 campus cluster is correct.
The important distinction is between fault-domain-aware data placement and the failure scenarios that a specific resiliency configuration is guaranteed to survive.
In a Storage Spaces Direct (S2D) campus cluster, defining rack fault domains and configuring rack awareness allows S2D to distribute data according to the rack topology. However, this does not mean that every two-copy mirror volume is guaranteed to remain online after an entire rack is lost.
For a 2+2 campus cluster, with two nodes in each rack, Microsoft’s published guidance states:
- Two-copy mirror, 50% efficiency: Survives the loss of one node, but does not survive the loss of an entire rack.
- Four-copy mirror, 25% efficiency: Survives the loss of an entire rack plus one additional node.
The documentation therefore recommends using four-copy mirrored volumes for a 2+2 campus cluster when rack-level resiliency is required.
Answers to your questions:
1. Does a two-copy mirror remain online after an entire rack fails in a 2+2 configuration?
No. In a 2+2 campus cluster with StorageRack fault-domain awareness, a two-copy mirror is supported as a volume configuration, but it is not considered rack resilient.
If an entire rack containing two nodes is lost, the two-copy volume is not guaranteed to remain online. Rack-aware placement does not increase the failure tolerance of the underlying two-copy resiliency configuration.
2. Is the rack failure treated as one fault-domain failure or as the loss of two nodes?
The cluster recognizes the rack as a fault domain. However, Storage Spaces resiliency remains limited by the number and placement of available data copies.
In practical terms, losing one rack in a 2+2 deployment also means losing two nodes and all storage attached to those nodes. For a two-copy mirror, that failure exceeds the documented guarantee of surviving one node failure.
3. Can the witness keep the storage pool and volumes online?
The witness contributes to cluster quorum, helping the surviving nodes maintain cluster membership and avoid a split-brain condition.
However, the witness does not store volume data, provide an additional mirror copy, or compensate for unavailable S2D storage. Maintaining cluster quorum does not, by itself, guarantee that the storage pool and its volumes can remain online after losing half of the storage-bearing nodes.
4. Would a three-way mirror in a 3+3 campus cluster survive the loss of one rack?
The current campus-cluster documentation does not provide the same explicit resiliency matrix for a 3+3 layout that it provides for the 2+2 layout. It would therefore be better not to state that a standard three-way mirror is rack resilient in this topology unless Microsoft publishes an explicit support statement or confirms the design through a support case.
The documented Windows Server 2025 campus-cluster model uses exactly two rack fault domains and describes two-copy and four-copy volume options. The published guidance explicitly identifies four-copy mirror as the rack-resilient choice for the 2+2 topology.
5. Is there a supported option that provides rack resiliency at 50% storage efficiency?
Based on the currently published guidance, Microsoft does not document a 2+2 campus-cluster configuration that combines:
- 50% storage efficiency,
- Survival of an entire rack failure, and
- Continued availability of all volumes.
For a 2+2 campus cluster, the documented recommendation for surviving an entire rack failure is four-copy mirror, with 25% storage efficiency.
In short:
- 2+2 with two-copy mirror: Supported as a volume configuration but not rack resilient. It survives one node failure, not the documented loss of an entire rack.
- 2+2 with four-copy mirror: Supported and recommended when rack resiliency is required. It survives the loss of an entire rack plus one additional node.
- Rack fault-domain awareness: Influences data placement but does not make a two-copy mirror capable of surviving every rack-level failure.
- Cluster witness: Helps maintain cluster quorum but does not provide storage redundancy or replace unavailable data copies.
- 3+3 with three-way mirror: The available documentation does not explicitly confirm the rack-failure behavior, so this should be validated with Microsoft before using it as a supported design assumption.
- 50% efficiency with rack resiliency: This is not documented as a supported option for the published 2+2 campus-cluster topology
For additional information, please visit: Create a Storage Spaces Direct Campus Cluster | Microsoft Learn
Deploy Storage Spaces Direct on Windows Server | Microsoft Learn
Fault domain awareness | Microsoft Learn
If you find this information helpful, please click Accept Answer.
Thank you for using Microsoft Q&A.