An NSX Manager node that has been isolated or powered off for an extended period (e.g., several days) needs reintroduction to a functional two-node cluster. There is a concern regarding potential database corruption or cluster instability when the isolated node is powered back on.
get cluster status on healthy nodes shows the isolated node as DOWN or UNAVAILABLE in the DATASTORE group.VMware NSX
The NSX Manager uses a Corfu distributed database. When a node is isolated, it falls behind in the database epoch. The cluster maintains quorum with the remaining two nodes, but the isolated node lacks recent state updates.
The NSX Manager cluster architecture is designed to handle the reintroduction of isolated nodes through automatic synchronization. Follow these steps to recover the three-node cluster:
get cluster status
STARTING or DEGRADED to UP as it syncs with the active nodes.DataCorruptionException in /var/log/corfu/corfu.9000.log, the node must be detached and redeployed.For further assistance, Contact Broadcom Support or see Support Phone Numbers.