Following a storage outage or vSAN partitioning event, the NSX Manager cluster may fail to initialize. This typically manifests as an inaccessible UI and services reporting as UNAVAILABLE.
Symptoms
get cluster status shows services as UNAVAILABLE or UNKNOWN.Failed to start default target: Transaction for nsx-custom.target/start is destructive./var/log/corfu/corfu.9000.log) report DataCorruptionException: Checksum mismatch.VMware NSX 4.x
An underlying storage interruption or improper shutdown causes file system inconsistencies and unrecoverable corruption in the Corfu database binary files.
Workaround:
For immediate recovery of a corrupted cluster:
fsck from the GRUB emergency menu to repair partition errors. Follow Resolution section in KB Manager node disk/partition is mounted as read-only alarm in NSX Manager - Disk Corruption correction using FSCKfsck, run cat /config/corfu/LAYOUT_CURRENT.ds on all nodes. A node with a mismatched or non-replicating layout must be replaced.detach node <UUID>get cluster status reports the cluster as STABLE.