You observe the following symptoms across multiple ESXi hosts or independent vCenter environments:
vSphere HA failover operation in progress.Partitioned or FDMUnreachable.vSphere HA
ESX 8.x
vCenter 8.x
NSX 4.x
This issue occurs when a transient, external physical network disruption—such as a packet storm or severe network congestion—simultaneously isolates ESXi hosts from their management gateways and storage heartbeat datastores.
Based on the fdm.log and physical switch statistics, the VMware HA mechanism functions as designed by detecting the loss of communication and initiating failover. High discard or throttle statistics on physical switches (e.g., Dell, Cisco) typically correlate with these events.
There is no configuration change required within the VMware software layer. To resolve and prevent recurrence, investigate the physical network infrastructure:
fdm.log for the following patterns:/var/run/log/vmkernel.log for overlay flaps indicating control detection time expiration:/var/run/log/vmkernel.log for the following types of messagespktcap-uw) during the next occurrence.