Symptoms
kubectl get pods fail with Error from server: etcdserver: request timed out.Lost access to volume 54f89e21-########-####-##########98 (Datastore_Name) due to connectivity issues. Recovery attempt is in progress and outcome will be reported shortly./var/log/hostd.log: info hostd[ID] [sub=Vimsvc.ha-eventmgr] Event 15734 : Lost access to volume 5d9b401e-########-####-##########42 (Datastore_Name) due to connectivity issues./var/log/vobd.log: [vmfsCorrelator] [esx.problem.vmfs.heartbeat.timedout] 66fdce5d-########-####-##########44 VMFS_Volume_Name/var/log/vmkernel.log: ScsiDeviceIO: 4686: Cmd(0x45...) 0x2a ... failed H:0x7 D:0x0 P:0x0 WARNING: NMP: nmp_DeviceRequestFastDeviceProbe:235: NMP device "naa.600..." state in doubt .task blocked for more than 120 seconds or kernel: watchdog: BUG: soft lockup.vmkwarning.log contains: Device performance has deteriorated. I/O latency increased... to 2076000 microseconds.
ESXi hosts monitor VMFS datastores via heartbeats issued every 3 seconds. If an I/O operation does not complete within 16 seconds, the datastore is marked offline.
DEVICE BUSY status until the heartbeat is reclaimed. Guest operating systems remain online only if they can sustain these high-latency periods.etcd timeouts, as the underlying storage becomes unresponsive to the Kubernetes control plane.vmkernel.log for SCSI status H:0x7 (Host error), indicating the host cannot communicate with the storage device.hostd timeouts.etcd services will recover automatically.For further assistance, see . Scroll to the bottom of the page and click on your respective region.