ESXi 8.0 hosts running on HPE ProLiant Gen11 hardware experience intermittent or total loss of access to VMFS datastores. This results in all virtual machines on the impacted host becoming unresponsive or powering off.
Symptoms:
Lost access to volume <Datastore_Name> due to connectivity issues.ScsiDeviceIO: 4115: Cmd(0x45d988625dc8) 0x8a, CmdSN 0x800e000c from world 3129234 to dev "naa.####" failed H:0x0 D:0x8 P:0x0WARNING: NMP: nmp_DeviceRequestFastDeviceProbe:235: NMP device "naa.####" state in doubtDevice going All paths down
The connectivity loss is caused by transport-layer instability at the Fibre Channel Host Bus Adapter (HBA) level. This is frequently associated with outdated system BIOS on HPE Gen11 platforms and physical layer errors (CRC errors) on the FC vmhba, leading to SCSI H:0x8 (Busy/Retry) status and eventual All Paths Down (APD) states.
Perform the following steps to resolve and prevent recurrence:
"/usr/lib/vmware/vmkmgmt_nic/vmkmgmt_nic -sh vmhba<X>"For assistance with log collection or further hardware analysis, see Contact Broadcom Support.