Microsoft Windows Server Failover Clustering (WSFC) nodes using shared VMDKs on vSphere 8.x experience cluster timeouts or "Production SQL is down" symptoms following storage path redundancy degradation.
Device <naa.id> performance has deteriorated or redundancy degraded.When storage paths fail or become degraded, the ESXi host attempts to scan LUNs for metadata. If these LUNs are held by a WSFC node via SCSI-3 Persistent Reservations and the "Perennial Reservation" flag is not set on the host, the scan operation hangs or delays, leading to cluster heartbeats failing.
esxcli storage core device setconfig -d <naa.id> --perennial-reservation=trueesxcli storage core device list -d <naa.id>Additional Information: Refer to if storage path redundancy remains degraded after host-level configuration. Refer to KB for log collection procedures.