ESXi hosts may experience persistent management stack instability, characterized by the host entering a "Not Responding" state in vCenter and recurring failures of core management services (hostd, vpxa, vsanmgmtd). This document details how to identify and resolve this issue when it is caused by physical CPU hardware degradation.
hostd) and vCenter agent (vpxa)VMware vSphere ESXi
This issue is rooted in physical hardware degradation of the server’s CPU package. When a specific CPU package or core suffers from localized memory corruption or internal logic faults, it causes fatal exceptions that propagate to the host management stack, leading to service exhaustion and host isolation.
Log Evidence:
hostd-zdump) show recurrent crashes on the same specific physical CPU core or package.