VMware ESXi hosts equipped with Mellanox Technologies ConnectX-6 Lx network adapters utilizing the nmlx5_core driver experience packet drops, throughput degradation, or link flapping. Interrogating the hardware counters reveals an active, rapid accumulation of physical layer frame check sequence errors and lane symbol errors.
Symptoms include:
Critical counts for rxCrcErrorsPhy and rxSymbolErrorsPhy across interface lanes.
VMware ESXi 8.0U3
Defective or degraded upstream physical network switch ports causing physical layer (Layer 1) signal attenuation and data corruption prior to host ingestion.
Schedule a maintenance window and place the affected ESXi host into Maintenance Mode via vSphere Client or SDDC Manager.
Physically migrate the fiber optic patch cabling and transceivers from the degraded switch ports to known functional alternative ports on the upstream network switch.
Coordinate with the network engineering team to administratively disable (shutdown) and flag the defective switch ports to prevent future workload allocations.
Clear the host hardware interface statistics.
Monitor the network interface statistics over a 24-hour baseline period to confirm that the error counters remain stationary at zero by running:
https://knowledge.broadcom.com/external/article/401149/receive-length-errors-detected-on-mellan.html