Two or more ESXi 8.0 Update 3 hosts experience a total loss of network connectivity following an upstream network change or physical switch hang (e.g., during an iOS upgrade).
Review the vobd.log and vmkernel.log for the following signatures around the time of the outage:
Validate if the host detected a physical link state change at the time of the upstream ToR switch failure:
In(14) vobd[#####]: [netCorrelator] #####us: [vob.net.vmnic.linkstate.down] vmnic vmnic1 linkstate downIn(14) vobd[#####]: [netCorrelator] #####us: [esx.problem.net.vmnic.linkstate.down] Physical NIC vmnic1 linkstate is down...In(14) vobd[#####]: [netCorrelator] #####us: [esx.clear.net.vmnic.linkstate.up] Physical NIC vmnic1 linkstate is upIn(182) vmkernel: cpu55:2098132)bnxtnet: bnxtnet_display_link:####: [vmnic6 : 0x########] NIC Link is down...In(182) vmkernel: cpu48:2098132)bnxtnet: bnxtnet_display_link:####: [vmnic6 : 0x########] NIC Link is Up, 10000 Mbps (NRZ) full duplex, Flow control: noneThe failure domain resides at the physical network layer. While the ESXi host eventually detects the link down event and broadcasts failover notifications (RARP) over the healthy secondary link, traffic fails to route if the secondary switch drops these updates or the upstream fabric is logically isolated.
If you identify these log signatures coinciding with an upstream switch event, engage your physical network team to investigate the secondary (surviving) switch:
For guidance on uplink allocation, see KB 420924 How to Allocate Specific Uplinks to Port Groups on a vSphere Distributed Switch.