This article addresses an issue where virtual machine backups using the Network Block Device (NBD) transport mode experience severe performance degradation, specifically characterized by extremely low read speeds.
Backup read speeds drop as low as 0.005 MB/sec for virtual machines on a specific host. Review of the Commvault vsbkp.log shows below entries.
12956 2208 07/01 12:04:12 11730629 stat- ID [readdisk], Bytes [312475977], Time [53659. 726612] Sec(s), Average Speed [0.005554] MB/Sec12956 2208 07/01 12:04:12 11730629 stat- ID [Datastore Read [<Datastore_name>]], Bytes [1594884775], Time [53689.737842] Sec(s), Average Speed [0.028329] MB/SecNFC sessions are established successfully but stall or cease data transfer after a period of time.
Backups may run for 15+ hours before being manually canceled.
The issue is localized to specific network paths between a particular Backup Media Agent and ESXi host.
The backup media agent is a virtual server agent running on a Windows Virtual Machine.
ESXi vmkernel.log and hostd.log shows successful snapshot creation and NFC handshake completion without errors.
Alternative Media Agents on different network segments are able to back up the same VM at normal speeds.
The affected Media Agent can back up virtual machines on other hosts successfully, indicating the bottleneck is unique to the specific route between the agent and the target host.
The underlying cause is an external network transit path bottleneck unique to the route between the virtual Media Agent and the ESXi management interface .
Host-side stability is confirmed when:
To resolve the performance bottleneck, follow these steps:
Isolate Backup Traffic (Best Practice): Create a dedicated VMkernel adapter for backup traffic to prevent contention with management traffic.
Verify Network Path Integrity: Engage the networkteam to perform a bi-directional path trace between the Media Agent IP and the ESXi VMkernel IP.
If an immediate fix for the network bottleneck is not available, modify the backup selection policy: