<TIMESTAMP> In(05) vcpu-0 - Migrate_Open: Migrating to <SOURCE_ESXI_HOST> with migration id <MIGRATION_ID>
<TIMESTAMP> In(05) vcpu-0 - Progress -1% (msg.checkpoint.saveStatus)
<TIMESTAMP> In(05) vcpu-0 - Checkpointed in VMware ESX, 8.0.3, build-24859861, Linux Host
<TIMESTAMP> In(05) vthread-40866634 - vmiop_log: (0x0): Start saving vGPU state ...
<TIMESTAMP> In(05) vthread-40866634 - vmiop_log: (0x0): Finish saving vGPU state ...
<TIMESTAMP> In(05) vcpu-0 - DeviceIfCheckPointMigrate 0000:46:00.5 SAVING: DevCptReadBytes 59686824 ret=3
<TIMESTAMP> In(05) worker-40866650 - MigrateDevCptStreamWorker: Finish transmitting device checkpoint data. (endOfStreams true failureCode 0 isListRetired false)
ESXi 8.0.3
The vmware.log shows the destination source driver version as 580.126.08 and the source host driver version as 535.274.03
According to the NVIDIA documentation (https://docs.nvidia.com/vgpu/19.0/known-issues/bug-200602087-suspend-resume-different-vgpu-manager-versions-fails.html), these are not compatible:
“Suspending a VM configured with vGPU on a host running one version of the vGPU manager and resuming the VM on a host running a version from an older main release branch fails. For example, suspending a VM on a host that is running the vGPU manager from release 19.5 and resuming the VM on a host running the vGPU manager from release 18.6 fails.”
This incompatibility is what causes “vmiop_log: (0x0): Unsupported block header version (0x0) of migration data encountered”.
Contact NVIDIA support troubleshooting this issue. If NVIDIA needs to collaborate with Broadcom support, create a Broadcom support case and reference NVIDIA's case number.