Troubleshooting Storage Issues in VCF Operations
search cancel

Troubleshooting Storage Issues in VCF Operations

book

Article ID: 318408

calendar_today

Updated On:

Products

VCF Operations/Automation (formerly VMware Aria Suite)

Issue/Introduction

This article provides troubleshooting steps for resolving storage capacity issues in VCF / Aria Operations.

You may be experiencing a storage issue if one or more file systems on your VCF / Aria Operations nodes exceed 80% utilization (verifiable by running the df -h command on each node).



UI Errors:

  • Data retriever is not initialized yet

  • Unable to connect Platform Services

  • Waiting for Analytics

  • The Cluster was shut down because one node was out of disk space

  • ERR_TOO_MANY_REDIRECTS

  • Disk space on this appliance is running low and filling up rapidly. If no action is taken immediately, services and UI access will go down. Add more disk space immediately to prevent service disruption

Active Alerts:

  • Disk space on node is low

  • Fsdb is running critically low on disk space (or is estimated to run out soon)

  • vRealize Operations Cluster database is running out of disk space

Log Errors:

  • You may see below exceptions in the /storage/log/vcops/log/analytics-<uuid>.log file. 

    ERROR [FsdbDataRetentionManager]  com.vmware.vcops.fsdb.FsdbDataRetentionManager.deleteResourceData - Delete resource data request is failed for resource #### :com.integrien.alive.FSDB.LowDiskSpaceException: Failed to increase file size. File /usr/lib/vmware-vcops/data/#/####/####_##_####.dat. Reason: FSDB is running low on disk space

    ERROR [Regular Data storage worker thread 1]  com.vmware.vcops.fsdb.datareceiver.StorageWorkItem.saveObservation - Save metric data failed: Failed to increase file size. File /usr/lib/vmware-vcops/data/#/##/####_##_##.dat. Reason: FSDB is running low on disk space

Environment

VMware Aria Operations 8.18.x

VMware Cloud Foundation Operations 9.0.x

Resolution

Proceed with the relevant solution below:

  1. /storage/db is out of space
  2. /storage/log is out of space

  3. / is out of space

Note : Expanding an existing disk is not supported and may result in cluster instability.